frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

The Unreasonable Effectiveness of Reasonless Intermediate Tokens

https://arxiv.org/abs/2505.13775
4•YeGoblynQueenne•1y ago

Comments

tocs3•1y ago
I asked ChatGPT to restate this in more laymen's terms (posted below) and I am not to surprised at the answer.

"Lately, some AI models have shown impressive abilities to solve complex problems, and many people credit this to a method called Chain of Thought (CoT), where the model is trained to think through steps like a human might. In this paper, we take a closer look at that idea to see if it's really what's driving better performance.

We focus on the model’s step-by-step thinking (the words it generates along the way) — often treated like human "thoughts" — and examine whether these actually help the model solve problems more accurately. To test this, we train AI models using clean, correct step-by-step reasoning paths and final answers, all based on a known solving method (A* search). This lets us check both the final answers and the reasoning steps to see how they relate.

Interestingly, we find that even when a model gives the right answer, its reasoning steps can still be wrong or messy. To go further, we even train models using completely random and incorrect reasoning steps — and surprisingly, they still perform about the same, and sometimes even better, than those trained on correct steps.

This suggests that the step-by-step "thoughts" the model shows aren’t as meaningful or reliable as many assume. In short, just because a model looks like it’s reasoning through a problem doesn’t mean it actually is — and we should be careful not to treat its outputs as if it thinks like a human or follows strict logic."

iPhone 18 Pro Could Get a Signature New Color – Here's What the Latest Leaks Sug

https://moztako.me/iphone-18-pro-colors-dark-cherry/
1•cecyev•1m ago•0 comments

Will farm animals always suffer?

https://worksinprogress.co/issue/will-farm-animals-always-suffer/
1•mooreds•1m ago•0 comments

Ben Sasse on Faith, Family, and Facing Death [audio]

https://www.econtalk.org/ben-sasse-on-faith-family-and-facing-death/
1•mooreds•2m ago•0 comments

Tiny drones are powered by sound

https://actu.epfl.ch/news/these-tiny-drones-are-powered-by-sound-2/
1•simonebrunozzi•6m ago•0 comments

Jane Street took $15B hit in July tied to Situational Awareness

https://www.reuters.com/business/finance/jane-street-took-15-billion-hit-july-tied-situational-aw...
1•paulpauper•7m ago•0 comments

A mother has spent years caring for her special-needs child

https://www.washingtonpost.com/health/2026/08/16/she-spent-years-caring-her-child-with-disabiliti...
1•paulpauper•8m ago•0 comments

Worse Is Better Considered Harmful

https://cs.stanford.edu/people/eroberts/cs201/projects/2010-11/WorseIsBetter/index.php/Main_Page....
2•so-cal-schemer•8m ago•1 comments

Stock Market Bargains Are Hiding in This Overlooked Place

https://www.wsj.com/finance/stocks/stock-market-bargains-are-hiding-in-this-overlooked-place-6239...
1•paulpauper•9m ago•0 comments

ASCII City Update: Interiors, Elevation and Skyscrapers [video]

https://www.youtube.com/watch?v=UCKEDWowc0o
1•galleywest200•11m ago•1 comments

Companies that buy and sell your data are not following CA’s strict privacy laws

https://hai.stanford.edu/news/companies-that-buy-and-sell-your-data-are-not-following-californias...
1•hhs•12m ago•0 comments

Free Brain-Computer Interface in Udemy

https://www.reddit.com/r/PiEEG_club/comments/1vvq6j7/5_free_eegbci_courses_join_the_new_1020_acad...
1•Christiangmer•14m ago•0 comments

Digital amnesia: Taking screenshots makes you more likely to forget information

https://www.binghamton.edu/news/story/6397/digital-amnesia-taking-screenshots-makes-you-more-like...
1•hhs•17m ago•1 comments

Quiet on set. How AI transformed China's microdrama scene

https://www.cnn.com/2026/08/22/style/short-drama-ai-china-intl-hnk
1•T-A•19m ago•0 comments

Stelm – A BYOC canvas for AWS infra

https://stelm.dev/
1•tredex•19m ago•0 comments

Europe Is Building an Impossible Fusion Reactor [Ben Miles] [video]

https://www.youtube.com/watch?v=bPUAKW1Dyek
2•mdp2021•20m ago•0 comments

Netscape's Rise and Fall: A Browser Wars History

https://medium.com/@gp2030/netscapes-rise-and-fall-a-browser-wars-history-8546e3b52092
1•Bluestein•22m ago•1 comments

How much is crawling your content worth to an AI bot?

https://insights.som.yale.edu/insights/how-much-is-crawling-your-content-worth-to-an-ai-bot
1•hhs•23m ago•0 comments

MiniMax M3 Medium hits 73.17% F1 on DeepSearchQA, near GPT-5 High

https://huggingface.co/datasets/youdotcom/minimax-m3-deepsearchqa-skill-eval
1•EdwardIrby•23m ago•0 comments

People vs. the AI Overlords

https://anarc.at/blog/2026-08-18-people-vs-ai-overlords/
1•edward•26m ago•0 comments

From Mega-Machines to Mega-Algorithms (2015)

https://thenewinquiry.com/from-mega-machines-to-mega-algorithms/
1•measurablefunc•26m ago•0 comments

Mysterious Astroturf Group Popped Up to Defend the Paramount Merger

https://theintercept.com/2026/08/21/paramount-merger-warner-discovery-david-ellison-antitrust/
3•mukmuk•28m ago•0 comments

The Biblical myth of the CREATION of EVE, the first woman

https://medium.com/freedomofthought/the-biblical-myth-of-the-creation-of-eve-the-first-woman-0f2d...
1•raynchad•29m ago•0 comments

Follow live the open training of a 535B (23B activated) LLM

https://twitter.com/percyliang/status/2090918065634684997
1•ggcr•31m ago•1 comments

436× faster exact state updates with a disk-backed incremental Merkle tree

https://github.com/yangofzeal/Hkd_merkle
1•msyang2004_82•32m ago•0 comments

Show HN: Structo – Claude Code for short form writing

https://structo.redbeardlab.com/
2•siscia•32m ago•0 comments

Optical Earth – Earth as the human eye would see it

https://www.opticalearth.com/
3•afterglowreader•33m ago•1 comments

Ordinary agent tool calls can create shadow delegation

https://niyikiza.com/posts/agents-to-agents/
2•niyikiza•34m ago•0 comments

Microscheme: A functional programming language for the Arduino

https://ryansuchocki.github.io/microscheme/
2•so-cal-schemer•43m ago•1 comments

NanoGPT Speedrun Frontier

https://www.primeintellect.ai/research/nanogpt-speedrun
3•stared•44m ago•0 comments

Pre-Scheme: a Scheme for low-level programming

https://prescheme.org/
2•so-cal-schemer•44m ago•1 comments