frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

The Unreasonable Effectiveness of Reasonless Intermediate Tokens

https://arxiv.org/abs/2505.13775
4•YeGoblynQueenne•1y ago

Comments

tocs3•1y ago
I asked ChatGPT to restate this in more laymen's terms (posted below) and I am not to surprised at the answer.

"Lately, some AI models have shown impressive abilities to solve complex problems, and many people credit this to a method called Chain of Thought (CoT), where the model is trained to think through steps like a human might. In this paper, we take a closer look at that idea to see if it's really what's driving better performance.

We focus on the model’s step-by-step thinking (the words it generates along the way) — often treated like human "thoughts" — and examine whether these actually help the model solve problems more accurately. To test this, we train AI models using clean, correct step-by-step reasoning paths and final answers, all based on a known solving method (A* search). This lets us check both the final answers and the reasoning steps to see how they relate.

Interestingly, we find that even when a model gives the right answer, its reasoning steps can still be wrong or messy. To go further, we even train models using completely random and incorrect reasoning steps — and surprisingly, they still perform about the same, and sometimes even better, than those trained on correct steps.

This suggests that the step-by-step "thoughts" the model shows aren’t as meaningful or reliable as many assume. In short, just because a model looks like it’s reasoning through a problem doesn’t mean it actually is — and we should be careful not to treat its outputs as if it thinks like a human or follows strict logic."

Brookfield's Bruce Flatt discusses Nvidia strategic partnership on CNBC [video]

https://www.youtube.com/watch?v=Uxy4s2FP2EY
1•KasianFranks•2m ago•0 comments

Show HN: A nutrition companion for endurance athletes

1•bravang•2m ago•0 comments

An Interview with Arthur Mensch

https://www.economist.com/insider/inside-tech/an-interview-with-arthur-mensch
1•nmat•2m ago•0 comments

Show HN: PageSieve, a web scraping browser extension

https://julius383.github.io/PageSieve/
1•kajm•5m ago•0 comments

CINC: Computing in Cardiology

https://cinc.org/
1•teleforce•8m ago•0 comments

Astronomers Discover the Existence of a Black Hole Star

https://www.wired.com/story/black-hole-stars-are-becoming-less-hypothetical/
1•beardyw•9m ago•0 comments

A look into the harmful effects of Ocean Ramsey and shark influencers

https://www.drjackacooper.com/writing/shark-whisperer-is-lying-to-you
1•mindracer•10m ago•0 comments

Words Are Disappearing in the New Trump administration (2025)

https://www.nytimes.com/interactive/2025/03/07/us/trump-federal-agencies-websites-words-dei.html
1•donohoe•12m ago•1 comments

Democracy in Music

https://democracyinmusic.com/
1•skoodge•13m ago•0 comments

Slater gets full-text BM25 indexing and Graphiti Support

https://github.com/Hikari-Systems/graphiti-slater
1•rickkjp•13m ago•1 comments

Built a clicker game for my classroom – need honest opinions about it

https://www.buttonblam.com/
1•useful-warlock•16m ago•1 comments

RL for LLM Reasoning Is Sparse Policy Selection, Not Capability Learning

https://arxiv.org/abs/2605.06241
1•BlackGlory•18m ago•0 comments

What 50 open source projects taught us about security in the AI era

https://github.blog/open-source/maintainers/what-50-open-source-projects-taught-us-about-security...
2•yruzin•20m ago•1 comments

System design rules I always come back to

https://www.thetrueengineer.com/p/10-system-design-rules-i-always-come
2•adletbalzhanov•23m ago•0 comments

The Seven Weaknesses of AI Today

https://medium.com/tech-ai-chat/the-seven-weaknesses-of-ai-today-eee66aaa9c10
1•elye•27m ago•1 comments

Show HN: A public AI whose memory is shared across all users

https://wildstatic.com/
3•adjohu•28m ago•0 comments

Ranking the Prettiest Birds with Math

https://moultano.wordpress.com/2026/08/14/fairly-ranking-the-most-brilliant-birds/
2•moultano•29m ago•0 comments

Scientists Identify 3-Minute Exercise Better than 90M one

https://www.newsweek.com/scientists-say-this-3-minute-exercise-beats-90-minute-workouts-12322062
3•guytv•31m ago•1 comments

Clues by Sam

https://cluesbysam.com/
2•abelanger•34m ago•1 comments

Our Reality Is Shifting and It's Just the Start

https://guustaaf.substack.com/p/our-reality-is-shifting-and-its-just
3•Guustaaf•36m ago•1 comments

A recipe for drone racing with reinforcement learning

https://mrandri19.github.io/2026/07/19/drone-racing-with-reinforcement-learning.html
2•daww•38m ago•1 comments

How a Sony Veteran Is Overhauling the Company He Grew Up In

https://www.wsj.com/business/media/sony-ceo-hiroki-totoki-efc8923f
6•thm•38m ago•0 comments

The Mystery of Dark Oxygen

https://www.newyorker.com/science/elements/the-mystery-of-dark-oxygen
2•rbanffy•39m ago•0 comments

A U.S. Strategy to Prevent the Creation of Mirror Life

https://www.rand.org/pubs/research_reports/RRA4335-1.html
3•thunderbong•41m ago•0 comments

Woman claims stepfather used Grok to transform childhood photo into CSAM imagery

https://techcrunch.com/2026/08/15/woman-claims-her-stepfather-used-grok-to-transform-childhood-ph...
3•ilamont•45m ago•0 comments

Ranking Kickstarter Games I Backed but Never Finished or Ended Up Playing at All

https://blog.curiousquail.com/ranking-kickstarter-games-i-backed-but-never-finished-or-ended-up-p...
3•ajdude•46m ago•0 comments

A.I. Agents Are Taking Online Courses for Cheating Students

https://www.nytimes.com/2026/08/10/us/ai-cheating-online-degrees.html
3•bookofjoe•47m ago•1 comments

Show HN: 3D Graph of DeepSWE, via Three.js

https://icauro.com/toabundance
3•icauroboros•47m ago•2 comments

AGI will likely be silent

https://www.reddit.com/r/agi/s/sZECbecJVS
4•qikouki•48m ago•0 comments

WTF Are You Loading?

https://x-x.codes/posts/wtf-are-you-loading
2•alex_x•48m ago•0 comments