frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

The Unreasonable Effectiveness of Reasonless Intermediate Tokens

https://arxiv.org/abs/2505.13775
4•YeGoblynQueenne•1y ago

Comments

tocs3•1y ago
I asked ChatGPT to restate this in more laymen's terms (posted below) and I am not to surprised at the answer.

"Lately, some AI models have shown impressive abilities to solve complex problems, and many people credit this to a method called Chain of Thought (CoT), where the model is trained to think through steps like a human might. In this paper, we take a closer look at that idea to see if it's really what's driving better performance.

We focus on the model’s step-by-step thinking (the words it generates along the way) — often treated like human "thoughts" — and examine whether these actually help the model solve problems more accurately. To test this, we train AI models using clean, correct step-by-step reasoning paths and final answers, all based on a known solving method (A* search). This lets us check both the final answers and the reasoning steps to see how they relate.

Interestingly, we find that even when a model gives the right answer, its reasoning steps can still be wrong or messy. To go further, we even train models using completely random and incorrect reasoning steps — and surprisingly, they still perform about the same, and sometimes even better, than those trained on correct steps.

This suggests that the step-by-step "thoughts" the model shows aren’t as meaningful or reliable as many assume. In short, just because a model looks like it’s reasoning through a problem doesn’t mean it actually is — and we should be careful not to treat its outputs as if it thinks like a human or follows strict logic."

Tales O' Tech: Suffering Fools

https://zeldman.com/2026/09/01/tales-o-tech-suffering-fools/
1•wahnfrieden•3m ago•0 comments

Can you solve the three cloaks enigma?

https://threecloaks.com/
1•ahmedchemkhy•3m ago•0 comments

Don't Be Fooled by this Summer of AI Hype

https://www.technologyreview.com/2026/09/22/1144867/dont-be-fooled-summer-ai-hype/
1•jacquesm•4m ago•0 comments

Si Sheppard – How did a few hundred Spanish soldiers topple two empires?

https://www.dwarkesh.com/p/si-sheppard
2•gmays•5m ago•0 comments

Tech in cars can be used to snoop on you, Dutch spy chiefs warn

https://www.theguardian.com/technology/2026/oct/01/cars-smart-features-spy-gps-chinese-imports
1•iamnothere•6m ago•0 comments

Show HN: Graphene – Data analysis toolkit for your coding agent

https://github.com/graphene-data/graphene
2•kcmarr•8m ago•0 comments

The Note – A Markdown notebook with runnable code and formula tables

https://tobwil.github.io/THENote/en/
1•tobias_digital•9m ago•0 comments

Crafting Express – isolated local environments for AI coding agents

https://github.com/crafting-dev/express
1•neham25•10m ago•0 comments

Slot Machine Programming and the Hidden Curriculum

https://thelastsoftwareengineer.substack.com/p/slot-machine-programming-and-the
1•azhenley•15m ago•0 comments

Show HN: A 100% Client-Side HAR Analyzer That Parses Logs in Seconds

https://har-analyzer.onlinetool.workers.dev/
5•kvmsolutionscon•15m ago•0 comments

Sampling to Reinforce

https://srush.github.io/sampling-to-reinforce/
1•sharma-arjun•16m ago•0 comments

CSS Bed: Classless CSS themes to use as starting points in web development

https://www.cssbed.com
3•sea-gold•17m ago•0 comments

Hermes ChatGPT Extension – Your Hermes Agents Inside Codex/ChatGPT

https://github.com/intellectronica/hermes-chatgpt-extension
2•intellectronica•17m ago•0 comments

Show HN: DOM-defense – turn your current webpage into a tower defense level

https://chromewebstore.google.com/detail/dom-defense/bglbhkdakecceocdpnegdchjmmomnopg
1•skellertor•18m ago•0 comments

Ask HN: What are you working on? (October 2026)

1•meerita•20m ago•4 comments

Saturn 1 (SA-5) Camera Inside Kerosene Tank During Launch [video]

https://www.youtube.com/watch?v=fL-Oi9m2beA
2•Teever•23m ago•0 comments

Micronaut Framework 5.2.0 Released

https://micronaut.io/2026/09/27/micronaut-framework-5-2-0/
1•ejboy•26m ago•0 comments

Monitoring my Samsung smart washing machine without a cloud

https://tildeweb.nl/~michiel/samsung-monitoring-my-washer-without-a-cloud.html
1•roywashere•27m ago•0 comments

The Specter of Neuralese

https://www.astralcodexten.com/p/the-specter-of-neuralese
2•eatitraw•28m ago•0 comments

Pretraining Latent Information Feedback Transformers with Teacher Supervision

https://arxiv.org/abs/2609.38149
1•gmays•29m ago•0 comments

Researchers can reconstruct what you're looking at from a brain scan

https://www.technologyreview.com/2026/10/01/1145588/ai-mind-reading-reconstructs-what-youre-looki...
4•kinjba11•30m ago•1 comments

Reducing the cognitive load of AI changes

https://amoffat.github.io/blog/cognitive-load.html
2•ibobev•30m ago•0 comments

Rust 1.99.0

https://blog.rust-lang.org/2026/10/01/Rust-1.99.0/
2•ibobev•30m ago•3 comments

The death of web development education

https://molily.de/web-dev-education/
20•ibobev•30m ago•7 comments

Single CPU Language Modeling Benchmark Shattered

https://spasim.org/docs/HutterPrize/TwentyYearBenchmarkShattered.html
1•jabowery•31m ago•1 comments

Show HN: Fumble – keybr, but it learns from everything you type (macOS)

https://fumble.alexmcconnell.ie
1•alexmcconnellyc•33m ago•0 comments

Who Remembers Etch a Sketch?

2•cozyss•36m ago•1 comments

Claude.dev Blog / technical writing for people building with Claude

https://claude.dev/
1•duck•42m ago•0 comments

AI as Normal Technology

https://knightcolumbia.org/content/ai-as-normal-technology
3•api•43m ago•0 comments

Vue – A WebAssembly Port of TeXmacs

https://mgubi.github.io/texmacs/
1•mgubi•48m ago•2 comments