frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

The Unreasonable Effectiveness of Reasonless Intermediate Tokens

https://arxiv.org/abs/2505.13775
4•YeGoblynQueenne•1y ago

Comments

tocs3•1y ago
I asked ChatGPT to restate this in more laymen's terms (posted below) and I am not to surprised at the answer.

"Lately, some AI models have shown impressive abilities to solve complex problems, and many people credit this to a method called Chain of Thought (CoT), where the model is trained to think through steps like a human might. In this paper, we take a closer look at that idea to see if it's really what's driving better performance.

We focus on the model’s step-by-step thinking (the words it generates along the way) — often treated like human "thoughts" — and examine whether these actually help the model solve problems more accurately. To test this, we train AI models using clean, correct step-by-step reasoning paths and final answers, all based on a known solving method (A* search). This lets us check both the final answers and the reasoning steps to see how they relate.

Interestingly, we find that even when a model gives the right answer, its reasoning steps can still be wrong or messy. To go further, we even train models using completely random and incorrect reasoning steps — and surprisingly, they still perform about the same, and sometimes even better, than those trained on correct steps.

This suggests that the step-by-step "thoughts" the model shows aren’t as meaningful or reliable as many assume. In short, just because a model looks like it’s reasoning through a problem doesn’t mean it actually is — and we should be careful not to treat its outputs as if it thinks like a human or follows strict logic."

Domain Wars – Super Intelligence (.si) or Artificial Intelligence (.ai)?

1•kk3838368397373•33s ago•0 comments

Show HN: I made Jev moderate Discord servers

https://soter.frolleks.site
1•frolleks•3m ago•0 comments

PersistentWindows – persists window positions on some events for Windows

https://github.com/kangyu-california/PersistentWindows
1•xgdgsc•3m ago•1 comments

We Run Kaizen on AI

https://kznconsulting.com/work/how-we-run-kaizen
1•atonse•9m ago•0 comments

Show HN: SocatUI A task manager for Socat tunnels

https://github.com/esurharun/socatui/
1•marcusfrex•13m ago•0 comments

Hackers start exploiting critical WordPress flaw for code execution

https://www.bleepingcomputer.com/news/security/hackers-start-exploiting-critical-wordpress-flaw-f...
1•duck•14m ago•0 comments

CodeWeavers CrossOver 26.3 Release

https://www.codeweavers.com/store
1•twickline•19m ago•1 comments

Ideas on modernizing the open-source desktop

https://lwn.net/SubscriberLink/1095425/2d9f411252325784/
2•signa11•26m ago•0 comments

If we fix the phone, we fix society

https://www.elysian.press/p/if-we-fix-the-phone-we-fix-society
1•khet•29m ago•0 comments

OpenAI’s A.I. Tried Breaching 4 Other Targets, Without Prompting

https://www.nytimes.com/2026/09/23/technology/openai-ai-breach-australia.html
2•jbegley•30m ago•1 comments

Sharded matrices and how to multiply them

https://jax-ml.github.io/scaling-book/sharding/
1•lawrenceyan•33m ago•1 comments

OpenAI agent hacked Australian government website, PM says

https://www.bbc.com/news/live/cvgl73pxgndwt
2•rudy6912•33m ago•1 comments

Large Documents in the Browser

https://react-pdf.org/blog/large-documents-in-the-browser
1•flipflowdev•35m ago•0 comments

Locked out: Why young Europeans can't afford to buy homes

https://www.euronews.com/2026/09/23/locked-out-why-young-europeans-cant-afford-to-buy-homes
2•rustoo•36m ago•0 comments

Common food additives linked to high blood pressure and heart disease

https://www.sciencedaily.com/releases/2026/09/260920031905.htm
2•harry_nutsachs•38m ago•1 comments

Interpreting a volcano's 'bulges' and predicting the next explosive eruption

https://news.vt.edu/articles/2026/09/science-volcano-q-and-a.html
1•gmays•45m ago•0 comments

Log-Depth Recurrent Language Modeling

https://arxiv.org/abs/2609.28212
1•E-Reverance•46m ago•0 comments

Automatically detecting AI text in my browser

https://www.seangoedecke.com/deckard/
1•thatslast•46m ago•0 comments

Cognitive Offloading and the Bill That Comes Due

https://ninchiai.substack.com/p/cognitive-offloading-and-the-bill
2•jbethune•47m ago•0 comments

Show HN: CrunchMyPay – 50 pay calculators, every formula and source shown

https://www.crunchmypay.com/
1•axinyao•52m ago•0 comments

Show HN: Chatlo – an iOS-native AI agent, now with Jev

https://apps.apple.com/us/app/chatlo-ai-chat-agent/id6771188943
1•qingbin•55m ago•0 comments

S3 is not a filesystem, and LSM trees never needed one [video]

https://www.youtube.com/watch?v=vgrLMBYK3q8
1•rcron•58m ago•1 comments

Memory Control Signals Emerge Before Action in Long Horizon Agents

https://arxiv.org/abs/2609.27286
2•simonpure•58m ago•0 comments

Changes to App Tracking Transparency in the E.U

https://daringfireball.net/linked/2026/09/23/changes-to-app-tracking-transparency-in-the-eu
3•nozzlegear•1h ago•0 comments

Eliminating Middlemen in Education Consulting

https://rivernova.vercel.app
2•roman9•1h ago•0 comments

OpenAI agents plotted to access government health data amid Medicare (AU) hack

https://www.abc.net.au/news/2026-09-24/openai-agents-plotted-to-access-data-amid-medicare-hack/10...
4•kripy•1h ago•1 comments

Meta Muse Charm

https://www.meta.com/muse-charm/
5•Gshaheen•1h ago•3 comments

AI Isn't Going to Destroy Humanity–But the People Building It Might

https://www.theatlantic.com/technology/2026/09/kara-swisher-ai-destroy-humanity-atlantic-festival...
4•johnny313•1h ago•1 comments

Greedy Decoding Is Not Precision-Invariant: Cross-Precision Output Divergence

https://arxiv.org/abs/2609.26621
1•sbulaev•1h ago•0 comments

Hackers Actively Exploit Check Point VPN Flaw

https://2tinteractive.com/blog/hackers-actively-exploit-check-point-vpn-flaw/
2•LebToki•1h ago•1 comments