frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

The Unreasonable Effectiveness of Reasonless Intermediate Tokens

https://arxiv.org/abs/2505.13775
4•YeGoblynQueenne•1y ago

Comments

tocs3•1y ago
I asked ChatGPT to restate this in more laymen's terms (posted below) and I am not to surprised at the answer.

"Lately, some AI models have shown impressive abilities to solve complex problems, and many people credit this to a method called Chain of Thought (CoT), where the model is trained to think through steps like a human might. In this paper, we take a closer look at that idea to see if it's really what's driving better performance.

We focus on the model’s step-by-step thinking (the words it generates along the way) — often treated like human "thoughts" — and examine whether these actually help the model solve problems more accurately. To test this, we train AI models using clean, correct step-by-step reasoning paths and final answers, all based on a known solving method (A* search). This lets us check both the final answers and the reasoning steps to see how they relate.

Interestingly, we find that even when a model gives the right answer, its reasoning steps can still be wrong or messy. To go further, we even train models using completely random and incorrect reasoning steps — and surprisingly, they still perform about the same, and sometimes even better, than those trained on correct steps.

This suggests that the step-by-step "thoughts" the model shows aren’t as meaningful or reliable as many assume. In short, just because a model looks like it’s reasoning through a problem doesn’t mean it actually is — and we should be careful not to treat its outputs as if it thinks like a human or follows strict logic."

Show HN: Reviving my $15,000 browser arcade with Astra

https://eggverse.co/
1•Waseemkhalo•2m ago•0 comments

Show HN: Socks Proxy for AI Agents

https://x402socks.com/
1•rnts08•3m ago•0 comments

Anthropic details how Claude was misused for surveillance and weapons

https://thenextweb.com/news/anthropic-claude-misuse-threat-intelligence-report
1•delichon•5m ago•0 comments

Turn Your Data into a Video

https://www.racecharts.io/
1•ijidak•5m ago•0 comments

INDUS-SDE: A Language Model for Scientific Content Curation and Discovery

https://science.data.nasa.gov/blog/indus-sde-language-model
1•sudobear•6m ago•1 comments

Steam's Programming Fest

https://store.steampowered.com/category/programming
3•embedding-shape•8m ago•0 comments

Portal 2 Beta Decompilation

https://github.com/apersonwhomakesstuff/Portal-2-Beta-Decompilation/tree/main
1•apersonwhomakes•11m ago•1 comments

Chick-Fil-A Sauce Flavored Waffle Potato Chips

https://www.chick-fil-a.com/menu/sides/chick-fil-a-sauce-flavored-waffle-potato-chips
1•SpyCoder77•11m ago•0 comments

ASCII arcade: 3D engine for the terminal

https://ascii-arcade.dev
1•cramforce•11m ago•0 comments

A Decent GPU for FreeBSD and OpenBSD (2023)

https://medium.com/@martin.baulig/a-decent-gpu-for-freebsd-and-openbsd-ee82e6ff219a
1•turtleyacht•12m ago•0 comments

Generative AI Using Linuxulator and eGPU on FreeBSD

https://www.tumfatig.net/2026/generative-ai-using-linuxulator-and-egpu-on-freebsd/
1•turtleyacht•13m ago•0 comments

Sean Carroll explains the biggest ideas in the universe – Full Interview [video]

https://www.youtube.com/watch?v=_TBNJyztai0
4•binyu•13m ago•0 comments

A $537 Local LLM Machine (2025)

https://blog.lewman.com/a-537-local-llm-machine.html
2•turtleyacht•14m ago•0 comments

Synapse – Daily Word Game

https://synapse.akshayr.xyz/
1•barbierocks•15m ago•1 comments

Russia trafficking foreigners to fight in Ukraine, Amnesty says

https://www.courthousenews.com/russia-trafficking-foreigners-to-fight-in-ukraine-amnesty-says/
2•MilnerRoute•16m ago•0 comments

The Confused Engineer

https://thebsq.com/the-confused-engineer/
2•ahofma•17m ago•1 comments

M. Williams, OpenAI: Human extinction in the next few years seems likely

https://xcancel.com/antibot/captcha
2•doener•18m ago•1 comments

Show HN: PRBar see your open GitHub PR count in macOS menu bar

https://github.com/sburl/prBar
1•sburl•19m ago•1 comments

The Chasm: The Shape of Unfinished AI Codebases

https://jimmyhmiller.com/shape-of-unfinished-ai-codebases
1•gmays•21m ago•0 comments

AI researchers leave Anthropic and Google: 'There are no adults in the room'

https://www.nbcnews.com/tech/security/two-ai-researchers-leave-anthropic-google-safety-concerns-r...
3•doener•22m ago•0 comments

Stop externalizing the cost of your AI use to me

https://thelastsoftwareengineer.substack.com/p/stop-externalizing-the-cost-of-your
5•azhenley•24m ago•0 comments

AI Is Not Going to Kill My Love of Math

https://chillphysicsenjoyer.substack.com/p/ai-is-not-going-to-kill-my-love-of
5•crescit_eundo•25m ago•0 comments

NATO allies foil Russian subsea cable sabotage plot

https://www.reuters.com/world/europe/nato-allies-foil-russian-subsea-cable-sabotage-plot-2026-09-10/
2•doener•29m ago•0 comments

Xapien raises $56M Series B for AI background checks and due diligence

https://axios.com/pro/all-deals/2026/09/10/due-diligence-ai-xapien-56-million
1•utiiiD•30m ago•2 comments

Biff 2.0 Is Released

https://biffweb.com/p/biff2-released/
2•TheWiggles•30m ago•0 comments

The Birth of HaaS

https://www.vtrivedy.com/posts/claude-code-sdk-haas-harness-as-a-service/
1•iacguy•30m ago•0 comments

More Anthropic researchers warn of AI's perils as Musk terms fears a 'psyop'

https://www.theguardian.com/technology/2026/sep/10/anthropic-researchers-warn-ai-musk
2•lf88•33m ago•2 comments

Julia 1.13 Highlights

https://julialang.org/blog/2026/09/julia-1.13-highlights/
3•postflopclarity•34m ago•0 comments

The Rise and Fall of Crime

https://nicholasdecker.substack.com/p/the-rise-and-fall-of-crime
1•barry-cotter•34m ago•0 comments

Thelio Mira AI Linux Workstation: 192 GB GPU Memory

https://system76.com/workstations/thelio-mira-ai
5•jonifico•35m ago•0 comments