frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

The Unreasonable Effectiveness of Reasonless Intermediate Tokens

https://arxiv.org/abs/2505.13775
4•YeGoblynQueenne•1y ago

Comments

tocs3•1y ago
I asked ChatGPT to restate this in more laymen's terms (posted below) and I am not to surprised at the answer.

"Lately, some AI models have shown impressive abilities to solve complex problems, and many people credit this to a method called Chain of Thought (CoT), where the model is trained to think through steps like a human might. In this paper, we take a closer look at that idea to see if it's really what's driving better performance.

We focus on the model’s step-by-step thinking (the words it generates along the way) — often treated like human "thoughts" — and examine whether these actually help the model solve problems more accurately. To test this, we train AI models using clean, correct step-by-step reasoning paths and final answers, all based on a known solving method (A* search). This lets us check both the final answers and the reasoning steps to see how they relate.

Interestingly, we find that even when a model gives the right answer, its reasoning steps can still be wrong or messy. To go further, we even train models using completely random and incorrect reasoning steps — and surprisingly, they still perform about the same, and sometimes even better, than those trained on correct steps.

This suggests that the step-by-step "thoughts" the model shows aren’t as meaningful or reliable as many assume. In short, just because a model looks like it’s reasoning through a problem doesn’t mean it actually is — and we should be careful not to treat its outputs as if it thinks like a human or follows strict logic."

Vipassana, Volini, and Viktor Frankl

https://jatin564991.substack.com/p/vipassana-volini-and-viktor-frankl
1•jatinarora26•7m ago•0 comments

If math is more than proof, we need to better celebrate the rest of it

https://terrytao.wordpress.com/2026/09/18/if-math-is-more-than-proof-we-need-to-better-celebrate-...
1•num42•9m ago•0 comments

Apple M6 Pro Achieves the Highest Single-Core CPU Score in Geekbench 7

https://browser.geekbench.com/v7/cpu/389219
2•gainsurier•17m ago•0 comments

Conway's Law and Programming Languages

https://danieltan.weblog.lol/2026/09/conways-law-and-programming-languages
1•signa11•18m ago•0 comments

Members of right‑leaning parties prefer leaders with dark triad personality

https://theconversation.com/members-of-right-leaning-parties-prefer-leaders-with-dark-triad-perso...
2•jbotz•20m ago•0 comments

Being a Responsible Human in the Loop

https://raahelbaig.com/entry/responsible-human-in-the-loop/
2•raahelb•23m ago•0 comments

Google's Gemini AI hacked three companies in security test

https://www.bbc.co.uk/news/articles/c607l0k72rlvo
3•luxpir•27m ago•0 comments

Splash: A Local Engine Built Around the Model

https://inco.ai/blog/splash/
1•namjh•30m ago•1 comments

Ask HN: DeepSeek v4.1 Flash on Local hardware, what tok/s do you see?

2•Argonautlabs•31m ago•0 comments

Sudo and OpenDoas Timestamp Files

https://xn--1xa.duncano.de/sudo-doas-timestamp-files
1•signa11•32m ago•0 comments

The C++20's u8/char8_t Backward-Compatibility Fiasco

https://giodicanio.com/2026/09/11/the-c-plus-plus-20-s-u8-char8_t-fiasco/
1•signa11•38m ago•0 comments

Show HN: Gotedo Impress – Free, Modern Presentation Software

https://about.gotedo.com/en/products/gotedo-impress
1•_ndianabasi•39m ago•1 comments

Grammarly will send unhinged messages to all your users if you try to cancel

https://old.reddit.com/r/sysadmin/comments/1wjdpgx/psa_grammarly_will_send_unhinged_messages_to_all/
2•johnnyApplePRNG•40m ago•0 comments

Inside the AI "Kill Chain" That Destroyed an Iranian School

https://www.bloomberg.com/graphics/2026-iran-school-attack/
2•xeonmc•41m ago•1 comments

China's token consumption to reach 100 quadrillion in 2026, says report

https://www.wicinternet.org/2026-09/15/c_1213473.htm
2•geox•42m ago•2 comments

Show HN: Nordstjernen Web Browser 1.0.24

2•roschdal•45m ago•0 comments

Microsoft agentically ports Copilot runtime to Rust for $120K

https://www.theregister.com/devops/2026/09/18/microsoft-agentically-ports-copilot-runtime-to-rust...
1•satvikpendem•47m ago•2 comments

Human brain is two separate organs, Stanford Medicine-led research finds

https://med.stanford.edu/news/all-news/2026/09/two-separate-brains.html
42•emigre•48m ago•7 comments

From 150 people to one system

https://chrisveleris.com/voice-of-the-machine/from-150-people-to-one-system/
2•cvicpp123•54m ago•0 comments

Flet 1.0 Is Here

https://flet.dev/blog/flet-1-0/
5•ferryth•54m ago•0 comments

Stepfun Step 5 Preview (LLM): On AA Pareto frontier

https://artificialanalysis.ai/models/step-5
2•AnodicElegy•54m ago•1 comments

OmniFire - All-in-one platform for social media, business profile, & mobile apps

https://www.omnifire.app/
1•xianglinkong•55m ago•0 comments

The Evidence for AI Consciousness, Today (2025)

https://newsletter.ai-frontiers.org/p/the-evidence-for-ai-consciousness
4•mellosouls•1h ago•1 comments

The original Linux distro source

https://www.kernel.org/pub/linux/kernel/v1.0/
2•maxxxxxxxxxxxxx•1h ago•0 comments

Speaking with the Mind – Neuralink [video]

https://www.youtube.com/watch?v=_j806JHhCRo
1•Eridanus2•1h ago•0 comments

Dylan Patel's Information Machine

https://substratemag.com/dylanpatel/
1•gmays•1h ago•0 comments

Ask HN: Remember N8n? Anyone?

3•sankalpdomore•1h ago•0 comments

Extract all comments from a Google Sheet

1•bsunter2•1h ago•0 comments

Show HN: I made Niri-like terminal colm (cmux alternative)

https://colm.sh/
2•al3rez•1h ago•0 comments

NASA-IBM Lunar Foundation open-Source Geospatial AI Model

https://newsroom.usra.edu/usra-contributes-planetary-science-expertise-to-nasa-ibm-lunar-foundati...
10•noobplus•1h ago•0 comments