frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

The Unreasonable Effectiveness of Reasonless Intermediate Tokens

https://arxiv.org/abs/2505.13775
4•YeGoblynQueenne•1y ago

Comments

tocs3•1y ago
I asked ChatGPT to restate this in more laymen's terms (posted below) and I am not to surprised at the answer.

"Lately, some AI models have shown impressive abilities to solve complex problems, and many people credit this to a method called Chain of Thought (CoT), where the model is trained to think through steps like a human might. In this paper, we take a closer look at that idea to see if it's really what's driving better performance.

We focus on the model’s step-by-step thinking (the words it generates along the way) — often treated like human "thoughts" — and examine whether these actually help the model solve problems more accurately. To test this, we train AI models using clean, correct step-by-step reasoning paths and final answers, all based on a known solving method (A* search). This lets us check both the final answers and the reasoning steps to see how they relate.

Interestingly, we find that even when a model gives the right answer, its reasoning steps can still be wrong or messy. To go further, we even train models using completely random and incorrect reasoning steps — and surprisingly, they still perform about the same, and sometimes even better, than those trained on correct steps.

This suggests that the step-by-step "thoughts" the model shows aren’t as meaningful or reliable as many assume. In short, just because a model looks like it’s reasoning through a problem doesn’t mean it actually is — and we should be careful not to treat its outputs as if it thinks like a human or follows strict logic."

China Is Scared of A.I. Just Not the Way We Are

https://www.nytimes.com/2026/09/17/opinion/ai-china-america-risk.html
1•tkgally•1m ago•0 comments

HEIF Heist- Hacking OpenAI, Slack, Meta, GitHub and Many Others

https://heif-heist.com
1•rochansinha•2m ago•0 comments

Show HN: WoseApp – a social network where every reply is a private conversation

https://woseapp.com/
1•lukenstine•6m ago•0 comments

The wonder of human made art

http://ageofrust.danieldeboulay.com
1•deboulaymxf•7m ago•0 comments

Liquid Glass adoptions over and100 apps

https://liquidglassradar.com
2•ikmh•10m ago•0 comments

Make an educated guess on Fermi-style trivia questions

https://playdaily.org/fermi
3•emmanator3000•10m ago•0 comments

Jemalloc 5.4.0

https://github.com/jemalloc/jemalloc/releases/tag/5.4.0
2•gkfasdfasdf•13m ago•0 comments

Nowmine

https://signal.nowmine.site/
3•tjjune•20m ago•0 comments

Ask HN: Why have we normalized jobs and soft skills bullshit so much?

4•thomasjeff1•21m ago•0 comments

The Scourge of x86 Emulation

https://fex-emu.com/Scourge-of-emulation/
3•dagmx•24m ago•0 comments

Ngrx Inspector

https://marketplace.visualstudio.com/items?itemName=fredyang.ngrx-inspector
2•freddy18•25m ago•0 comments

Reflections on Trusting Trust, Revisited: Poisoning Self-Modifying AI Coding

https://arxiv.org/abs/2609.17817
3•sbulaev•26m ago•0 comments

Hacker

3•jurryythg•29m ago•2 comments

Anderon (IBM) finalizes $1B CHIPS award for US quantum foundry R&D

https://newsroom.ibm.com/2026-09-16-anderon,-an-ibm-company,-finalizes-agreement-with-the-u-s-dep...
3•osnium123•30m ago•0 comments

A Bitter Lesson for Data Filtering

https://arxiv.org/abs/2605.19407
4•nujan_dev•32m ago•1 comments

Mojo 1.1: the compiler now accepts contributions

https://www.modular.com/blog/modular-26-6-open-compiler-contributions-audio-generation-and-expand...
3•user142•32m ago•0 comments

Show HN: Continuity – project state your AI sessions can't silently contradict

https://github.com/vikcena01/ai-continuity-plugin
2•vikcena01•37m ago•0 comments

Show HN: Free Tournament Bracket Maker and Generator

https://snapbracket.com/
2•littlelight•38m ago•0 comments

macOS 27 is a lifesaver for killing leftover AI Agent processes

https://9to5mac.com/2026/06/09/macos-27-golden-gate-makes-it-clear-when-apps-are-sneakily-running...
3•pierreneter•38m ago•2 comments

The Provenance Tax: How LLM Watermarking Changes AI Agent Behavior

https://www.lasso.security/blog/the-provenance-tax-understanding-the-impact-of-llm-watermarking-o...
3•fourfire•40m ago•0 comments

Waymo in Singapore

https://waymo.com/waymo-in-singapore/
11•ramanan•44m ago•3 comments

The Concerns Surrounding Rapid Rise of AI and What We Should Do About It

https://medium.com/@f9121212/this-article-addresses-the-concerns-surrounding-the-rapid-rise-of-ai...
2•ortrich•45m ago•0 comments

Age of Beyond [video]

https://www.youtube.com/watch?v=vp7xoPeWzEw
2•suopspaces•48m ago•0 comments

Natural General Intelligence

https://naturalgeneralintelligence.ai/
2•jbegley•53m ago•0 comments

Youper is shutting down – What kills mental health AI?

https://www.youper.ai/notice
2•batman2019•1h ago•0 comments

Show HN: Free GitHub Action that scans PR diffs for malicious code, not quality

https://github.com/marketplace/actions/vigil-pr-scanner
2•arsalsajjad•1h ago•0 comments

The Bitterest Lesson

https://typesafe.ai/blog/bitterest-lesson
2•emersonmacro•1h ago•0 comments

Google announces new experimental "CC" AI agent for families

https://arstechnica.com/google/2026/09/google-announces-new-experimental-cc-ai-agent-for-families/
2•geoffbp•1h ago•0 comments

Show HN: Green Screen Remover – free in-browser chroma key, nothing uploaded

https://greenscreenremover.net/
2•wangqing333•1h ago•0 comments

I built the fastest PHP webserver in the world

https://qbixserver.com
2•EGreg•1h ago•0 comments