frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

The Unreasonable Effectiveness of Reasonless Intermediate Tokens

https://arxiv.org/abs/2505.13775
4•YeGoblynQueenne•1y ago

Comments

tocs3•1y ago
I asked ChatGPT to restate this in more laymen's terms (posted below) and I am not to surprised at the answer.

"Lately, some AI models have shown impressive abilities to solve complex problems, and many people credit this to a method called Chain of Thought (CoT), where the model is trained to think through steps like a human might. In this paper, we take a closer look at that idea to see if it's really what's driving better performance.

We focus on the model’s step-by-step thinking (the words it generates along the way) — often treated like human "thoughts" — and examine whether these actually help the model solve problems more accurately. To test this, we train AI models using clean, correct step-by-step reasoning paths and final answers, all based on a known solving method (A* search). This lets us check both the final answers and the reasoning steps to see how they relate.

Interestingly, we find that even when a model gives the right answer, its reasoning steps can still be wrong or messy. To go further, we even train models using completely random and incorrect reasoning steps — and surprisingly, they still perform about the same, and sometimes even better, than those trained on correct steps.

This suggests that the step-by-step "thoughts" the model shows aren’t as meaningful or reliable as many assume. In short, just because a model looks like it’s reasoning through a problem doesn’t mean it actually is — and we should be careful not to treat its outputs as if it thinks like a human or follows strict logic."

CortextAI – A private, offline AI operating system for personal productivity

https://cortextai.saposs.com
1•jimmy_lee•28s ago•0 comments

Kindle Direct publishing limits to 2 titles per format per week

https://www.kdpcommunity.com/s/article/KDP-Title-Creation-Limits-Update?language=en_US
1•speckx•1m ago•0 comments

Show HN: Arcentry – Create 3D infrastructure diagrams

https://arcentry.com/?hn
1•wolframhempel•3m ago•1 comments

I got 2.2x more tokens per second from llama.cpp on Intel Arc

https://grigio.org/how-i-got-2-2x-more-tokens-per-second-from-llama-cpp-on-intel-arc/
1•grigio•3m ago•0 comments

Show HN: Replaceme – a native Mac agent that texts you what needs your attention

https://replaceme.app
1•eliswed•4m ago•0 comments

Show HN: Subscrio – OSS subscription entitlements for .NET and TypeScript

https://github.com/subscrio/subscrio
1•feech•5m ago•0 comments

Show HN: Hntui – A TUI for Hacker News

https://github.com/ahmd-sh/hntui
1•ahmd-sh•8m ago•0 comments

The End of Hand-Written Code: A Conversation with David Heinemeier Hansson

https://thoughteconomics.com/david-heinemeier-hansson/
1•cebert•10m ago•0 comments

Neural Drive

https://mlx-optiq.com/blog/neural-drive-game
1•codelion•11m ago•0 comments

Chatbots give a narrow slice of knowledge:Researchers warn of knowledge collapse

https://news.ku.dk/all_news/2026/09/ai-chatbots-give-us-a-narrow-slice-of-knowledge-researchers-w...
1•DeepLogin•12m ago•0 comments

China Broadens Travel Curbs to Encompass Family of Top AI Talent

https://www.bloomberg.com/news/articles/2026-09-28/china-broadens-travel-curbs-to-encompass-famil...
1•sbulaev•12m ago•1 comments

Cohesix – find out what happened to a local AI job

https://github.com/lukeb-aidev/cohesix/releases/tag/v1.2.0
1•Cohesix•14m ago•0 comments

Richardson Maturity Model, Steps Toward the Glory of REST

https://martinfowler.com/articles/richardsonMaturityModel.html
1•vladde•17m ago•1 comments

Nvidia Launches Open Agent Safety Platform

https://mrkt30.com/nvidia-open-agent-safety-platform-openshell-sentry/
1•newscomAI•17m ago•0 comments

Vite+ 1.0 Is Out

https://voidzero.dev/posts/announcing-vite-plus-1-0
2•vanyle•17m ago•1 comments

The Largest Roman Mosaic Ever Excavated Opens to the Public

https://www.thisiscolossal.com/2026/09/largest-roman-mosaic-museum-rome/
1•surprisetalk•19m ago•0 comments

Space X Startship first orbital flight

https://www.nytimes.com/2026/09/28/science/space/spacex-starship-first-orbital-flight.html
1•ltononro•19m ago•0 comments

SlopTotal, a Self-hosted AI text detector that runs 23 open models

https://github.com/pablocaeg/sloptotal
3•sloptotal•19m ago•0 comments

Geely to Buy 30% Stake in Rival Nio's Battery-Swapping Unit

https://www.bloomberg.com/news/articles/2026-09-28/geely-to-buy-30-stake-in-rival-nio-s-battery-s...
1•teleforce•20m ago•0 comments

Fingerprints of Jev

https://jdhornsby.com/fingerprints-of-jev/
1•time0ut•21m ago•0 comments

Why Software Factories Fail [video]

https://www.youtube.com/watch?v=Ib5GBkD555M
1•mpweiher•22m ago•0 comments

Intro to Clojure Workshop [video]

https://www.youtube.com/watch?v=KwM1c7vb-fg
1•doubleg•22m ago•0 comments

Is an Agentic Bank Run Coming?

https://www.apollo.com/wealth/insights-news/insights/daily-spark/is-an-agentic-bank-run-coming
2•theanonymousone•23m ago•0 comments

My coding agent now records the demo video for its own PRs

https://github.com/half144/cutaway
1•half144•23m ago•1 comments

#59 Hans Kelsen and Carl Schmitt: Friend and Enemy

https://voelkerrechtsblog.org/59-hans-kelsen-and-carl-schmitt-friend-and-enemy/
1•jruohonen•25m ago•0 comments

Hospitals use AI to find more things to bill for. Insurers use AI to deny them.

https://twitter.com/HedgieMarkets/status/2104266774766039301
2•MrBuddyCasino•26m ago•0 comments

Parisians on Hunt for Baguettes as Bakers Get Nod to Take Vacation (2015)

https://www.npr.org/sections/thesalt/2015/08/25/433005099/parisians-on-hunt-for-baguettes-as-bake...
1•thunderbong•28m ago•0 comments

How NOT to detect residential proxies

https://blog.truesign.ai/posts/how-NOT-to-detect-residential-proxies/
9•kanokano•28m ago•3 comments

Oracle triggers 'force majeure' on data center project over power delays

https://www.reuters.com/business/oracle-cites-force-majeure-shield-itself-controversial-data-cent...
1•andyjohnson0•28m ago•0 comments

The case for Wile E. Coyote, engineer: how it shaped the way we build

https://school.coyotiv.com/blog-wile-e-coyote-engineer-en
1•arbayi•29m ago•0 comments