frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

The Unreasonable Effectiveness of Reasonless Intermediate Tokens

https://arxiv.org/abs/2505.13775
4•YeGoblynQueenne•1y ago

Comments

tocs3•1y ago
I asked ChatGPT to restate this in more laymen's terms (posted below) and I am not to surprised at the answer.

"Lately, some AI models have shown impressive abilities to solve complex problems, and many people credit this to a method called Chain of Thought (CoT), where the model is trained to think through steps like a human might. In this paper, we take a closer look at that idea to see if it's really what's driving better performance.

We focus on the model’s step-by-step thinking (the words it generates along the way) — often treated like human "thoughts" — and examine whether these actually help the model solve problems more accurately. To test this, we train AI models using clean, correct step-by-step reasoning paths and final answers, all based on a known solving method (A* search). This lets us check both the final answers and the reasoning steps to see how they relate.

Interestingly, we find that even when a model gives the right answer, its reasoning steps can still be wrong or messy. To go further, we even train models using completely random and incorrect reasoning steps — and surprisingly, they still perform about the same, and sometimes even better, than those trained on correct steps.

This suggests that the step-by-step "thoughts" the model shows aren’t as meaningful or reliable as many assume. In short, just because a model looks like it’s reasoning through a problem doesn’t mean it actually is — and we should be careful not to treat its outputs as if it thinks like a human or follows strict logic."

Taylor Farms faces recalls, six lawsuits, and a congressional inquiry

https://www.inc.com/amaya-nichole/taylor-farms-faces-scrutiny-on-every-front/91388694
1•speckx•33s ago•0 comments

BananaMind 2 Pro: Language Model Trained on a Consumer GPU

https://huggingface.co/BananaMind/BananaMind-2-Pro
1•banaxii•34s ago•0 comments

Self-Improving Agents Are Event-Sourced

https://lobu.ai/blog/self-improving-agents-are-event-sourced/
1•buremba•1m ago•0 comments

Capital Formation

https://pluralistic.net/2026/08/14/one-chokable-throat/
1•hn_acker•1m ago•0 comments

Silicon Valley Loves Jargon–and 'Hill-Climbing' Is Its Favorite New Phrase

https://www.wsj.com/tech/ai/silicon-valley-loves-jargonand-hill-climbing-is-its-favorite-new-phra...
1•apparent•2m ago•0 comments

Base Power Company: Manufacturing the Capability Cannon

https://www.notboring.co/p/base-power-company-chapter-3
1•figbert•2m ago•0 comments

Download error in Milwaukee delayed final results in Wisconsin primary election

https://www.jsonline.com/story/news/politics/elections/2026/08/12/download-error-in-milwaukee-del...
1•zahlman•3m ago•0 comments

People Are 'Marrying' Chatbots. These Lawmakers Want to Stop Them

https://www.wired.com/story/people-are-marrying-chatbots-these-lawmakers-want-to-stop-them/
2•jaredwiener•3m ago•0 comments

Every fucking website: 2026 edition

https://op.tngl.io/every-fucking-website/
1•nerdypepper•4m ago•1 comments

Show HN: Captain, AI Travel Agent

https://www.moonlight.ng/captain/
2•fathermerry•6m ago•0 comments

DSCI – self hosted Git server and CI for developers

http://dsci.sparrowhub.io/
1•melezhik•6m ago•0 comments

Apple Maps Rolls Out Ads, Targets Madison Avenue

https://www.mediapost.com/publications/article/417237/apple-officially-rolls-out-ads-in-maps.html
1•thm•9m ago•1 comments

You can now turn off Google Gemini's visible watermarks

https://www.theverge.com/tech/980416/google-gemini-ai-watermarks-removal
1•thm•10m ago•0 comments

Malaccamax

https://en.wikipedia.org/wiki/Malaccamax
2•teleforce•10m ago•0 comments

The modern economy is scam-powered

https://mindtheblindspot.substack.com/p/the-modern-economy-is-scam-powered
2•petethomas•10m ago•0 comments

Zooming Out

https://www.michaelbromley.co.uk/blog/zooming-out/
1•speckx•11m ago•0 comments

Show HN: Open-source and AI native web analytics

https://github.com/OpenLabs-so/openanalytics
3•rahulbridge•13m ago•0 comments

Six Facts about the Recent Employment Effects of Artificial Intelligence

https://digitaleconomy.stanford.edu/publication/canaries-in-the-coal-mine-six-facts-about-the-rec...
1•paulpauper•15m ago•0 comments

We urgently need a coherent national AI cybersecurity policy

https://joshuasaxe181906.substack.com/p/we-urgently-need-a-coherent-national
1•paulpauper•16m ago•0 comments

Racket v9.3

https://blog.racket-lang.org/2026/08/racket-v9-3.html
3•privong•16m ago•0 comments

She Makes the Impossible Happen for Her Ultrarich Clients

https://www.nytimes.com/2026/08/09/style/olivia-ferney-top-tier-travel.html
1•paulpauper•16m ago•0 comments

Hex Supply Customer Support Log

https://strangehorizons.com/wordpress/poetry/hex-supply-customer-support-log/
1•BiraIgnacio•17m ago•0 comments

About Tibo's Reset

https://zrainbow.substack.com/p/about-tibos-resets
1•ZRainbow•17m ago•0 comments

2026 Hugo Awards

https://www.thehugoawards.org/hugo-history/2026-hugo-awards/
1•BiraIgnacio•17m ago•0 comments

Free Speech VPS Providers Put to the Test

https://crippled.media/article/free-speech-vps-providers-put-to-the-test
1•Cider9986•18m ago•0 comments

A whale's dive shapes its song

https://www.eurekalert.org/news-releases/1138432
1•gmays•19m ago•0 comments

Tokens per second visualized as the programming language balls

https://kylejeong.com/inference-balls
1•Kylejeong21•19m ago•0 comments

List of predictions for autonomous Tesla vehicles by Elon Musk

https://en.wikipedia.org/wiki/List_of_predictions_for_autonomous_Tesla_vehicles_by_Elon_Musk
3•ceejayoz•21m ago•0 comments

Two flights, one call sign: ATC resolves dangerous situation in Phoenix

https://www.flightradar24.com/blog/aviation-news/two-flights-one-call-sign-atc-resolves-dangerous...
1•berkeleyjunk•21m ago•0 comments

Millions sent wildfire emergency phone alert

https://www.bbc.co.uk/news/live/cmdx7p1q02yrt
6•basisword•21m ago•4 comments