frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

The Unreasonable Effectiveness of Reasonless Intermediate Tokens

https://arxiv.org/abs/2505.13775
4•YeGoblynQueenne•1y ago

Comments

tocs3•1y ago
I asked ChatGPT to restate this in more laymen's terms (posted below) and I am not to surprised at the answer.

"Lately, some AI models have shown impressive abilities to solve complex problems, and many people credit this to a method called Chain of Thought (CoT), where the model is trained to think through steps like a human might. In this paper, we take a closer look at that idea to see if it's really what's driving better performance.

We focus on the model’s step-by-step thinking (the words it generates along the way) — often treated like human "thoughts" — and examine whether these actually help the model solve problems more accurately. To test this, we train AI models using clean, correct step-by-step reasoning paths and final answers, all based on a known solving method (A* search). This lets us check both the final answers and the reasoning steps to see how they relate.

Interestingly, we find that even when a model gives the right answer, its reasoning steps can still be wrong or messy. To go further, we even train models using completely random and incorrect reasoning steps — and surprisingly, they still perform about the same, and sometimes even better, than those trained on correct steps.

This suggests that the step-by-step "thoughts" the model shows aren’t as meaningful or reliable as many assume. In short, just because a model looks like it’s reasoning through a problem doesn’t mean it actually is — and we should be careful not to treat its outputs as if it thinks like a human or follows strict logic."

LP Voting and Investor Consent for AIFs

https://taghash.io/blog/lp-voting-and-investor-consent-for-aifs-a-practical-workflow
1•koolhead17•1m ago•0 comments

Ukraine's secretive tech world building the future of war

https://www.fastcompany.com/91611099/inside-ukraines-secretive-tech-world-building-the-future-of-war
1•SLHamlet•1m ago•0 comments

Ezra Klein's Podcast with Jensen Huang

https://thezvi.substack.com/p/on-ezra-kleins-podcast-with-jensen
1•paulpauper•1m ago•0 comments

Were 1970s Korean Orphanages a Great Place to Raise Kids?

https://www.richardhanania.com/p/were-1970s-korean-orphanages-a-great
3•paulpauper•2m ago•0 comments

The 50-Year Hangover

https://freddiedeboer.substack.com/p/the-50-year-hangover
2•paulpauper•2m ago•0 comments

Dear Humanity – Can We Teach AI Compassion?

https://www.hussmanfunds.com/comment/ai_alignment_260917/
2•alch-•4m ago•0 comments

The Post-AGI Era

https://www.avidfayaz.com/writings/post-agi/the-post-agi-era
5•Avid_F•5m ago•0 comments

Intrinsic Uniqueness and Reconstruction Across Mathematical Presentations

https://zenodo.org/records/22775358
5•alex_albert•6m ago•0 comments

Sprites Connectors is a great example of Terrible API Design

https://www.freestyle.sh/blog/opinion/sprites-connectors-terrible-api-design
4•benswerd•8m ago•0 comments

Jev vs. Kev: open-source Jev alternative tested side by side

https://opper.ai/blog/jev-vs-kev-open-decision-model
6•felix089•8m ago•0 comments

MacTube 2 YouTube as a real Mac app – without Shorts and without AI slop

https://github.com/depo23/MacTube2
1•edib•8m ago•0 comments

Brutalist Asia

https://www.wallpaper.com/architecture/brutalist-asia-book
3•keiferski•8m ago•0 comments

Color generator based on modified Tree(3)

https://leviathan-colors.emergent.host/
1•Leosawin•9m ago•1 comments

"Thin Air," Real Money: The Trump Crypto Game

https://thinairrealmoney.com/
2•surprisetalk•10m ago•0 comments

Show HN: Reliopt – Pareto-frontier, contract-gated optimization for LLM programs

https://github.com/obielin/reliopt
1•arabking•12m ago•0 comments

Ask HN: Has anyone built a leaderless multi-agent system?

1•har-ki•12m ago•0 comments

Snap Wants to be a State Actor??–Kansas v. Snap

https://blog.ericgoldman.org/archives/2026/09/snap-wants-to-be-a-state-actor-kansas-v-snap.htm
3•hn_acker•12m ago•0 comments

Risk factors for androgenetic alopecia: a systematic review and analysis (2015)

https://link.springer.com/article/10.1186/s12889-026-26258-y
1•OutOfHere•14m ago•0 comments

A valuation-dislocation tracker for 34 public companies

https://michaelhillaert.com/
1•michaelhillaert•16m ago•0 comments

Prompt2ELF: An LLM wrote a 449-byte HTTP server without a compiler

https://github.com/faustinoaq/prompt2elf
1•totakaro•16m ago•0 comments

The CFTC Is Tying Its Own Hands on Prediction Markets

https://www.lawfaremedia.org/article/the-cftc-is-tying-its-own-hands-on-prediction-markets
1•hn_acker•16m ago•0 comments

The cheap new AI model taking aim at OpenAI and Anthropic

https://www.ft.com/content/456884ea-2558-4648-8036-a77b73733430
3•michaelhillaert•18m ago•0 comments

What Jev will do to data engineering

https://www.astronomer.io/blog/what-jev-will-do-to-data-engineering/
1•jlaneve•19m ago•0 comments

Show HN: I made a WWI dogfight game that runs in the browser

https://www.gamedev.pl/ay/biplane-skirmish
2•fullstackwife•19m ago•1 comments

Meteor M2-4: Receiving Weather Satellite Images from My Garden in San Francisco

https://gdamdam.github.io/meteor-satellite-reception/meteor-m2-4-2026-09-23.html
1•gidam•21m ago•0 comments

The Download: The Pentagon's AI-powered lie detector and young organ limits

https://www.technologyreview.com/2026/09/25/1145157/the-download-pentagon-ai-lie-detector-young-o...
1•joozio•23m ago•0 comments

Jupyter AI: A Map of 100 Jupyter Extensions for AI

https://openteams.com/awesome-jupyter-ai-extensions/
1•redsquirrel12•26m ago•0 comments

Kakao Entertainment to Shut Down N. American Webtoon Platform Tapas

https://www.animenewsnetwork.com/news/2026-09-22/kakao-entertainment-to-shut-down-n-american-webt...
1•speckx•27m ago•0 comments

First Triangulation Results in UAP Search by the Galileo Project Observatories

https://www.youtube.com/watch?v=HrvEbbqsN-s
1•musha68k•27m ago•1 comments

If the Work Is So Meh That It Might Be AI, Who Cares How It Was Made?

https://novelarcade.substack.com/p/whats-worse-than-ai-slop-trad-slop
1•richardatlarge•28m ago•1 comments