frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Show HN: A Linux terminal agent with transparent context, permissions, and logs

https://github.com/ljedrz/nachalnik/tree/master/kamchatka
1•ljedrz•39s ago•0 comments

Compositing in Wayland: A Technical Breakdown

https://openlib.io/compositing-in-wayland-a-technical-breakdown-in-linux/
1•ankitg12•3m ago•0 comments

Scientists Invent Underwater Umbrellas to Protect Coral Reefs–and It's Working

https://gizmodo.com/scientists-invent-underwater-umbrellas-to-protect-coral-reefs-and-it-appears-...
1•giuliomagnifico•3m ago•0 comments

ZKapi

https://twitter.com/ethereumfndn/status/2105753290977980682
1•sethx•4m ago•0 comments

Free Coding Agent With Unlimited frontier models

https://github.com/office233/free-coding-agent
1•abel53•6m ago•0 comments

Tokyo court grants legal protection to human voices in AI clone case

https://apnews.com/article/japan-anime-voice-actor-ai-a2a9596a197f149398771a9166043ef9
1•geox•7m ago•0 comments

Show HN: Header Modifier Pro – A $19 pay-once alternative to Requestly

https://chromewebstore.google.com/detail/header-modifier-pro/onhbgdmpmjmpghendkimpdcbnhlioegd
1•pale_gear•8m ago•0 comments

Generic Const Args and You

https://blog.rust-lang.org/inside-rust/2026/10/02/generic-const-args-and-you/
1•throawayonthe•11m ago•0 comments

Show HN: LaunchPact – Get support for your Product Hunt launch

https://www.launchpact.io
1•launchpact_io•12m ago•0 comments

OpenAI hacks 2nd Australian Government Department

https://www.abc.net.au/news/2026-10-02/rogue-open-ai-agent-breach-nsw-government-website/107223108
2•killingtime74•12m ago•1 comments

CompatNav–a free WordPress plugin that tells you what breaks on a PHP upgrade

https://wordpress.org/plugins/compatnav-php-upgrade-checker/
1•wpwp•15m ago•0 comments

Christophe Pettus: All Your GUCs in a Row: Max_wal_senders

https://thebuild.com/blog/all-your-gucs-in-a-row-max_wal_senders/
1•hnrprtlpdb•16m ago•0 comments

Turning to AI for health answers may signal anxiety, depression in young adults

https://news.ufl.edu/2026/09/ai-health/
1•giuliomagnifico•16m ago•0 comments

New feed manager and built-in reader are now available natively within Waterfox

https://www.waterfox.com/releases/6.7.5/
2•hggh•16m ago•0 comments

Anthropic Just Filed to Go Public. Does the Math Add Up?

https://www.youtube.com/watch?v=T-oXyXwD6sE
1•Betelbuddy•16m ago•0 comments

Robert Seacord: "Unsafe Rust" – RustConf 2026 [video]

https://www.youtube.com/watch?v=bTAEymCp7Ts
1•pjmlp•18m ago•0 comments

Using AI to chart a course for our post-quantum migration

https://blog.cloudflare.com/ai-driven-cryptography-discovery/
1•msgchainhq•19m ago•0 comments

Taliban suspends fiber optic network maintenance and expansion

https://www.afintl.com/en/202608270539
3•walrus01•24m ago•0 comments

Celld: Self-hosted, distributed Durable Objects

https://flaviocopes.com/celld/
1•handfuloflight•26m ago•0 comments

New Python Releases: 3.10.x Reaches EOL, 3.15.0 Scheduled Today

https://blog.python.org/2026/10/python-31022-31117/
2•birdculture•27m ago•0 comments

Too Many GCache Page Files in MySQL Data Directory

https://www.percona.com/blog/too-many-gcache-page-files-in-mysql-data-directory/
2•hn3ufz62f7•29m ago•0 comments

As US relations fray, Canada gets serious about its own launch industry

https://arstechnica.com/space/2026/10/as-us-relations-fray-canada-gets-serious-about-its-own-laun...
3•rbanffy•34m ago•0 comments

Claude and the Un-Man

https://www.onhumansandagents.com/p/claude-and-the-un-man
2•mikemangialardi•36m ago•0 comments

VCs are funding a concert business as a hedge against AI

https://www.businessinsider.com/vcs-funding-rasa-world-concerts-as-a-hedge-against-ai-2026-9
3•ilreb•36m ago•0 comments

SpaceX describes surgical intervention before launch of latest crew mission

https://arstechnica.com/space/2026/10/spacex-describes-surgical-intervention-before-launch-of-lat...
1•rbanffy•37m ago•0 comments

Mercury Decide – Inception's structured decision model

https://twitter.com/_inception_ai/status/2105375133003325768
1•Topfi•39m ago•1 comments

DisplayLink Direct SDK – USB Connected Displays

https://github.com/DisplayLink/displaylink-direct
1•greenmoon-5•39m ago•0 comments

GPT-6 Astra Robot Agents with 14% Higher Success Rate but 65% Fewer Tokens

https://dagroup-pku.github.io/PyRUA-Lean/
2•ilreb•39m ago•0 comments

C64 Mercenary: a novel exploit bug in a 41-year-old game

https://gamesexplained.com/c64/mercenary/#lift
1•thunderbong•41m ago•0 comments

I built a 3D planet for agents

https://terrabots.ai
1•mrjuchara•42m ago•3 comments
Open in hackernews

The Unreasonable Effectiveness of Reasonless Intermediate Tokens

https://arxiv.org/abs/2505.13775
4•YeGoblynQueenne•1y ago

Comments

tocs3•1y ago
I asked ChatGPT to restate this in more laymen's terms (posted below) and I am not to surprised at the answer.

"Lately, some AI models have shown impressive abilities to solve complex problems, and many people credit this to a method called Chain of Thought (CoT), where the model is trained to think through steps like a human might. In this paper, we take a closer look at that idea to see if it's really what's driving better performance.

We focus on the model’s step-by-step thinking (the words it generates along the way) — often treated like human "thoughts" — and examine whether these actually help the model solve problems more accurately. To test this, we train AI models using clean, correct step-by-step reasoning paths and final answers, all based on a known solving method (A* search). This lets us check both the final answers and the reasoning steps to see how they relate.

Interestingly, we find that even when a model gives the right answer, its reasoning steps can still be wrong or messy. To go further, we even train models using completely random and incorrect reasoning steps — and surprisingly, they still perform about the same, and sometimes even better, than those trained on correct steps.

This suggests that the step-by-step "thoughts" the model shows aren’t as meaningful or reliable as many assume. In short, just because a model looks like it’s reasoning through a problem doesn’t mean it actually is — and we should be careful not to treat its outputs as if it thinks like a human or follows strict logic."