frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Controlling Reasoning Effort in LLMs

https://magazine.sebastianraschka.com/p/controlling-reasoning-effort-in-llms
64•ibobev•14h ago

Comments

simonw•10h ago
I'm amused by how the whole reasoning model thing feels like a formalization of the old "think step by step" prompting hack, which was discovered against GPT-3 two years after that model was first released.

My favorite trick for controlling the reasoning level is the hack where you look at the output token stream and spot the token for "the model has concluded reasoning"... and then replace that with the tokens for "wait, but" and force it to keep going!

embedding-shape•7h ago
> My favorite trick for controlling the reasoning level is the hack

First described here I think, specifically for Qwen models: https://arxiv.org/abs/2501.19393, although it seems unclear how much it actually improves the final response unless you create a new fine-tuned model with this in mind, and also if it's possible to replicate for other models. Still, got inspired and gonna give it a try with DiffusionGemma although slightly differently

sva_•9h ago
I recommend his book, "Build a Reasoning Model (From Scratch)", which is also linked in the article.

https://sebastianraschka.com/books/#build-a-reasoning-model-...

razorbeamz•4h ago
LLMs don't actually reason, they just create the illusion of reasoning by repeating things over and over.
cyanydeez•2h ago
mmm, practically, llamacpp solved this with reasoning budget for tokens and a customized message.

ive tailored my local models to match agent with a message that pushes to use subagents and context compression.

it works pretty smoothly if theres proper vector and scope to the project. theres probably a way to also get it to record memories but ive not seen a memory system that includes spatial type reasoning which would create memories associated with file paths down to AST graphlike leaves.

then we could include remembering details about the path.

llamacpp also has a header for setting these that an intelligent harness could tailor per round message and budget. ideally you start with a small budget and expand based on some complexity criteria.

alas, i want to build things not related to AI.

A Koi Pond Mosaic Made from 10 Pounds of 3D Printer Waste

https://www.instructables.com/A-Koi-Pond-Mosaic-Made-From-10-Pounds-of-3D-Printe/
14•sudo_cowsay•46m ago•1 comments

Who's afraid of Chinese models?

https://stratechery.com/2026/whos-afraid-of-chinese-models/
421•mfiguiere•17h ago•288 comments

Running Doom on Our Custom CPU and Going Viral

https://www.armaangomes.com/blogs/doom/
11•arghunter•44m ago•0 comments

Five US tech giants' hidden debts soar to $1.65T on opaque AI funding

https://asia.nikkei.com/business/technology/five-us-tech-giants-hidden-debts-soar-to-1.65tn-on-op...
45•NordStreamYacht•42m ago•1 comments

Kimi Work

https://www.kimi.com/products/kimi-work
450•ms7892•11h ago•197 comments

Jelly UI: Soft-body physics for native HTML form controls

https://jelly-ui.com/
401•baldvinmar•11h ago•143 comments

Human mathematicians are being outcounterexampled

https://xenaproject.wordpress.com/2026/07/20/human-mathematicians-are-being-outcounterexampled/
272•artninja1988•9h ago•94 comments

Hacker wipes Romania's land registry database

https://news.risky.biz/risky-bulletin-hacker-wipes-romanias-entire-land-registry-database/
604•speckx•15h ago•337 comments

Jane Street: Incremental

https://github.com/janestreet/incremental
11•handfuloflight•47m ago•1 comments

Jellyfin founder Andrew leaves team

https://forum.jellyfin.org/t-project-leadership-changes
161•swat535•5h ago•102 comments

Flock Credibility Lost as It Repeatedly Lies to City Councils, Police, & Public

https://www.aclu.org/news/privacy-technology/tracking-alpr-cameras/flock-safety-credibility-lost-...
226•StatsAreFun•4h ago•31 comments

Nativ: Run frontier open models locally on your Mac

https://blaizzy.github.io/nativ/
221•aratahikaru5•10h ago•80 comments

I wrote an bash enumerator because I was sick of xargs

https://numerlab.org/2025/07/20/bashumerate-enumerator/
95•wallach-game•8h ago•67 comments

Show HN: Immersive Gaussian Splat tour of grace cathedral, San Francisco

https://vincentwoo.com/3d/grace_cathedral/
116•akanet•8h ago•24 comments

Is surveillance risk chilling your online speech?

39•Webstir•1h ago•28 comments

Agent swarms and the new model economics

https://cursor.com/blog/agent-swarm-model-economics
152•jlaneve•10h ago•65 comments

Launch HN: Bloomy (YC S26) – AI-powered mastery learning for K-12

73•alexsouthmayd•12h ago•76 comments

China’s open-weights AI strategy is winning

https://werd.io/american-ai-is-locked-down-and-proprietary-its-losing/
1018•benwerd•14h ago•813 comments

The Psychology of Software Teams

https://www.routledge.com/The-Psychology-of-Software-Teams/Hicks/p/book/9781032963389
64•dcre•5d ago•16 comments

LEDs’ potential to save our night skies

https://spectrum.ieee.org/led-light-pollution
223•defrost•15h ago•170 comments

A Mathematical Tribute to the Soccer Ball

https://www.nytimes.com/2026/07/17/science/mathematical-tribute-soccer-ball.html
7•igonvalue•3d ago•2 comments

My two year old taught me constraint solving

https://thecomputersciencebook.com/posts/how-my-2yo-taught-me-constraint-solving/
53•bambataa•1w ago•19 comments

The Power of Awareness: Overcoming Surveillance Capitalism

https://www.scottrlarson.com/presentations/overcoming-surveillance-capitalism-with-awareness/
71•trinsic2•8h ago•8 comments

Perfection is not over-engineering

https://var0.xyz/posts/perfection-is-not-over-engineering.html
215•var0xyz•14h ago•92 comments

Corners Don't Look Like That: Regarding Screenspace Ambient Occlusion (2012)

https://nothings.org/gamedev/ssao/
158•firephox•13h ago•65 comments

What most histories get wrong about MUMPS's first standard

https://github.com/rochus-keller/MUMPS/blob/main/docs/First_MUMPS_Standard_Article.md
8•Rochus•1w ago•8 comments

You only need the frontier model for one single edit

https://stencil.so/blog/prewalk
96•jxmorris12•5d ago•25 comments

Shinjuku Station in 3D

https://satoshi7190.github.io/Shinjuku-indoor-threejs-demo/
176•Gecko4072•14h ago•36 comments

How we measured AI writing across arXiv, and where the measurement breaks

https://unslop.run/blog/measuring-ai-writing-on-arxiv
201•dopamine_daddy•12h ago•145 comments

85.3 GFlops: Optimizing FP32 Matrix Multiplication on a Single AMD Zen 3 Core

https://github.com/houslast3/85.30-GFLOPS-Single-Core-FP32-Matrix-Multiplication-on-AMD-Zen-3
61•houslast•3d ago•17 comments