frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

RTK reports token savings, but our cost benchmarks disagree

https://quesma.com/blog/does-rtk-make-ai-coding-cheaper/
27•michalwarda•1h ago

Comments

vrighter•49m ago
well yeah.... now you're giving it output it wasn't trained on.
nextaccountic•5m ago
What if the next-gen models are trained on RTK output as well? Then you will actually have less tokens in the context window, and the model won't become confused (which would require more turns, wasting tokens)
aeneas_ory•48m ago
All of these "hacks" are snakeoil and I think deep down we all know. Whether it's caveman, RTK, or whatever other vibe-coded productivity/token cost saving hacks/skills/claude.md.

What I had success with (although benchmarks are older) is to index the codebase with a dedicated local code embedding model. It's a bit expensive on the CPU side but in my benchmarks it reduced token use and wall clock time significantly. Of course, it's always dependent on statistical noise + host system load, and running sufficiently large benchmarks is simply too expensive, so take em with a grain of salt.

Why does it work you may ask? Well, LLMs basically brute force words/phrases and pipe that into find/grep/pgrep/whatever (or as recently discussed here write a python script for it - https://news.ycombinator.com/item?id=49654229). Semantic search looks for similarities so you have to do less brute forcing. Comes of course at the cost of indexing everything first.

You can find the project here: https://github.com/ory/lumen

sreekanth850•44m ago
i don't know if such hacks works, but in C# if you use roslyn mcp, you save a lot.
VulgarExigency•15m ago
I don't think they're comparable. RTK just modifies the output of CLI tools to reduce the number of tokens, a Roslyn MCP gives the agent a fundamentally superior way of interacting with a C# codebase.
sreekanth850•13m ago
Yes. and i find model makes less errors and reasoning the codebase well, especially when you do a large refactor.
antupis•5m ago
My main issue with rtk is that rtk somehow messes modification and agent start polling same tool continuously.
fwlr•37m ago
This makes sense. “Don’t try to penny-pinch your employees” is a lesson most managers learn eventually, and I guess agent-orchestrators will have to learn it too.
gillesjacobs•22m ago
Main takeaway:

``` Average cost per attempt, without → with RTK:

Claude/Fable: $1.72 → $1.64 (~5% cheaper) DeepSeek: $0.115 → $0.121 (~5% more expensive)

Almost all Claude savings came from a single task. Excluding it, savings were under 1%. ```

It took me a few rereads to parse out the top-line. This article really buries the lede.

fleetfox•6m ago
I don't understand how this or all these magic skill bundles and methodologies get traction and why they are so popular. It's either plain worse or has serious trade offs.

Show HN: Number Checker for WA – Check and Verify WhatsApp Numbers

https://chromewebstore.google.com/detail/number-checker-for-wa-che/iiblpglkglfmefhlcdlcibphegpjfgpo
1•qwikhost•1m ago•0 comments

Evolution of Attention over the Years

https://gargighosh1.substack.com/p/evolution-of-attention-over-the-years
1•sonabinu•2m ago•0 comments

Homebrew now has an official GUI app

https://github.com/Homebrew/BrewUI/
2•kozika•3m ago•0 comments

What is JSONC and is it any better than JSON?

https://www.howtogeek.com/what-is-jsonc-is-it-any-better-than-json/
1•theanonymousone•6m ago•0 comments

Airflow re-engineered for speed and scale

https://www.astronomer.io/blog/astro-airflow-re-engineered-for-speed-and-scale/
1•jlaneve•8m ago•0 comments

Conversations with JJ

https://laurmaedje.github.io/posts/jj/
1•haeseong•8m ago•0 comments

SSH for Developers

https://flaviocopes.com/ssh-for-developers/
1•ibobev•9m ago•0 comments

Stacked solar panels are pushing beyond the limits of standard photovoltaics [video]

https://www.youtube.com/watch?v=y5Pt4nSZMnA
1•thelastgallon•10m ago•0 comments

A Deep Dive into Tmux

https://flaviocopes.com/tmux/
1•ibobev•10m ago•0 comments

A Deep Dive into Ghostty

https://flaviocopes.com/ghostty/
1•ibobev•10m ago•0 comments

Show HN: NanoVector – A 120KB zero-dependency vector search engine in C and SIMD

https://github.com/eminsk/nanovector
1•eminskinfo•12m ago•0 comments

Hanoi Hannah

https://en.wikipedia.org/wiki/Hanoi_Hannah
1•thunderbong•13m ago•0 comments

We're temporarily pausing new sign-ups and upgrades to the ChatGPT Pro $200 plan

https://help.openai.com/en/articles/9793128-about-chatgpt-pro-tiers
1•embedding-shape•14m ago•0 comments

Buying GameCube games from other regions

https://sethmlarson.dev/buying-gamecube-games-from-another-region
1•surprisetalk•16m ago•0 comments

Culture slop: how AI gave Brazil an American 1980s

https://albertoescarlate.substack.com/p/culture-slop-how-ai-gave-brazil-an
1•Anon84•16m ago•0 comments

Show HN: Foldelight – the iPhone Duo folding effect the MacBook was owed

https://lufzle.dev/foldelight/
3•riffonio•17m ago•0 comments

Europe's difficult choices on AI – Mario Draghi

https://www.ft.com/content/f054f927-b512-452a-b494-ea53f5ac1079
1•nkjoep•18m ago•0 comments

Could United Launch Alliance's money problems force its owners to sell?

https://arstechnica.com/space/2026/09/could-united-launch-alliances-money-problems-finally-force-...
2•rbanffy•20m ago•0 comments

Twelve search APIs ran through the same AI agent to see which one helps it most

https://artificialanalysis.ai/agents/search-api
3•2001asjad•20m ago•0 comments

The Oldest Architecture in Computing

https://www.allthingsdistributed.com/2026/09/the-oldest-architecture-in-computing.html
1•jreynar•21m ago•0 comments

Scheduling is NP-Hard. We need responses in milliseconds

https://cal.com/blog/scheduling-is-np-hard-we-need-responses-in-milliseconds
2•mirzap•26m ago•0 comments

Seeing high-dimensional data in two dimensions

https://stochastic.blog/seeing-high-dimensional-data-in-two-dimensions/
2•Anon84•27m ago•1 comments

Make Claude Code Faster and Cheaper with Ory Lumen

https://www.ory.com/blog/ory-lumen-semantic-search-claude-code
3•Bluestein•27m ago•0 comments

Parody on tech non-profits: The Association Organized

https://www.inventati.org/diaphragmwp/theass/
2•dmwpnski•29m ago•1 comments

Show HN: Light Converter – background removal that runs in the browser

https://github.com/lightScout/light-converter
1•lightScout•29m ago•0 comments

We need to build more data centers

https://hunchfox.substack.com/p/the-cost-of-abundance-ai-2030
1•sahil_chaudhary•30m ago•1 comments

Code Contracts: Structured specifications colocated with code

https://code-contracts.cc/
1•spolu•31m ago•0 comments

PythonOS – A tiny OS that boots into Python

https://github.com/jordanhubbard/pythonos
2•indigodaddy•31m ago•0 comments

Show HN: Hopya – Open-Source Project Management Platform

https://github.com/wnzn/hopya
1•wenzani•33m ago•0 comments

Greenland

https://nk412.com/photoblog/greenland.html
1•grilledchickenw•33m ago•0 comments