frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Show HN: Relay – a self-hosted LLM gateway with smart routing and request pacing

https://github.com/anchorshell/relay
2•pavelmelnichuk•49m ago
Hi, I’m Pavel. I’ve been building Relay, a self-hostable AI gateway. The main thing I wanted to solve was smart routing without turning the gateway itself into a bottleneck. Relay can classify an incoming request, decide what kind of model capabilities the request needs, then route it across your configured providers while also accounting for limits and available capacity. It ships with routing classification models I’ve trained specifically for this problem, with the option of swapping that piece for Laya or other System One models.

This started when I was trying to string together several providers’ free tiers. I kept hitting rate limits at different intervals, which broke some of my agent clients. Some agents were also getting greedy with shared resources, so I needed a way to manage how they used the available capacity.

That led to Relay’s queue-first approach. Sometimes it’s better to wait a second or two for your preferred model than immediately fall back to another one. Relay queues and paces requests against configured provider limits, aiming to make use of available capacity without repeatedly hitting rate-limit errors or needlessly falling back to worse models.

Relay’s classification model also looks for signals that a request needs specific capabilities, such as coding or more complex reasoning, while classifying the request’s main intent before routing anything.

One thing I’ve obsessed over is keeping that decision layer cheap. Relay’s built-in classifier runs in the single digit millisecond range. It has a deliberately narrow job: classifying LLM requests and helping decide where to send them. It probably won’t be playing DOOM, but that’s a trade off I’m happy with for a routing layer. In the routing tests I've run so far, the built-in classifier is considerably faster than Laya while producing broadly similar routing decisions. Working on getting Jev up and running, and will report back to see how that stacks up as well.

The community version is available now, with a public repo, a built-in dashboard, and a local classifier. It’s written in Go, and you can run it with npx @anchorshell/relay or build it from source. The gateway itself is lightweight; it also ships with small classification models you can use locally.

There’s also a hosted version with a free tier if you don’t want to run it yourself. It offers our more capable classification models, along with team features, and separate limits for individual agents.

Vortex: One Format for Any Shape

https://spiraldb.com/blog/vortex-one-format-for-any-shape
1•surprisetalk•1m ago•0 comments

Prompt_jev(): Bringing Jev to Motherduck SQL

https://motherduck.com/blog/motherduck-supports-jev/
1•theanonymousone•2m ago•0 comments

Tons of Radioactive Waste Was Dumped in the Atlantic. Decades Later It's Leaking

https://www.sciencealert.com/150000-tons-of-radioactive-waste-were-dumped-in-the-atlantic-decades...
1•ck2•2m ago•0 comments

Show HN: Blink – A high-performance Jev like decision model for C and WASM

1•marcobambini•5m ago•0 comments

Claudisms AI models overuse the same words. This is a running catalog of them

https://claudisms.ai/
1•Bluestein•6m ago•0 comments

'Turn the cameras off:' London's growing privacy pushback against smart glasses

https://www.cnn.com/2026/09/22/tech/london-privacy-pushback-meta-smart-glasses
2•thunderbong•6m ago•0 comments

The Personal Internet

https://yenecys.bearblog.dev/the-personal-internet/
1•Emerald_dreamer•7m ago•1 comments

Runtime Alignment with Synchronous Control Monitoring

https://max.ax/writing/synchronous-control-monitoring/
1•k5hp•8m ago•0 comments

Values, conditions, intent: People should declare them, not the model

https://github.com/Jang-woo-AnnaSoft/execution-state-preflight/blob/main/spec.en.md
1•offaxis•9m ago•0 comments

Show HN: Epure – Self-hosted exception tracker in 2 containers (Rust, Postgres)

https://github.com/epure-sh/epure
1•nonmaskable•10m ago•0 comments

How to Create a Crypto Token in 2026: Complete Step-by-Step Guide

https://www.promoteproject.com/article/228685/how-to-create-a-crypto-token-in-2026-complete-step-...
1•kodegridmarketi•15m ago•0 comments

I Maxed Out Fable 5 and Regretted It

https://thoughts.jock.pl/p/fable-5-efficiency-high-not-max-2026
2•joozio•15m ago•0 comments

Up to 100k in Taiwan spying for China: expert

https://www.taipeitimes.com/News/front/archives/2026/09/16/2003864338
2•harry_nutsachs•15m ago•0 comments

Show HN: ReacherX – Open-source Apollo alternative for founders and devs

https://www.reacherx.com/home
1•noobships•17m ago•0 comments

GPU is starving – LLM host dispatch at 191k steps/s on 1 vCPU

https://github.com/cortexLab011/floria-serving
1•cortexlab1•17m ago•0 comments

Apple Music Hall

https://www.apple.com/newsroom/2026/09/apple-opens-apple-music-hall-a-state-of-the-art-live-music...
3•soheilpro•17m ago•1 comments

Why the hard part of route optimization is the modeling, not the algorithm

https://kardinal.ai/why-the-hard-part-of-route-optimization-is-the-modeling-not-the-algorithm/
1•cedrichervet•18m ago•0 comments

I'm sick of Claudisms, and on AI-based software development

https://www.polso.info/im-sick-of-claudisms-future-ai-software-development
2•Risse•19m ago•0 comments

In Trump, Xi Sees 'Rare Window' to Convince Taiwan That Resistance Is Futile

https://www.wsj.com/world/china/trump-xi-taiwan-f236f228
1•Betelbuddy•20m ago•0 comments

JevBench, a reproducible benchmark for typed decision models

https://benchmarkheaven.com/jev-models
1•florianstandhar•20m ago•1 comments

British Columbia Sues OpenAI, Alleging ChatGPT Aided Mass School Shooting

https://www.wsj.com/tech/ai/british-columbia-sues-openai-alleging-chatgpt-aided-mass-school-shoot...
2•fortran77•20m ago•1 comments

Training a model to identify AI-generated web content from structure alone

https://arxiv.org/abs/2609.15369
3•jochenmadler•21m ago•1 comments

Show HN: ScopeTrail – audit receipts for multi-hop agent delegation

https://github.com/scopetrail/scopetrail
1•mrjimmy1•21m ago•0 comments

Coulomb's law remains tricky to test at home

https://chillphysicsenjoyer.substack.com/p/coulombs-law-remains-tricky-to-test
1•surprisetalk•21m ago•0 comments

Show HN: Gemma 3 4B as a typed decision function in Rust (47 ms/decision)

https://github.com/zozo123/gemma-to-jev
1•zozo123-IB-IL2•22m ago•0 comments

The American Prison Experiment: A History of Good Intentions and Bad Results

https://reason.com/2026/09/08/the-american-prison-experiment/
1•mooreds•23m ago•0 comments

Create digital product stores in minutes and start earning your first dollars

https://cloud.freshlimepay.com/?lang=en
2•jimmy_lee•23m ago•1 comments

PEP 823 – None-aware access operators

https://peps.python.org/pep-0823/
1•Ravencentric•23m ago•0 comments

ForgeMT: Multi-tenant platform for self-hosted GitHub Actions runners on AWS

https://github.com/cisco-open/forge
1•handfuloflight•24m ago•0 comments

Unfinished Work in Package Security

https://nesbitt.io/2026/09/22/unfinished-work-in-package-security.html
1•lumpa•24m ago•0 comments