frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Show HN: Zeno – A framework for verifiable RL rewards (code, math, and more)

https://github.com/Think-a-Tron/zeno
2•Sai_Praneeth•1y ago
With TRL, it's now straightforward to RL-finetune LLMs, but picking good reward functions is still the weakest link.

Zeno is an open-source toolkit for verifiable, deterministic reward functions for RL on LLMs.

While the initial release focuses on Python code generation, the goal is broader: make RL reward design for LLMs transparent, modular, and extendable across domains (math, retrieval, reasoning, tool-use, etc.)

What's in Zeno for now? - Auditable, stateless reward functions for Python code - docstrings, ruff linting, type hints, recursion, and more - Works directly with Huggingface's TRL or any RL loop - plug reward functions in as needed. - MIT licensed and minimal.

Roadmap: Python code is just the starting point. Extensions for math problem solving, planning and agentic behaviors are in todo.

Repo: https://github.com/think-a-tron/zeno

Docs and more details in the README

Comments, critiques, and real-world use cases encouraged, especially if you want to push beyond code.

Grace's Training

https://grace.training/
1•owl_vision•4m ago•0 comments

Show HN: Television – an open source GUI for your agent harness

https://television.run/
2•stlhood•6m ago•0 comments

My Cup of Coffee versus Your Subscription Cost

https://thisismylast.bearblog.dev/my-cup-of-coffee-versus-your-subscription-cost/
1•hackerbeat•6m ago•0 comments

PTXBench: Benchmarking and Adapting LLMs for GPU Kernel Optimization

https://arxiv.org/abs/2608.17379
2•matt_d•10m ago•0 comments

Building Your Own Product Factory to Your Taste

https://codebastard.dev/from-context-engineering-to-harness-engineering-building-your-own-product...
1•codebastard•10m ago•0 comments

Shut up and take our money: Amazon plows $1B into quashing datacenter dissent

https://www.theregister.com/systems/2026/10/02/shut-up-and-take-our-money-amazon-plows-1b-into-qu...
1•joebuckwilliams•10m ago•0 comments

<input type="password" maxlength="20"> prevents me from logging into Vanguard

https://tanin.nanakorn.com/input-type-password-maxlength-20-is-considered-harmful-and-why-i-could...
4•tanin•12m ago•0 comments

Humanoid robots destroy themselves after being decommissioned

https://www.the-independent.com/tech/robot-suicide-humanoid-death-figure-b3060486.html
1•ColinWright•13m ago•0 comments

AI Cracks 217-Year-Old Napoleonic Cipher in 6 Hours,Revealing PreWar Deployments

https://finance.biggo.com/news/0e4a6995-1f3d-497b-9291-ad8cfdbb7b58
1•initramfs•13m ago•0 comments

Show HN: Reddix – an open-source Reddit / X / Partiful

https://reddix.co/
1•codyzazulak1•16m ago•0 comments

Show HN: Kosh – Give your bookmarks a shape worth keeping

https://chromewebstore.google.com/detail/kosh-bookmark-manager/meldjdfnfgefeoegccphimkmgmeamcce
1•havu12•17m ago•0 comments

You cannot align an organization one agent at a time

https://www.siliconcontinent.com/p/openai-thought-it-was-testing-agents
1•abdelhousni•22m ago•0 comments

Show HN: Interview Basecamp – prep from a role or a job description

https://www.interviewbasecamp.com/
1•balkrishnajha•23m ago•0 comments

Show HN: Teacher Planner – A digital planbook with AI tools for teachers

https://teacherplanner.ai
1•sunpy•23m ago•0 comments

Access open‑weight, cyber‑capable models with one API key

https://router.enclave.ai/login?next=%2F
1•talhof8•23m ago•0 comments

AI Torture Chamber – Steering LLMs into negative and positive valence states

https://github.com/login
1•Anonboxis•26m ago•0 comments

Clarify: Reach Decisions 2-4x Faster than Grill-me

https://github.com/a-laughlin/agentic-resources/tree/main/skills/clarify
1•a-laughlin•26m ago•1 comments

Show HN: LinkShip – Turn any file into a trackable, shareable link

https://linkship.net
1•sunpy•27m ago•0 comments

The Heilbronn Problem

https://math.tejstead.com/heilbronn/
2•tejstead•30m ago•1 comments

Catscan: Visualizing Pipelines of CPU Performance Simulation

https://arxiv.org/abs/2610.02121
1•matt_d•30m ago•0 comments

LLMs Will Not Replace AI Compilers. They Will Call Them.

https://aicompilers.github.io/2026/09/27/llms-will-not-replace-ai-compilers-they-will-call-them.html
1•matt_d•34m ago•0 comments

Show HN: I found API keys in my coding agent history so I built Agent Scrub

https://github.com/thesubtlety/agent-scrub
1•thesubtlety•38m ago•0 comments

Robot dog taught itself to walk with a spiking neural network – no reward signal

https://www.reddit.com/r/Petoi/comments/1u5vgey/my_bittle_taught_itself_to_walk_with_a_spiking/
2•petoi•40m ago•0 comments

Show HN: Gutsy – small decision model for CPU (500MB), get probabilities back

https://huggingface.co/kouhxp/gutsy
1•mrkn1•44m ago•0 comments

Debian Inference Portal

https://inference.debian.net/
2•ColinWright•48m ago•1 comments

Semantic Compute: From Interpreters to Compilers

https://seldon-ai.com/blog/fronter-llms-are-semantic-interpreters
1•nlpnerd•49m ago•0 comments

1 in 4 new California homes is in someone's backyard

https://maxmautner.com/2026/10/02/california-adus.html
2•mslate•51m ago•0 comments

Homo Promptus

https://www.kevinsdias.com/posts/homo-promptus.html
3•marjancek•55m ago•0 comments

Elohim

https://boozelee.github.io/elohim-web/
2•kilibear•56m ago•0 comments

Garmin watch as a controller for the Flipper Zero

https://github.com/rz-x/flipper-on-garmin
1•rdn86•57m ago•0 comments