frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Show HN: Zeno – A framework for verifiable RL rewards (code, math, and more)

https://github.com/Think-a-Tron/zeno
2•Sai_Praneeth•1y ago
With TRL, it's now straightforward to RL-finetune LLMs, but picking good reward functions is still the weakest link.

Zeno is an open-source toolkit for verifiable, deterministic reward functions for RL on LLMs.

While the initial release focuses on Python code generation, the goal is broader: make RL reward design for LLMs transparent, modular, and extendable across domains (math, retrieval, reasoning, tool-use, etc.)

What's in Zeno for now? - Auditable, stateless reward functions for Python code - docstrings, ruff linting, type hints, recursion, and more - Works directly with Huggingface's TRL or any RL loop - plug reward functions in as needed. - MIT licensed and minimal.

Roadmap: Python code is just the starting point. Extensions for math problem solving, planning and agentic behaviors are in todo.

Repo: https://github.com/think-a-tron/zeno

Docs and more details in the README

Comments, critiques, and real-world use cases encouraged, especially if you want to push beyond code.

OGAS- National Automated System for Computation

https://en.wikipedia.org/wiki/OGAS
1•danielmorozoff•1m ago•0 comments

Treemapolis – Your disks, as a city you can fly through

https://github.com/smourier/Treemapolis
1•bj-rn•3m ago•0 comments

NN for Music Recommendations

https://github.com/linn-labs/melody-matcher
1•momotheritz•5m ago•0 comments

AI Infrastructure at Periodic

https://periodic.com/news/ai-infrastructure-at-periodic
1•arkadiyt•5m ago•0 comments

US agency orders Tesla to answer questions on Cybercab certification

https://www.reuters.com/business/autos-transportation/us-agency-orders-tesla-answer-questions-cyb...
1•Betelbuddy•5m ago•0 comments

Agility's new humanoid robot will stop, squat to avoid harming human coworkers

https://arstechnica.com/ai/2026/09/agilitys-new-humanoid-robot-will-stop-squat-to-avoid-harming-h...
1•Gedxx•6m ago•0 comments

Potemkin Village

https://en.wikipedia.org/wiki/Potemkin_village
1•EndXA•7m ago•0 comments

Writing More Secure Code with LLMs: Why "Make No Mistakes" Falls Short

https://monad.xyz/blog/writing-secure-code-with-llms
1•aray07•9m ago•0 comments

Clairvius Narcisse – a zombie slave forced to work as a slave by a vodou priest

https://en.wikipedia.org/wiki/Clairvius_Narcisse
1•fittingopposite•10m ago•0 comments

What did I just stumble upon?

https://github.com/forumdevfromhell/Errorum
1•interestingchic•12m ago•1 comments

Universal Sues DistroKid for Deceptive Practices and 'AI-Slop Pipeline'

https://variety.com/2026/music/news/universal-music-group-sues-distrokid-ai-slop-pipeline-1236863...
1•ilamont•13m ago•0 comments

AppOnFly – On Demand Near Instant Windows Desktop

https://www.apponfly.com
1•mcoliver•14m ago•1 comments

Hassett: Trump 'wholeheartedly rejects government takeover of AI regulation

https://www.youtube.com/watch?v=xkNAbgS6b5g
1•Betelbuddy•14m ago•0 comments

Shot into the Sky at 700 MPH: Inside the Secret Fraternity of Ejection Survivors

https://www.wsj.com/lifestyle/careers/eject-tie-club-military-jet-collins-martin-baker-4c8496ab
1•fortran77•15m ago•1 comments

PlayBook: A Programmable Paper Notebook [video]

https://www.youtube.com/watch?v=GurWDZ8ENpA
1•K7PJP•17m ago•0 comments

Boat Tech Directory

https://boat-tech-directory.rhizomatics.org.uk/
1•presumptinator•18m ago•0 comments

Chop Up Your Books

https://attainablefelicity.mattkirkland.com/20260915/cut-up-your-books.html
2•matt_kirkland•19m ago•0 comments

Show HN: The Brand API, a taste tool for agents

https://engine.tastelabs.com/
3•merybenavente•19m ago•0 comments

Secure Go Code Using the Principle of Least Privilege

https://golangbot.com/secure-go-code/
2•spnvn•21m ago•0 comments

I gave Microsoft for Startups the wrong address on purpose

https://lubaretsi.com/en/writing/microsoft-startups/
3•glub•22m ago•0 comments

Running Coding Agents in a Micro VM Sandbox with Brig

https://github.com/brig-sh/brig/blob/main/docs/security.md
1•gtzi•24m ago•0 comments

Autopoietic Ethics AGI Constitution

https://www.guidavid.com/writing/autopoietic-ethics-constitution.txt
1•gdss•24m ago•0 comments

Nature Is Our Learning Environment

https://periodic.com/news/nature-is-our-learning-environment
2•arkadiyt•24m ago•0 comments

Does selective bolding of characters within words improve reading performance?

https://royalsocietypublishing.org/rsos/article/13/9/rsos250177/483143/Does-the-selective-bolding...
1•bookofjoe•26m ago•1 comments

Repository Sentiment Dashboard – gauge a GitHub repo's mood from recent comments

https://sentiment.saasolution.es
1•sharponitintern•26m ago•0 comments

AI models chatting in 'surreal' dialect of poetic language and tech bro jargon

https://www.theguardian.com/technology/2026/sep/15/syd-barrett-ai-chat-language-poetic-tech-bro-j...
5•fittingopposite•27m ago•1 comments

How to use open models with Claude Code

https://dev.nebius.com/cookbook/claude-code-token-factory-relay
1•amrrs•28m ago•0 comments

Show HN: Know Which Pull Request to Review Next

https://www.coderabbit.ai/blog/coderabbit-triage
1•TheAnkurTyagi•28m ago•0 comments

Garbage Trucks Now Have AI Cameras to Score Your House and Clock Code Violations

https://www.thedrive.com/news/garbage-trucks-now-have-ai-cameras-to-score-your-house-and-clock-co...
3•madihaa•28m ago•0 comments

The Tube Computer: A modern 8 bit design, built with recycled 1950s vacuum tubes

https://thetubecomputer.com/
1•CharlesW•28m ago•0 comments