frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Show HN: ZeroShot: Agent session monitoring to make your team go faster

https://tryzeroshot.com
11•chaos_emergent•59m ago

Comments

kplawver•42m ago
I've been using it for three weeks and really like it - it's already made some great additions to our app's skills and improved my coding agents' workflow!
chaos_emergent•15m ago
thanks @kplawver! As you know it's early days and we're roughing things out but really excited that you're getting value from it :)
ChronoGawd•1m ago
Posting on behalf of Nikhil here, CTO @ BuildBetter, YC founder (W19). We're building ZeroShot, an agent session monitoring tool that builds and shares self-improving skills with your team so you can work with AI agents the way your most productive engineers do.

Download the Mac Desktop App at tryzeroshot.com/download or run curl -fsSL tryzeroshot.com | sh to give it a spin!

What ZeroShot does: it captures your coding-agent sessions locally (Claude, Codex, Cursor, Pi, OMP, more on request), then runs a classification pipeline over them in the background looking for friction such as retry loops, course corrections, and missing context. When the same friction shows up across multiple sessions or people, it clusters those and drafts a reusable skill. What's unique?

- Cross-session skill drafting. Agents can't see that they've walked the same failure trajectory five times across five sessions. Even agent-driven retrospectives miss what we catch because the same underlying friction looks different in each transcript, so anything analyzing raw transcripts comes up empty. Instead, we cluster at the pattern level.

- Team-level enrichment (paid tier). We look across your team's sessions, GitHub PR comments and contribution history, and weight skills toward the people who are experts in a given codebase or area. We're also starting to surface why some teammates get better mileage from their LLM use and share that across the team. This part is still in progress and I'd love feedback on which productivity metrics are most useful.

- The system maintains itself so you don't have to. Hand-written .md and rules files decay quickly. ZeroShot validates and regenerates skills as your sessions evolve, so the maintenance burden sits with us, not your team.

Knowledge sharing across the org. Skills live in ZeroShot-managed skillsets shared across the team, and the same accumulated skills follow you whether agents run locally or remote.

The problem we kept hitting: Between September and November of last year, I found myself getting more upset with the vibe-slop that made its way into main from engineers who knew better (and some who didn't). So we started harness engineering. At first we just wanted to make sure that the models wouldn't make the same dumb mistakes that we kept on repeating in PR comments and to our own agents. We realized that specific engineers were able to "hold" the models better and were more productive because of it. ZeroShot is born from that internal experimentation that led us to being 5x more productive as a team.

We were inspired to build the app by our own internal self-improving-skill system, which itself was inspired by systems that the Codex team has espoused on X.

Privacy: the app runs entirely on your machine for free. Session data doesn't leave your repo unless you explicitly opt into the team features. We use your local agents to parse through clusters.

Limitations:

- Skill extraction works best when there's repeated friction across multiple user sessions. If you're a solo dev doing highly varied work, the clustering has less signal to work with.

- The efficiency numbers on our site are from teams with heavy repeat work; your mileage will genuinely vary and I don't want to oversell it.

- Works with Claude Code, Codex, Cursor, Pi (including OMP) today, and we're adding support for more environments weekly. It's early days and we're constantly working on our pipelines; please give us feedback as you start using it!

Mac only for now, Windows in the works, Linux CLI.

Pricing: the local app is free, no account needed. Paid plans are for team features (shared skills, observability, cost attribution).

Happy to answer anything, especially skeptical questions about the token-savings claims. I'd rather explain the methodology than have you dismiss them outright. We're actively looking for design partners, so please hmu if you want a demo or run into any issues!

Stop Detecting Deepfakes: Make the Fraud Irrelevant by Design

https://smarterarticles.co.uk/stop-detecting-deepfakes-make-the-fraud-irrelevant-by-design
1•dxs•2m ago•0 comments

Show HN: I bet your hardest customer orders can be automated. Challenge Zentriq

https://www.zentriqai.com/en
1•Enes_Erdem•2m ago•0 comments

Show HN: NegativeEV – I built a tool to help people see how bad their bets are

https://negativeev.com/
2•qkwrv•2m ago•1 comments

Don't use (pancakes emoji) as menu icon

https://github.com/orgs/community/discussions/203497
2•gavide•2m ago•0 comments

Stevey's Google Platforms Rant (2011)

https://gist.github.com/chitchcock/1281611
2•asplake•2m ago•0 comments

Mark Zuckerberg is becoming the king of the 'side quest'

https://www.ft.com/content/6d206b85-bc50-4c89-973b-f18cd84e2a15
2•JumpCrisscross•3m ago•0 comments

My Email Hygiene

https://matanabudy.com/my-email-hygiene/
2•speckx•3m ago•0 comments

OpenAI revenue in July topped all of Q2 driven by GPT-5.6 release

https://www.cnbc.com/2026/07/29/openai-cfo-sarah-friar-tells-employees-arr-in-july-topped-all-of-...
3•giuliomagnifico•5m ago•0 comments

UEFA and its national associations will not participate in FIFA competitions

https://www.uefa.com/news-media/news/02a7-213a92896eb0-54dfbf454e3b-1000--statement-on-behalf-of-...
3•dickfickling•6m ago•0 comments

Postgres Queues Actually Scale

https://www.dbos.dev/blog/making-postgres-queues-scale
5•KraftyOne•8m ago•0 comments

I'm a senior estimator, I open sourced my construction takeoff tool

https://github.com/Kentucky-ai/opentakeoff
2•kentucky-ai•8m ago•0 comments

Great Teachers Are Underappreciated

https://twitter.com/Alfred_Lin/status/2082102711030546783
3•gmays•9m ago•0 comments

Why I Left Google (2015)

https://medium.com/@docjamesw/why-i-left-google-c170e6165f2a
2•sssilver•11m ago•0 comments

What Would It Mean to See a New Color?

https://www.newyorker.com/magazine/2026/08/03/what-would-it-mean-to-see-a-new-color
3•fortran77•11m ago•0 comments

Show HN: Leia – a CLI to oneshot playlists onto Yoto MYO cards

https://github.com/nsokin/leia
2•nsokin•11m ago•0 comments

Show HN: APIgen – Generate Fusio REST APIs and Angular UIs from TypeSchema

https://apigen.app/
2•k42b3•12m ago•0 comments

Floundering A.I. Hedge Fund Is Rescued by Rival

https://www.nytimes.com/2026/07/30/business/artificial-intelligence-situational-awareness-citadel...
2•mapping365•13m ago•0 comments

Turning GitHub Actions into an oracle (2025)

https://www.ethanheilman.com/x/35/index.html
2•EthanHeilman•13m ago•0 comments

Singles Are So Burned Out They're Letting Spreadsheets Evaluate Their Dates

https://www.wsj.com/lifestyle/relationships/dating-singles-love-hackers-spreadsheets-711a493b
2•gmays•14m ago•0 comments

IronClaw 1.0 – one secure agent on CLI, Slack and Telegram, one memory

https://near.ai/blog/introducing-ironclaw-1-0
2•dannycb•15m ago•0 comments

The Zeno's Paradox of AI

https://negativezone.github.io/2026/07/29/zeno.html
2•negativezone•15m ago•0 comments

The Impending, Inescapable Deluge of A.I

https://www.nytimes.com/interactive/2026/07/29/technology/ai-chips-data-center-boom.html
2•KasianFranks•15m ago•1 comments

So you want to use plants to reduce CO₂

https://dynomight.net/plants/
2•surprisetalk•16m ago•0 comments

Show HN: A single character turns primitive recursion into general recursion

https://github.com/raoofha/pr
2•raoof•16m ago•0 comments

Programming as Art

https://noumenalnotions.space/essays/programming_as_art/
3•speckx•17m ago•0 comments

I made a term-a11y: accessible spinners and progress bars for CLI tools

https://github.com/zaydea805/term-a11y
2•zay_dea•19m ago•1 comments

LLM re-reads the same text a thousand times. Here's what I measured

https://swellweb.github.io/reame/bytes/
3•targetbridge•19m ago•0 comments

AI productivity gains are closer to 10% than 10x

https://leaddev.com/reporting/ai-productivity-gains-are-closer-to-10-than-10x
4•champagnepapi•19m ago•1 comments

Thinking Machines: Inkling Small

https://twitter.com/ArtificialAnlys/status/2082894822180819057
4•tosh•20m ago•0 comments

Show HN: State of the Feed – TikTok usage analyzed

https://wrapped.vantezzen.io/state-of-the-feed
1•bennett_dev•20m ago•0 comments