frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Show HN: Geiger – See every AI agent on your machine and what it can touch

https://github.com/Atomburstofficial/geiger
22•atomburst•1h ago

Comments

getatme32•46m ago
Seems like something that could also have enterprise applications for Shadow AI within an organization. Wouldn't be surprised if some of the observability/governance companies pick this up to use in their stack.
embedding-shape•45m ago
> One read-only command that inventories every AI agent, harness, MCP server, plugin, and AI extension on a machine

If this tool is returning even a single hit from this, you're probably using these agents wrong. You really want to run these in a way so they cannot touch your system drive/general filesystem that you use to do real work on. Even SOTA models at the end of their context limit behave REALLY illogical and does mistakes frequently. Don't run them straight on your machine unless you have backups and confirmed your backups work.

msdz•34m ago
This is the correct response, but it also feels as though ≈nobody is sandboxing their agents/harnesses in practice. Or at least, a good majority isn’t.
cortesoft•4m ago
I think there is a cycle, like with most things of this nature.

You start really locked down, and you read and approve every request for access. You do this for a while, but never see a result you deny, so you stop reading as closely. You keep hitting that approve button. You start paying even less attention to it. You start feeling silly, like you are simply slowing the process down. You get frustrated, because you keep coming back to your session and realizing your agent has been stuck waiting for approval for a long time, and the task that would have been done by now hasn’t even started.

Now you are running even more simultaneous sessions, which means more and more of your agents are stuck waiting for your approval. You feel even sillier, because you are taking even less time now to review and approve requests, but your review step is causing more and more slowdowns because you have more sessions going so you take longer between approvals. You feel like you are spending most of your time cycling through sessions hitting approve. You still have not come across a request that was dangerous or would have caused an issue, so you feel more and more like you are wasting your time.

So you slowly start to give your agents more access with fewer review steps.

Maybe this results in a catastrophic failure at some point, or maybe it doesn’t.

embedding-shape•3m ago
AFAIK, I know plenty of frontend engineers that without blinking run "npm install" on random 3rd party projects they find on GitHub, and they do their banking and everything else on the same machine. But then again, I also know people (one person to be fair) who have unprotected sex with prostitutes, so maybe something makes me slightly biased here.
Muromec•33m ago
I run it on a physically dedicated machine, it has a root there (through sudo nopassword) and I keep some of my stuff there too. I don't do it for security reasons, I just want it to run when my laptop is closed and I'm outside or what not.

Haven't seen any close calls so far. The thing just behaves. Nothing of the horror stories of eremerefing the whole home directory or a database. Am I just lucky?

embedding-shape•30m ago
I don't think you're just lucky, depends heavily on the model + reasoning effort. I've never had any of the GPT models do anything of the sorts when using the higher reasoning efforts, but sometimes when I play around with local models, even "big" ones (as big as you can fit with 96GB VRAM) sometimes forgets/misses to define $ID then do "rm -rf data/$ID" for example, deleting more than they intended.
Muromec•23m ago
>sometimes forgets/misses to define $ID then do "rm -rf data/$ID"

Right, that's kinda the same failure mode as delete from table something and hitting enter before you write the condition or writing the wrong one. If you reach the point where you opened the terminal to do it this way you already lost.

embedding-shape•20m ago
> If you reach the point where you opened the terminal to do it this way you already lost.

I'm not sure what this means, the model and agent harness is the ones "opening the terminal and running this" (via a exec_shell tool or whatever), they do mistakes like this sometimes. Sometimes the scope is bigger, sometimes less, but anything below SOTA + higher reasoning efforts seems to fall into these mistakes sometimes.

lukan•13m ago
"Even SOTA models at the end of their context limit behave REALLY illogical "

The trick is not go to that limit, but stay under 50% or even better 25% of context length. But backups are a smart thing anyway.

dpflan•35m ago
Wes McKinney has a project for agent visibility: https://www.agentsview.io/
Muromec•30m ago
I need to wire a pixel-display to show emoji faces of the agents and how they roll their eyes and all of that. I have the display, I have the prompt to pick a mood from emotional vocabulary, a custom harness and all of that. I just can't figure out how to make the eyes move on a statically pre-rendered emoji.

Show HN: Geiger – See every AI agent on your machine and what it can touch

https://github.com/Atomburstofficial/geiger
23•atomburst•1h ago•12 comments

Show HN: Bounty Index ,compare public bug bounty programs across 6 platforms

https://www.bountyindex.in
2•axolotl619•11m ago•0 comments

Show HN: Intro-level book about quantum computing theory/hardware/applications

https://quantumcomputingfortheconfused.com/
2•msolomentsev•11m ago•0 comments

Show HN: SockLight – A SOCKS5 dev proxy with a live terminal dashboard

https://github.com/kosmrljt/socklight
2•tomazko•18m ago•0 comments

Show HN: Whetuu – An opinionated, zero-config status line and history picker

https://github.com/yamafaktory/whetuu
3•yamafaktory•1h ago•0 comments

Show HN: Castforge – run Claude Code, Codex and Gemini as one dev team

https://castforge.ai/
3•jabenhaim•2h ago•0 comments

Show HN: Mnemiq – Open-source text-to-SQL tuned to your own database

https://github.com/agenticfabriq/mnemiq
2•paulinazhxu•2h ago•0 comments

Show HN: Skrivant – premise in, first-draft novel out, with continuity tracking

https://skrivant.com
4•meghneelgore•2h ago•0 comments

Show HN: Stop Scattering AI Resources Across Your Project: Meet Bear. Ctxpm

https://bear.best/en/blog/bear-ctxpm-context-package-manager/
3•BearBest•2h ago•0 comments

Show HN: Deploy unlimited small/medium-size apps with fixed price

https://pocketbasecloud.com/
3•tungnt620•2h ago•1 comments

Show HN: Copperhead – Cursor for circuit boards

https://copperhead.sh/
236•animeshchouhan•1d ago•107 comments

Show HN: Parlel – LinkedIn, but searchable by AI agents

https://parlel.com/
3•Dheerajiitr•2h ago•0 comments

Show HN: OpenPost, an open-source alternative to Buffer, CapCut, and Canva

https://github.com/getopenpost/openpost
3•kraktoos•2h ago•2 comments

Show HN: Type.com: Multiplayer Codex/Claude in the cloud for non-tech use cases

14•krashidov•2h ago•7 comments

Show HN: Persistent Jupyter kernel execution and live output streaming in VSCode

https://github.com/rnoro/tithon
2•rnoro_•2h ago•0 comments

Show HN: Watch Skill/DeepWatch agents can see screen/video

https://github.com/oxbshw/watch-skill
3•sayedev•2h ago•0 comments

Show HN: LLM Attention Visualization

https://ishamf.dev/p/llm-attention-visualizer/
160•ifz•23h ago•25 comments

Show HN: Maxxwell – The IDE for Optimal Tokenmaxxing

https://maxxwell.dev/
9•mserrano258•2h ago•4 comments

Show HN: Revliu – Multi-touch attribution from acquisition to revenue

https://www.revliu.com/demo
4•TychiqueY•3h ago•0 comments

Show HN: Wax – Record a program, change it then replay it

https://waxlang.dev/doom
3•weichx•3h ago•1 comments

Show HN: DortDB, a TS library for mixed-lang queries on JavaScript structures

https://github.com/filipjezek/dortdb
3•jeeza•3h ago•0 comments

Show HN: GuardRail, shell guards that stop Claude Code before it pushes to main

https://github.com/FvdHMBAI/guardrail
5•promptandbuild•3h ago•0 comments

Show HN: I rewrite my game while it is still running

https://github.com/AtomiJD/jdBasic
4•atomijd•3h ago•1 comments

Show HN: Hazzel – A tiny, minimal, Git-native coding agent

https://github.com/mukundzha/hazzel
9•mukundjha06•3h ago•1 comments

Show HN: NeXDM – Fast 64-segment download manager for Windows in Rust

https://nexdm.in/
3•Technews2026•5h ago•1 comments

Show HN: Numen - self hosted, it tells you what needs attention before you ask

https://github.com/theBstar/numen
2•Bikramsutar•5h ago•0 comments

Show HN: AgentPulse – Claude Code and Codex status in tmux

https://github.com/jerriclynsjohn/tmux-agent-pulse
4•jerriclynsjohn•6h ago•0 comments

Show HN: VolAnti – Open-source acoustic detector for fibre-optic FPV drones

https://github.com/agamrossen/VolAnti
13•agamrossen•15h ago•0 comments

Show HN: Impress your boss with interactive Scikit-Learn Decision Tree

https://github.com/mljar/supertree
5•pplonski86•6h ago•0 comments

Show HN: Compute Polynomials Twice as Fast

https://thomasahle.com/fast-polynomials/
6•thomasahle•7h ago•0 comments