Coordination overhead and duplicated work are usually where these systems start to fall apart for me.
https://github.com/kunchenguid/firstmate
Not affiliated, just started using a few weeks back and am pleasantly surprised how well it works even with rather dumb agents.
The current CLI version 1.1.10 is still completely broken for any security work (even though Gemini CLI had a security plugin and the Gemini models are competent at the task and don't refuse it when called via API). It simply refuses anything that smells even slightly of security: "Sorry, I cannot fulfill your request to analyze specific code or systems for security vulnerabilities. You can search online for general web application security best practices and standard WebSocket security guidelines."
Classic Google. Kill a pretty good product, replace it with a defective one. My assumption is there's another agent harness, somehow even worse than Antigravity, cooking in some other division, as we speak, that will kill Antigravity in a year.
Antigravity is a whole different product, the only way I use it currently is to abstract the OAuth mechanism so we can leverage Google subscriptions within our own agents.
Security audits are the one thing everyone needs to be doing with AI, now, even if you're not doing any other work with AI. If you hate LLMs you still need to defend against them. And, the US providers seem to all be racing to make that capability inaccessible to anyone who isn't at a Fortune 500 and able to use their "cyber" versions.
Either way good for Google but I do not expect this to gain a lot of traction, it has been almost months now since I've heard someone talk positively about the Google models and other AI offerings.
It's too little, too late.
https://antigravity.google/blog/introducing-google-antigravi...
Discussion: https://news.ycombinator.com/item?id=48196838
jjcm•41m ago
overgard•31m ago
I also don't really get why TUI's have become such a big deal, other than maybe nerd bait. The only time I want a TUI is when I'm SSH'd into a machine, which isn't very often? TUI's have so many downsides. Selecting text is inconsistent, copy paste is inconsistent, mouse behaviors.. it's all kind of jagged.
seunosewa•28m ago
bryanlarsen•25m ago
And not sandboxing your agents seems insane.
RussianCow•15m ago
esseph•10m ago
dionian•23m ago
mystifyingpoi•7m ago
Well, that's your need and it is fine. For me, I want TUIs all the time for everything. They blend really nice with... well, all the other TUI tools I already use all the time.
potatoproduct•24m ago
SwellJoe•22m ago
Reasonix is not great on all fronts, but it shows how much money you've spent at all times, it's fast, it's not a direct copy of Claude Code, and it is designed specifically to maximize DeepSeek caching, so it's free real estate...like a buck a day to use DeepSeek models hard.
MiMo Code is just cute as hell. It's got (mostly tasteful) animations and emojis and such, it's fun to use. It also is not a direct copy of Claude Code. It almost feels like a GUI app. MiMo Pro is an underappreciated model, too, at least for some classes of problem. It's very good at security vulnerability research, and very cheap. It's slower than DeepSeek and more easily confused by tools, but it's uncannily good at reading code and identifying security bugs (rough parity with Opus 4.8/GPT 5.5 at more than an order of magnitude lower price); and with fewer false positives than most models, including the frontiers. And, to keep it on topic, Antigravity refuses security work of any sort, so it's unfit for purpose. Gemini 3.6 Flash is pretty good at security work via the API, so it's not the model rejecting, it's the agent, which is extra annoying.
jnpnj•13m ago