frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Gemini 4 Argon (High): Intelligence, Performance and Price Analysis

https://artificialanalysis.ai/models/gemini-4-argon
42•theanonymousone•1h ago

Comments

anuragdaram•57m ago
Will have to check what the pricing would be for this model.
krat0sprakhar•40m ago
> Argon will launch at an introductory price of $2 per million input tokens and $10 per million output tokens, with cached input tokens priced at 95% off input token price.

https://blog.google/innovation-and-ai/models-and-research/ge...

1/5th the price of Astra and Fable

A_D_E_P_T•42m ago
Trading punches in the benchmarks with Mimo v2.6 and 6.1-Sol (both very cheap!), and decidedly inferior to Opus 5.5. I'm afraid this looks unimpressive. Rather comical that they're delaying its launch "for safety reasons".
thereitgoes456•34m ago
What are you talking about? It’s comparable to Opus 5.5 on “high” (54 vs 53; $1.82 vs $1.99), crushes every model except the most modern OAI/Ant ones, has way lower hallucination than every existing model and probably broader support for multimodal like existing Gemini models. This is so ludicrously off base.
asdfasgasdgasdg•33m ago
From the charts, it's similar in intelligence and cost per task to both Opus 5.5 (high) and 6-Astra (max). It would be better if it were more intelligent and less expensive, but I don't see a reason to expect it to have better performance than models released around the same time.
godbox•29m ago
To play the Devil's advocate, Claude loves chugging tokens, while Gemini appears to be quite a bit more conservative and efficient. I believe AA's price per task breakdown reflects this.
yipinwong•30m ago
I am still not sold on Gemini 4 Argon yet from the chart.

The price is enticing for cost per tasks, but let's see how it goes.

I have montly (cheapy) sub to gemini models and has been underwelming and lowered the tier.

dom96•24m ago
I'd love to run it on my benchmark but alas, Google not making it public prevents this.
godbox•24m ago
What a snooze fest. Another model that does not meaningfully improve on intelligence or price compared to its peers. Google has basically announced that they've "caught up" with the rest. I think they've been doing great work in the Flash department so seeing this is... underwhelming?
enraged_camel•19m ago
Based on some... rumors I've heard, this is their "Pro" offering. There is supposed to be an Ultra coming as well.
losvedir•16m ago
The only models beating it are on "max", while this is "high". There's no guarantee that those effort / reasoning levels compare, but there's almost certainly an "xhigh" or "max" version later which will score higher there.
augment_me•10m ago
I found that Google does not benchmaxx as much as the other providers. Of you look at real-case evaluation like lm-arena, even the 3.8 flash is often near the top despite its benchmark index being worse.
ai-x•20m ago
Note: Google can sell their tokens at cost if they really want to drive out competition, but long term they are better off by everyone making a healthy margin (and Google does a double-dip by also selling compute, services).

So, like any optimal game theory move, they are better off not starting a price war

pooper•14m ago
A slightly different question would be can they afford to "sit out" of any price war?
aliljet•20m ago
It's hard to not see this as a gut punch for OpenAI. They're lead was largely captured by scoring on value (by way of reset after reset) and now they're getting eaten up on price and being bestes and equalled on performance. I'll still pay a premium for Opus 5.5 right now because it's nearly unlimited use, but Google is the quiet sleeping king Everyone is happy to watch everyone else, but I'd wager google burns more tokens through their search product than basically anyone else and now they're just quietly pacing the frontier...
tomrod•14m ago
They own their hardware. That vertical integration alone probably saves oodles because they can reconfigure to their needs as opposed to individually negotiating data centre plans. I can't imagine the complexity both OpenAI and Anthropic have to maintain for their deployments.
zozbot234•8m ago
It's not as smart as Claude Opus 5.5 High according to the AA benchmark. Looks like a big fat nothingburger so far, though it's possible that future fine-tuned checkpoints of the same pretrained model will do a lot better.
aleqs•8m ago
> Opus 5.5 right now because it's nearly unlimited use

Anthropic has some of the lowest usage per $ in general, not sure what you're taking about.

algoth1•14m ago
The most impressive jump for me is in the low hallucination rate, which is specially impressive given how bad Gemini current models are on this regard
dang•14m ago
Related ongoing thread:

Gemini 4 Argon - https://news.ycombinator.com/item?id=49913571

Gemini 4 Argon

https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-4-argon/
673•bradleyg223•2h ago•412 comments

Surprisingly complex waves reveal the brain's inner workings

https://www.quantamagazine.org/surprisingly-complex-waves-reveal-the-brains-inner-workings-20260930/
79•ibobev•3h ago•18 comments

EDG C++ front-end goes public

https://edgcpp.org/#transition
94•iandinwoodie•2h ago•33 comments

Launch HN: Magnitude (YC S25) – Self-optimizing inference engine for agents

https://github.com/magnitudedev/magnitude
103•anerli•4h ago•43 comments

5x faster Edge Functions: V8 isolates to Firecracker MicroVMs

https://www.netlify.com/blog/edge-functions-firecracker-microvms/
75•jbott•3h ago•25 comments

Halfspace experimental IDE for solid modeling with distance fields

https://www.mattkeeter.com/projects/halfspace/
36•luu•2h ago•6 comments

A brief history of the Bloomberg terminal

https://spectrum.ieee.org/bloomberg-terminal
195•rbanffy•7h ago•78 comments

CHOMPI portable sampler instrument is now open-source (hardware and software)

https://www.chompiclub.com/opensource
13•lashkari•4h ago•2 comments

You said no MCP

https://earendil.com/posts/you-said-no-mcp/
569•yarapavan•12h ago•325 comments

Singapore govt dating app uses Gale-Shapley stable marriage algorithm

https://twitter.com/tuakdotsol/status/2105105417760391258
63•rzk•12h ago•8 comments

Doing a Machine Learning PhD While Working in Japan

https://www.tokyodev.com/articles/doing-a-machine-learning-phd-while-working-in-japan
15•pwim•14h ago•4 comments

Functional Ultrasound Imaging (fUSI) from scratch

https://www.neuroai.science/p/functional-ultrasound-imaging-from
11•pminimax•1h ago•1 comments

I could've accessed 17T Microsoft records

https://blog.faav.net/how-i-couldve-accessed-17-trillion-microsoft-records
227•luispa•2d ago•98 comments

Great Dirhombicosidodecahedron ("Miller's Monster")

https://www.software3d.com/MillersMonster.php
12•cobbzilla•5h ago•1 comments

Why the Bronze Age Collapsed

https://www.worksinprogress.news/p/why-really-caused-the-bronze-age
13•AnodicElegy•1d ago•7 comments

Bild AI (YC W25) Is Hiring a Founding Product Engineer

https://www.ycombinator.com/companies/bild-ai/jobs/dAbC3Gd-founding-product-engineer
1•rooppal•5h ago

CS240 AI Cheating Retrospective

https://turkeyland.net/thoughts/ai.php
24•ArchAndStarch•2h ago•10 comments

Dear Software Makers

https://blog.jim-nielsen.com/2026/dear-software-makers/
26•speckx•2h ago•9 comments

The last time my family was replaced by technology

https://manuel.darcemont.fr/posts/the-last-time-my-family-was-replaced-by-technology/
154•megalomanu•9h ago•361 comments

What TLA+ can and can't check

https://buttondown.com/hillelwayne/archive/what-tla-can-and-cant-check/
115•b-man•8h ago•27 comments

Before pixels: Modular industrial dashboards

https://unsung.aresluna.org/before-pixels-modular-industrial-dashboards/
13•leephillips•3h ago•2 comments

Gitea 28.0

https://blog.gitea.com/release-of-28.0.0/
46•porridgeraisin•1h ago•16 comments

Burning Man death rates – A short lesson in statistics

https://ihavenapkinthoughts.substack.com/p/burning-man-death-rates-a-short-lesson
84•viraj_shah•2d ago•104 comments

Gemini 4 Argon (High): Intelligence, Performance and Price Analysis

https://artificialanalysis.ai/models/gemini-4-argon
42•theanonymousone•1h ago•20 comments

SDF vs. MSDF vs. Slug: GPU Text Rendering

https://alphapixeldev.com/sdf-vs-msdf-vs-slug-vs-rive-gpu-text-rendering/
117•ibobev•8h ago•48 comments

Commit description as a thinking tool

https://yedhu.me/posts/commit-description-as-a-thinking-tool/
102•yedhukrishnan•4h ago•58 comments

Dumb servers, smart clients: Distributing OS images with Quarry

https://amutable.com/blog/distributing-images-quarry
14•Levitating•9h ago•0 comments

Recursive `make` and `-j`

https://quuxplusone.github.io/blog/2026/09/30/recursive-make-j/
14•ibobev•4h ago•4 comments

Responsible Release of AI-Generated Mathematics

https://agmai.org/general-sep29/
60•aureianimus•19h ago•61 comments

SDF Public Access Unix System ... est. 1987

https://sdf.org/
73•kmstout•7h ago•13 comments