frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

SpaceXAI's Grok 4.6 Scores 61 on the Artificial Analysis Intelligence Index

https://artificialanalysis.ai/articles/grok-4-6-benchmarks-and-analysis
81•wertyk•1h ago

Comments

satvikpendem•50m ago
Cursor, since Grok 4.5, has had an incredible deal for frontier level models, their subscription now goes way further than OpenAI or Anthropic. Even on their lower tier plans you can use a lot tokens on their of their first party models (Grok and Composer) and not really run out comparatively. Combine them with an orchestrator and implementor type setup and it goes even further.
aliljet•47m ago
Can you explain what you mean? These days courtesy of an addictive reset game OpenAI is playing, I can't find anything with frontier intelligence that's more cost efficient...
jesse_dot_id•43m ago
Goes even further to exfiltrate your data, yeah.
greenavocado•24m ago
That would be Muse Spark Contributor Tier. 12-21x price reduction at the expense of your digital existence.
hmokiguess•13m ago
I believe they are the only western provider that has Kimi K3 on a subscription plan today as well. I would love to ditch Anthropic and be on Kimi if there were a subsidized plan like that with ZDR
nomilk•1m ago
How does Grok 4.5 compare to Opus >= 4.8 though?

I'm willing to pay 2x for a 10% smarter model. Intelligence matters that much (because 10% smarter probably saves, on average, several hours of human time).

thiago_fm•49m ago
I often wonder if there's a chance, even if minimal... that they stole the weights of the Anthropic models they run on their datacenter... or are actively destillating it.
connicpu•47m ago
I think the more likely explanation is that the Cursor data they effectively acquired for $10B was extremely valuable for their training when combined with the insane number of GB300s xAI has for training.
winstonp•39m ago
Cursor was 60B. The 10B number was the breakup fee if the deal fell through.
qudat•38m ago
> ... or are actively destillating it.

I just assumed every model manufacturer is distilling from the frontier models. If they aren't they are definitely trying to do it.

nylonstrung•47m ago
I have never met a single human being who uses Grok for coding
sidcool•46m ago
Hello. Nice to meet you.
supriyo-biswas•44m ago
I'm only being forced to use it at $WORK since some people overran their Cursor bill, so everyone gets Cursor Auto enabled by default which routes to Grok 4.5.
kvirani•43m ago
Folks working in US govt tend to, based on convos I've had with one such person.
visopsys•43m ago
I use. I used to be a Claude user. Since trying Grok 4.5 and especially Grok 4.6, I don't want to go back to Claude any more (I have early access to 4.6).

Grok is 3x+ faster than Claude and I can't tell the diff in engineering work quality. As an engineer, speed is important to me.

bigyabai•41m ago
For $30/month, I'd expect it to have higher usage limits than Claude Code and Codex.
visopsys•
pzo•39m ago
Seems the cache read pricing almost doubled from $0.30 in Grok 4.5 to $0.50 in Grok 4.6.

In my experience in heavy coding sessions most pricing is just cache read and cache write like 80% of my token bill.

sidcool•18m ago
Grok is not the best model around, but it's decent. It gets the basic job done at a low price. I don't think it can advance frontier Math, yet.
leerob•6m ago
Probably can't advance frontier math yet, yeah. But please let us know other places you want to see Grok improve for future models!
petesergeant•14m ago
Interesting. Grok 4.5 is a capable model, although not quite at Fable/Sol levels. Will be interesting to see how this holds up. Musk appears to have made a savvy choice buying Cursor's data.
32m ago
It really does. I felt like I could have spent $1000+ api token on claude for the amount of work on my $30 grok subscription.
drewnick•26m ago
An hour in, I've been running four terminals full bore on my $20/mo Grok sub and I'm at 9% for the week. Codex or Claude would easily have hit 5-hour or weekly limits.
DetroitThrow•42m ago
I've tried it on my "let's run every model in parallel and see which finds more edge cases" type of tasks, and Grok 4.5 was really behind Opus/ChatGPT but ahead of Gemini - despite having a strong showing on benchmarks.

That makes me really skeptical of it being GPT5.6-tier, much less Fable-tier, based on some of these benchmarks alone. But I'll test here shortly.

busch_j•38m ago
A bunch of SWEs at my work use it as their primary model.

We have Claude, ChatGPT, and Cursor with essentially no cap on spend (top guy is spending over 10K a month on AI at API prices), and he hasn't had his hand slapped.

So it's not like they are using it purely because it's cheaper.

I think people like to use it for its speaking style, pretty solid performance, and its speed.

LeBit•37m ago
I refuse to use that product because of the parent company.
Recurecur•19m ago
I wonder why folks feel the need to virtue signal like this…?

As an aside, do you use Apple equipment despite Apple’s use of Chinese slave labor?

wilburx3•2m ago
100%, it's an easy pass given that it is always playing catch up.
jm4•35m ago
I have. He was using it due to philosophical reasons the same way many people have philosophical reasons for avoiding it. I don't know how many people are like that, but it's not exactly where you want to position your product if you're a business.

Personally - and I know I'm not alone with this sentiment based on comments I see on this site - I wouldn't touch Grok no matter how good or cheap it is. I don't trust Elon and I don't want to give another dollar to the world's richest person who turns around and uses the money to interfere with elections. The guy I know uses it for essentially the same reason I won't use it.

peder•31m ago
Does it matter? Why turn it into a popularity contest?
Recurecur•24m ago
Hi! Grok’s worked quite well for my use cases…

It’s also a great deal!

locknitpicker•15m ago
> I have never met a single human being who uses Grok for coding

Me too. The only people I ever saw using grok were using it by accident as they used copilot in auto mode and noticed some prompts were thrown it's way.

I saw far more people using Mistral than grok.

treexs•6m ago
the new models are quite good, give it a shot

DeepSeek V4 Pro 0813

https://openrouter.ai/deepseek/deepseek-v4-pro-0813
271•explosion-s•1h ago•82 comments

Tailscale Traces Database Corruption to 16y/o SQLite WAL-Reset Bug

https://tailscale.com/blog/sqlite-wal-reset-bug
422•ropbear•3h ago•64 comments

2026 Eclipse Webcams

https://jonty.github.io/2026_eclipse_webcams/
381•zoenolan•6h ago•95 comments

Glaciers on the Climate Dashboard

https://climate.metoffice.cloud/glaciers.html
41•mooreds•1h ago•3 comments

Tim King, AmigaDOS developer, has died

https://amiga-news.de/en/news/AN-2026-08-00070-EN.html
121•doener•3h ago•22 comments

Someone is running mass vulnerability scans, spoofing AI bots like ClaudeBot

https://knownagents.com/insights
126•gavinhking•3h ago•76 comments

Why Tiny JPEGs Look Different in Chrome

https://guillaumetech.github.io/posts/jpg-scaling-chrome/
156•gutechh•3h ago•28 comments

Reflex (YC W23) Is hiring Growth and GTM Roles

https://www.ycombinator.com/companies/reflex/jobs/71x5GFb-growth-engineer
1•apetuskey•57m ago

Wednesday, August 12: GitHub, Incident with Pull Requests and Issues

https://www.githubstatus.com/incidents/76t89hbfb09h
34•arm32•1h ago•4 comments

SpaceXAI's Grok 4.6 Scores 61 on the Artificial Analysis Intelligence Index

https://artificialanalysis.ai/articles/grok-4-6-benchmarks-and-analysis
82•wertyk•1h ago•32 comments

HTML over WebSockets: real-time SPAs with barely any JavaScript

https://en.andros.dev/blog/ef4968f5/html-over-websockets-real-time-spas-with-barely-any-javascript/
18•redbell•1h ago•13 comments

License plate reader searches should require a warrant

https://andrewpwheeler.com/2026/08/12/license-plate-reader-searches-should-require-a-warrant/
347•apwheele•3h ago•214 comments

AI is removing the middle class of software engineering

https://blog.florianherrengt.com/ai-removing-middle-class-software-engineering.html
426•florianherrengt•4h ago•357 comments

Bike Bureau: Report Bike Lane Obstructions

https://loudbicycle.com/bb
25•kdmccormick•1h ago•14 comments

Qwen/Qwen3.8-2.4T-A95B

https://huggingface.co/Qwen/Qwen3.8-2.4T-A95B
200•Philpax•2h ago•60 comments

Pixel Watch 5

https://blog.google/products-and-platforms/devices/pixel/pixel-watch-5/
18•ortusdux•1h ago•9 comments

We just raised $400M in Series C

https://lovable.dev/blog/series-c
28•thoughtpeddler•1h ago•8 comments

What sort of maths are LLMs good at?

https://gowers.wordpress.com/2026/08/12/what-sort-of-maths-are-llms-good-at/
199•ColinWright•7h ago•96 comments

The Bit Player: My Father with Steve Zissou

https://www.theparisreview.org/blog/2026/07/27/the-bit-player-my-father-with-steve-zissou/
11•Thevet•4d ago•0 comments

Shade Map

https://shademap.app
68•fredley•4h ago•19 comments

Show HN: Woxi - Open-source Mathematica / Wolfram Language reimplementation

https://woxi.ad-si.com
211•adius•7h ago•30 comments

Hax – a minimalist, terminal-native coding agent written in C

https://usehax.dev/
39•OleksandrC•3h ago•11 comments

Qwen3.8-2.4T

https://huggingface.co/Qwen/Qwen3.8-2.4T-A95B-FP8
16•mmastrac•1h ago•0 comments

Delphi 13 Community Edition Is Now Available

https://blogs.embarcadero.com/delphi-13-community-edition-is-now-available/
122•layer8•6h ago•85 comments

Felix and I

https://jacobfilipp.com/felix/
48•surprisetalk•2d ago•3 comments

Automatic1111 for Apple metal, 40% speed up sd1.5

https://therad.ninja/from-8-10-seconds-to-3-7-teaching-automatic1111-to-speak-metal-on-an-m3-pro/
46•dmikey831•4h ago•18 comments

Solving the Shortest Vector Problem in $2^{0.6039n}$ Time via Mid-Point Hessian

https://arxiv.org/abs/2608.02478
25•sbulaev•1w ago•6 comments

High-Res Photo Shows Sand-Capped Butte Rising from Mars Plain of Polygons

https://petapixel.com/2026/08/04/amazing-high-res-photo-shows-a-butte-rising-from-mars/
127•bookofjoe•6d ago•10 comments

My Agent Setup

https://chad.cm/posts/2026-8-11-my-agent-setup
72•carimura•4h ago•39 comments

Bigos (Polish Hunter's Stew) Recipe Builder

https://chefsbinge.com/bigos-recipe-builder/
62•doublepg23•5d ago•21 comments