frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Morpion Solitaire

http://www.morpionsolitaire.com/
1•jackettm•43s ago•0 comments

High-mileage electric cars are more reliable than petrol ones, study finds

https://www.theguardian.com/environment/2026/sep/22/high-mileage-used-electric-cars-reliable-petr...
1•theanonymousone•1m ago•0 comments

A Call for Control of Frontier AI Models

https://www.government.nl/documents/2026/09/22/a-call-for-control-of-frontier-ai-models
1•geox•2m ago•0 comments

Why banning AI in the classroom will not be enough

https://theconversation.com/why-banning-ai-in-the-classroom-will-not-be-enough-to-remedy-poor-rea...
1•jruohonen•3m ago•0 comments

Show HN: Cookbook – a workspace for your team and agents

https://cookbook.team
1•dproz•6m ago•1 comments

Does an open-weight decision model beat a hosted one? Jev vs. Laya

https://astgl.com/p/local-laya-vs-hosted-jev-typed-decisions
1•Jmeg8r•7m ago•0 comments

Zero-downtime Linux kernel zero-day mitigation via eBPF and SECCOMP

https://github.com/mc493/linux-kernel-zero-day-mitigation-zero-downtime-kernel-defense-
2•mc493•10m ago•0 comments

Show HN: a0flow — paid micro-APIs for AI agents, x402/USDC, no signup

https://a0flow.com
1•updatesbyai•10m ago•0 comments

International Conference on Functional Programming (ICFP) 2026 talks released

https://www.youtube.com/@acmsigplan/playlists
1•matt_d•11m ago•0 comments

What does this assembly code do?

https://www.pagetable.com/301
1•chorlton2080•12m ago•1 comments

AgentsView

https://www.agentsview.io/
2•skadamat•13m ago•0 comments

Playing with Jev for Daily Questions

https://yesno.fyi
1•epv•15m ago•0 comments

MiMo-v2.6-Flash

https://mimo.mi.com/models/en-US/mimo-v2.6-flash
1•smokeeaasd•16m ago•0 comments

HTTP QUERY Method: The Grey Zone Between Get and Post

https://isc.sans.edu/diary/33352
1•jruohonen•17m ago•0 comments

The Apple Watch Has a Problem(YouTube – Marques Brownlee) [video]

https://www.youtube.com/watch?v=pOX1l1edBME
1•DASD•18m ago•0 comments

Tell HN: Substack obfuscating text to break reading mode

3•xnx•19m ago•0 comments

DoorDash Agrees to $131.5M Settlement for Shortchanging Workers

https://www.nytimes.com/2026/09/22/nyregion/doordash-settlement-delivery-drivers.html
4•datsci_est_2015•20m ago•1 comments

Claude Opus 5.5 (High Effort) Intelligence, Performance and Price Analysis

https://artificialanalysis.ai/models/claude-opus-5-5-high
1•Topfi•21m ago•0 comments

Faster and local Jev like model for Mac

https://github.com/mizorewww/laya-coreml
1•mkagenius•24m ago•0 comments

Show HN: Botzilla – Web automation using visual scripting

https://www.botzilla.dev
1•btzll•25m ago•0 comments

Show HN: Notes on Agentic AI – A text-first guide for practicing engineers

https://github.com/mrsachindixit/agentixit
1•tbaadm•25m ago•0 comments

Enjoy Every Sandwich

https://bradmontague.substack.com/p/enjoy-every-sandwich
1•NaOH•25m ago•0 comments

Shipping our game to twelve platforms on day one

https://www.m2h.nl/writing/shipping-to-nine-platforms-on-one-day/
1•MikeHer•27m ago•0 comments

Show HN: Last Internet Connection

https://github.com/Rubinoslaw/Last-Internet-Connection
1•PolishDevelop22•27m ago•0 comments

Harper Lee first edition novel found in Oxfam shop

https://www.bbc.com/news/articles/cmd79wn72ngjo
1•theanonymousone•27m ago•0 comments

Meta Tests Muse AI Agent Calls That Are Made by Humans in a Call Center

https://www.404media.co/meta-tests-muse-ai-agent-calls-that-are-actually-made-by-humans-in-a-call...
2•toomanyrichies•29m ago•0 comments

Unreal Agent

https://github.com/unreallabsai/unreal-agent
2•trollied•29m ago•1 comments

Show HN: Parametric, low-poly assets for BIM/CAD

https://3dassetstudio.com
1•matt3D•32m ago•0 comments

Show HN: Dragonfly – an API client that understands your Express/Next.js routes

https://www.usedragonfly.xyz
1•saurowankhade•32m ago•0 comments

ADIE – Beyond Prediction. Toward Verifiable Decisions

https://robertbudai.github.io/adie-website/
1•DIXOR•33m ago•0 comments
Open in hackernews

GPT-6 Sol and Luna

https://openai.com/index/introducing-gpt-6-sol-and-luna/
324•OfficialTurkey•44m ago

Comments

hehimself•42m ago
Love the price reductions across major players
madduci•40m ago
Because Qwen4 has been announced!
blovescoffee•16m ago
and to squeeze anthropic, and other research innovations, not just chinese models but those help bring price down
eloisant•9m ago
And GLM 5.3 works great
beardsciences•41m ago
There's no way this wasn't meant to coincide with Anthropic's release today.
jstummbillig•39m ago
They hinted this release last week for tuesday already, so if anything it would be Anthropic that tried to make this happen. But I doubt it.
jrflo•25m ago
Altman said it was launching last week on twitter, but they pushed it back to this week
Cu3PO42•41m ago
Cutting prices by 50% as compared to 5.6 prices is exciting. GPT-6 Luna at $0.10/Mio input tokens and $0.50/Mio output is positively insane.

EDIT: this doesn't say anything about availability on either Azure or AWS. I'm assuming it will show up later, but it would be interesting if it didn't.

potwinkle•40m ago
Very nice in cost/1mtok. Looks like more work is being done for efficient everyday helper models as time goes on.
Readerium•40m ago
Opus 5.5 seems better? Can someone attach both scores
hehimself•39m ago
Not the direct competitor to Opus 5.5, cuz 6 Sol is 50% cheaper.
Readerium•28m ago
Same price on Cache Reads 0.2/M So won't be 50 percent cheaper, more like 25% cheaper assuming half cost is cache read.
blovescoffee•17m ago
cost is dominated by non cached reads
droidjj•40m ago
Not only is GPT-6 Luna better, it's 50% cheaper. It was already practically free on a pro plan.
nickandbro•38m ago
Pricing is insane, can have Luna going after a goal for 10 days and not run into maxing out the limits.
physicallyIllfr•18m ago
Why would you do this though, surely these long running /goal tasks just like letting a wild animal out into your code base.

Does anyone care about code quality anymore?

blovescoffee•14m ago
1. you can have luna clean up after itself and improve code 2. you might be doing something like video-editing, cad modeling, artistic direction, pcb routing, etc. that need to run a long time to "converge"
pookieinc•38m ago
I don't see how anyone can be using Claude with prices like this, it's pretty incredible what the OpenAI team is doing, w.r.t model quality and pricing.

  Prices per 1M tokens     Claude Opus 5.5    Claude Opus 5
   Cache reads              $0.20              $0.50
   Input tokens             $4                 $5
   Output tokens            $20                $25
   Cache writes             $5                 $6.25


Model

Input

Output

Price reduction

GPT‑6 Sol vs. GPT‑5.6 Sol

$4 → $2

$20 → $10

50% cheaper

GPT‑6 Luna vs. GPT‑5.6 Luna

$0.20 → $0.10

$1.20 → $0.50

50% cheaper

thereitgoes456•37m ago
These don’t necessarily reflect actual costs, OpenAI is not profitable and nowhere near. They’ve lost their market lead and Sam may feel they need to get it back with any means necessary.
giancarlostoro•36m ago
The last time I gave GPT a shot, it ate all my tokens and got nothing meaningful done.
Readerium•36m ago
A vs B

Should be B vs A correct?

Else it's confusing

bitmasher9•35m ago
GPT would charge more if they could. Both companies need way way more revenue. GPT simply made a calculation that they can earn more money by charging less than their competitors.
mchusma•38m ago
What a day! I couldn't really use the last Luna for much (wasn't smart enough) or Astra (too expensive). So this release is really exciting. I can probably use Sol 6 as much as I want in the week, which as great.
sfkgtbor•37m ago
I'm glad both labs noticed and are trying to improve the models communication styles, they were getting closer and closer to meaningless gibberish.
Readerium•37m ago
6 Sol Performs worse than 5.6 Sol at DeepSwe?

Wierd!!

samuelknight•37m ago
No terra it seems? Luna 5.6 is great for token churning so it will be exciting to try the new one.
MrBuddyCasino•32m ago
Perhaps people realized that Luna Max is ~ Terra?
o_m•20m ago
Nah, Luna uses was more tokens and fills the context up way to fast. Terra is in the sweet spot where if feels like Opus 4.6. Competent but not too smart. It also uses lets you have longer sessions (back and forth) without filling the context too fast.
minimaxir•17m ago
Terra is the middle-child in more ways than one. It has much lower usage than Sol or Luna (going off OpenRouter).
cesarvarela•36m ago
It looks like the optimal pattern is to have Astra as the orchestrator and Sol as the implementer. Same as with Fable and Opus.
petesergeant•28m ago
I've found Astra to be horrible at making orchestration decisions. I will be trying to use Sol for both. Fable is very good at it though. Worst part of my week is when I hit my Fable usage limit and have to switch to Astra.
recitedropper•35m ago
[flagged]
wahnfrieden•34m ago
And comments complaining about other comments too
dominotw•32m ago
and comments complaining about other comments complaining too
rtaylorgarlock•32m ago
Pretty sure this is the reason HN exists ¯\_(ツ)_/¯ lol
ronsor•32m ago
Including this one, yes.

But the reason people say "Claude can't compete" is because Claude Opus has been going downhill since 4.7, and many have found Opus 5 intolerable. Fable is much better, but also much more expensive than OpenAI's offerings.

CaptWorld•28m ago
And comments complaining about how only big businesses get benefitted too
droidjj•33m ago
Is this comment about astroturfing or a decline in comment quality? To be honest, I was one of those early commenters, and I was just genuinely shocked at the price drop. I am also excited to try Opus 5.5!
theanonymousone•35m ago
Third-party inference providers will have a hard time to beat Luna in pricing with comparable open models.
badatnames•34m ago
It's asking a lot to trust they can or will maintain this new pricing. In any case it's exciting to think this might lead to further price cuts in the highly competent and competitive Chinese clones. I'm still using ChatGPT for interactive queries, but at this point pretty much only because of its familiar UI
chaos_emergent•28m ago
Curious why you think it's unsustainable?
cmrdporcupine•24m ago
Well, they do rug pull constantly. This week and last leading up to this the cost to use Codex was overwhelmingly perceived as terrible. People running out of usage all over the place. Reddit full of people crying. I noticed it myself.

Then they do a new model launch, issue quota resets all around, and it's a party for 2-3 weeks before things return to normal.

FergusArgyll•18m ago
Oh, I'm happy I'm not the only one. Astra was feasting on tokens!
cmrdporcupine•16m ago
It wasn't just Astra. Sol 5.6 was a hog, too. They futzed with the formula and it pissed people off royal.
badatnames•21m ago
Because at some point keeping it up involves filing an S-1 that doesn't look like a garbage fire
cmrdporcupine•32m ago
Looking at their own charts it seems like it's only small incremental improvement over 5.6 Sol, but with a massive cost reduction. And the better writing/communication style that Astra had.

Which... fine, I'll take that.

Ninjinka•30m ago
so opus 5.5 is smarter and cheaper than fable, and sol 6 is a little dumber and WAY cheaper than astra? is that right?
Readerium•26m ago
Sol 6 is also dumber than Sol 5.6 on some tasks (DeepSWE)
eyk19•27m ago
Luna really is "intelligence to cheap to meter" by now
mrdependable•15m ago
Wouldn't that mean the cost of metering it is more than they make from metering it? I don't think that is the case.
jrflo•26m ago
The only two benchmarks shared between the Opus 5.5 and Sol 6 launch seem to be frontier code and automation bench, looks like Sol wins on automation bench (same performance for half the cost) and Opus 5.5 wins on frontier code (2-5% better scores across the board for same cost)
m_fayer•26m ago
I've been working with agents all year, but 5.6 Sol was some sort of sweet spot for me. Something about how it communicated verbally and its engineering instincts just clicked for me, and I was able to somehow predict it and jam with it. Like a colleague you click with. It's the first model I've gotten attached to. I'm concerned that whatever model supercedes it, while technically better, just won't feel quite as natural to work with. And this makes me feel very professionally vulnerable to the labs. I miss the days when my crucial tooling came from companies as reliable and predictable as, say, Jetbrains.
redox99•22m ago
Same. In fact I found 6 Astra to be a downgrade in situations where I didn't need the extra intelligence.
cmrdporcupine•13m ago
Yeah.

Astra was/is superior for planning type tasks. It was capable of doing seemingly magic things with rather vague/lazy instructions ("I need to be able to test this on Windows, maybe a qemu VM or something? Shrug." ... 1 hour later "yeah i built you a whole qemu + eval windows image + harness of powershell scripts + shell scripts to retrieve & verify harness.").

And for UI work -- which is not something I do a lot of but do here and there -- it was clearly superior to 5.6 Sol.

But it also feels sloppier? Somehow. And too expensive to use.

We'll see how Sol 6 is.

jeffnash•4m ago
I felt this way with Sol in the 5.6 series and was one of the seemingly few people on this earth who liked Terra for that reason. I would often have a very specific code-manipulation ask, e.g. "add a parameter to this method, ensure all callers pass it in, if there is not a logical way to derive the parameter to be passed in a particular instance, flag this in your final response", and Sol would go on some rabbit hole side quest to refactor my codebase to determine some way to derive it rather than flagging it as I had asked.

Terra had the "workhorse" quality where it could do these changes in bulk and follow directions without being too 'smart' (but sloppy) as you described. Luna was a bit too dumb and would make sloppy mistakes; I see that more as a "run these tests and format the results" sort of model. Maybe 6 Luna will be better.

I also just reread your comment and realized the naming convention is still extremely confusing with respect to ordering of [Family]x[Model]x[Number].

meerita•23m ago
OpenAI, Antrophic and others are operating with 80% margins. They can lower the prices for a long while.
simianparrot•22m ago
Well at least it looks like OpenAI is dogfooding because their announcements, product names, and everything else looks and sounds like LLM-slop.
sehw•21m ago
sage
msh•21m ago
I dont understand why there is not a gpt-6 terra?
blovescoffee•18m ago
It wasn't really used enough and it sat in an awkward middle space between luna and sol where either luna high/xhigh or sol med were better cost/perf wise
OutOfHere•14m ago
I don't agree. In "none" thinking mode, Terra serves a useful purpose where high but not top-tier intelligence is needed. Luna doesn't cut it.
Tadpole9181•18m ago
It probably didn't see that much use, as it struggled to find a niche. If you wanted intelligence tasks, Sol was cheap enough and much smarter. If you wanted performance and cost-effectiveness, Luna was significantly better value while being only a little less intelligent.

Terra ended up just being an awkward middle ground that was not particularly suited for any workload.

OutOfHere•13m ago
The users of Terra disagree. Specifically, Terra is useful when high but not top-tier intelligence is needed in instant ("none" thinking) mode.
msh•13m ago
I have found it worked quite well as the workhorse model in my hermes agent.
devinprater•20m ago
Good. Maybe they can use GPT-6 to fix the accessibility of their iOS app. Output shows as text fields to VoiceOver, and the accessibility announcements have backslashes before seemingly every punctuation mark. And then bring accessibility announcements to the Android app so I don't have to make a whole new app just to add that through an accessibility service. Ugh the things I do for accessibility cause I'm blind. On a better note though, AI has done so much for the blind community, from image (and increasingly video) description to mods for video games like Final Fantasy 1 through 6 Pixel remaster, I have a ton to be grateful for.
markerbrod•20m ago
Does anyone know if the ~50% price reduction also implies x2 subscription usage? Or is it only for the API.

Edit: Yes, it applies also to subscriptions, source https://x.com/thsottiaux/status/2102463847714247142

fHr•16m ago
Luna is the goat for real, cost intelligence ratio is insane already and it is enough for most daily computer use.
yipinwong•16m ago
I've been raving about Luna 5.6 as it's dirt cheap, and "intelligent enough". Double quoted.

Now GPT 6 Luna is even cheaper, and more intelligent, there is no going back... to SOL 5.6 for intelligent layer.

dmazin•13m ago
Per the benchmarks in the post, Luna 6 is at best a couple points superior to Luna 5.6 and (unless I’m reading it wrong) xhigh has actually degraded in quality?

I was hoping for a serious Luna upgrade. It was already cheap enough. This feels more like a price reduction than an upgrade.

That said, if the new Luna is able to handle ultra mode and subagents v2 in codex cli, then at least that’s a win.

recitedropper•15m ago
This is the most blatantly astroturfed thread I have ever seen on Hacker News.

My previous comment--which suggested astroturfing--was the highest upvoted comment here until it got flagged. Which implies to me that atleast the other remaining humans on this forum see it as well.

Clearly organic: 26 minutes, 89 comments, upvoted instantly to the top, posted within one hour of the Opus 5.5 announcement. Please @Dang, after this comment is flagged as well, delete my account.

john_strinlai•7m ago
>My previous comment--which suggested astroturfing--was the highest upvoted comment here until it got flagged

fyi, i flagged it because it is boring reading and against the rules.

if you suspect astroturfing, flag the comments and contact the mods.

(complaining that your complaint got flagged is also tiresome. contact the mods. "@dang" doesnt work, use the email.)

recitedropper•4m ago
Do you think the "you can't mention astroturfing" rules are really serving HN these days? Do you think this thread hasn't been manipulated?

I respect you for replying here though, and yes I get that HN forum standards would suggest flagging my previous comment. But it is just sad to see a place used to be so vibrant get manipulated because of how much weight it holds for us in the industry.

And yea sure, I could go and flag all the bots and message Dang. But probably time to stop shouting into the void. :)

jeffnash•15m ago
At this point, the deciding factors for me between Claude Code 20x and Codex Pro 20x are:

1/ Usage limits: downstream of input/output cost, but resets and obscure windows and odd 20x plan / 5x plan != 4x usage math throw a wrench into it. Winner right now is Codex by a mile, especially when you factor in ChatGPT usage (even 6 Astra Pro) is essentially unmetered on the 20x plan. Always a bummer when asking if I should see a doctor about a rash means I can't code as much. It's also is a godsend if you use an MCP like oracle to automate the process of calling the Pro model on particularly tough problems, giving better planning results or deeper code analysis without burning usage.

2/ Context window in the harness. Claude Code wins on this. There used to be a toml file workaround for Codex to extend the GPT context window to 1m, but this stopped working on the plans and only on per-token billing. 252k is just not enough. Codex's compaction is very good, fwiw, but it happens so frequently that even a model as powerful as Astra sometimes loses the plot on long-running tasks.

3/ Ability to use the plan outside of the official harness. Codex wins. Anthropic does shit like bills requests as extra usage if it sees a hermes.md in a commit.

I've subscription hopped a bunch, and at times I've had both, but I keep coming back to Codex because it wins on 2/3.

spijdar•9m ago
I dunno about Codex-the-application itself, but you can definitely use e.g. Pi with the larger context windows with a Codex login. It puts a pretty large multiplier on credit usage, however.
noname120•9m ago
> It's also is a godsend if you use an MCP like oracle to automate the process of calling the Pro model on particularly tough problems

As far as I know Codex (at least the GUI) can automatically call the ChatGPT Chat models (including Astra 6 Pro), you just need to @ a ChatGPT Chat conversation from within Codex and tell it when to use it.

> There used to be a toml file workaround for Codex to extend the GPT context window to 1m, but this stopped working on the plans and only on per-token billing

Not true, it works again[1]. I confirm that it works both on 5.6 Sol and Astra 6, possibly other models too.

[1] https://x.com/thsottiaux/status/2089082893804896524

GodelNumbering•15m ago
Gpt 6 Luna is cheaper than Deepseek 4.1 flash! Today is wild in terms of intelligence/price across the board!
Someone1234•14m ago
Have they solved GPT5.6 SOL's propensity to over-engineer and over-complicate? You'd ask SOL to do something relatively simple, and find four single-use methods, an interface, and a factory-factory.

I actually preferred 5.6-Terra not because it is technically superior (it isn't) but because it had better instincts to NOT do this stuff.

PS - Speaking of better instincts, have they closed the UI-design gap at all? I keep a Claude subscription just because /design produces significantly higher quality UI design/UI feedback/UI refinement than anything I've seen from OpenAI.

cmrdporcupine•9m ago
Astra 6 was a huge improvement over Sol 5.6 for UI work. I haven't tried Sol 6 yet for it (it's only been a few minutes).

The GPT / Codex models have always been "overengineer" personalities. I prefer that to "I left a pile of race conditions lying around and big gaps in testing" though, which is what I was getting from Opus at times.

But yes both Astra and Sol veer on the side of paranoid. And honestly that's better for team work. For solo work where you just want to yeet something, it can be tiring.

You learn to tame the GPT "personality" on this front by combing over once a week and asking it to find and exterminate pointless tests, clean abstractions etc.

msp26•14m ago
This Luna pricing is obscene man. 5.6 was good enough for so many use cases (data analysis, structured extraction etc).

Incredible.

ggcr•14m ago
Live notification in Codex:

> GPT-5.6-Sol is retiring. This conversation will automatically switch to GPT-6-Sol

I don't recall OAI retiring a model so early lol. Similar arch?

scrlk•13m ago
Artificial Analysis is reporting that 6 Luna scores 2 points lower on their coding index than 5.6 Luna, but is 60% cheaper:

> In the Coding Agent Index, Sol improves but Luna regresses: In OpenAI's Codex harness, GPT-6 Sol (max) scores 57 in the Artificial Analysis Coding Agent Index, up 2 points from GPT-5.6 Sol (max), with gains in Terminal-Bench 4.0 (43% vs 37%) and SWE-Atlas-QnA (58% vs 54%). At $2.99 per task it costs ~50% less than GPT-5.6 Sol (max) and sits on the Pareto frontier of Coding Agent Index vs Cost per Task. GPT-6 Luna (max) scores 41, down 2 points from GPT-5.6 Luna (max), with lower scores in SWE-Atlas-QnA (44% vs 49%) and DeepSWE v1.1 (64% vs 66%), at ~60% lower cost per task.

https://x.com/ArtificialAnlys/status/2102462962758033624

Readerium•7m ago
Yup more like a 5.7 than a 6
dmitrygr•13m ago
Selling dollar bills for $0.40 to undercut the guys selling them for $0.50 is a bold move. Let's see if it pays off for them.
ghoshbishakh•12m ago
So opus 5.5 has reduced price. Who is winning then?
hamburglar1•11m ago
Code deception 10% at 5.6 to 1.3% for 6.0? So models are getting more safe rather than less safe? hmmm
OutOfHere•10m ago
As a user of 5.6-Terra, I am sick and tired of the inconsistencies in GPT model families.
m3kw9•9m ago
The new default is 6.0 Sol high. Escalate to Astra-medium. If usage is tight go luna6.0-max
seatac76•7m ago
Would be funny if Google drops Gemini 4 today.
zaik•7m ago
Why is Claude missing on the "Factuality" graph?
simonw•3m ago
GPT-6 Luna being half the price of GPT-5.6 Luna is a really big deal.

Here's GPT-6 Luna pelicans: https://tools.simonwillison.net/markdown-svg-renderer?url=ht...

And GPT-6 Sol: https://tools.simonwillison.net/markdown-svg-renderer?url=ht...

Scroll to the bottom for the GPT-6 Sol max one: https://tools.simonwillison.net/markdown-svg-renderer?url=ht...

For comparison, here are the pelicans I got for GPT-6 Astra: https://tools.simonwillison.net/markdown-svg-renderer?url=ht... - I still like the Astra Max one best.

jumploops•3m ago
I’m still finding context is king, even with the best models.

For example, I had Fable review Astra’s output yesterday, and it found some issues and fixed them. Passing the fixes back, Astra then uncovered additional issues with Fable’s fixes (and yes, this will go on ad infinitum if you let it, but these were “real” issues).

It seems the big story here is the reduced Luna pricing. It’s a fantastic model that can handle most automation needs (though I still use the big models for day-to-day development).

wyre•22m ago
Any business would charge more if they could. Jevon's paradox would mean that they can make more money by charging less because demand is going to keep growing.
atq2119•11m ago
FWIW, what you're describing is a simple demand curve, not Jevons paradox.

The "paradox" is when an increase in efficiency which would decrease the use of a resource all else equal, instead indirectly causes more use.

blovescoffee•21m ago
Of course they'd charge more if they could... Of course they're pricing to outcompete their competitor...
blubber•16m ago
They also have postponed their IPO. So they don't have to be profitable that soon. Anthropic on the other hand plans to do the IPO this fall.
vanuatu•6m ago
HN discovers competition leads to lower prices
mchusma•34m ago
Opus 5.5 is incredible so far, its going to get used. Fable is much better than Astra for me in practice, and Sol is not marketed as better.

Its a great release, I will use both heavily.

etothet•32m ago
For API usage, sure. But plenty of people have subscriptions where these differences effectively don’t matter.
mfiguiere•32m ago
Also, batch processing prices are still 50% off, which put GPT-6 Sol and GPT-6 Luna at $5 and $0.25 for output.

https://developers.openai.com/api/docs/pricing?latest-pricin...

onlyrealcuzzo•30m ago
I could already run Sol High on 3 concurrent side projects 24/7 and not run out of quota.

This is great, but practically, I'm not going to start working on more side projects.

Perhaps in another 6-12 months I'll be fine to drop down to $20/m instead of $200.

charliegoforit•25m ago
How much does it cost you per month to have that much sol high usage and what do you use, api? Through what? Thank you
wyre•21m ago
they said quota so i would imagine the $200 subscription. Probably through Codex or Pi coding agents.
onlyrealcuzzo•11m ago
$200/mo

A lot of what I'm doing has pretty expensive build/testing processes between iterations - even on a 40 core machine - so I'm not burning tokens 24/7 like some people may.

I'd guess I'm probably spending >50% of the time running tests & build processes & tooling and the remainder is purely burning tokens.

I also have some internal tooling (that I will hopefully open source soon) that makes LLMs substantially more correct (thus more efficient) - so there's that, too.

shmoil•27m ago
>> GPT‑6 Sol vs. GPT‑5.6 Sol

>> $4 → $2

>> $20 → $10

Do you mean 100% more expensive? GPT 6 is 100% more expensive than 5.6 per your post.

blovescoffee•21m ago
It's before and after following the arrow. 6 is the cheaper one.
yzydserd•16m ago
Yes very poor proofreading!
edf13•25m ago
You also need to compare allowances on Codex vs. Claude Code
dyauspitr•24m ago
Wtf is GPT-6 Sol, I though GPT-6 is Astra?
Readerium•23m ago
Number is generation Name is the size (Luna smallest to Astra largest)
dyauspitr•20m ago
Then what is Astra high-extra high-Ultra? That’s effort within each tier?
MaKey•13m ago
Exactly
Readerium•9m ago
Yes that is number of reasoning tokens used.

Performance increases both with larger model (Luna vs Sol)

And with more reasoning (low vs xhigh)

LZ_Khan•23m ago
Disagree. I would never use OpenAI cause they're probably just going to steal whatever I'm working on.
AustinDev•22m ago
and anthropic won't? or any other inference provider? Running your own inference either locally or remotely are probably the only ways to make sure that doesn't happen.
solenoid0937•17m ago
Well we know for a fact that OpenAI steals Millennium Problem work from researchers. Have we seen anything similar from Anthropic?
vanuatu•6m ago
source? p sure they said they were confident they did not access the researcher's chats
nradov•20m ago
What are you working on? Is any of it actually worth stealing?
OutOfHere•8m ago
And why is that bad? As your brain gets older, it will not remain so clever, so you'll be grateful for an AI that thinks like you do when it comes to your line of work.
an0malous•21m ago
These are the pre rug pull prices. They'll increase prices 10x and nerf the models after they IPO.
solenoid0937•20m ago
Before IPO. This is why Anthropic isn't playing the same games
blovescoffee•10m ago
there are still competitive market forces for co's post IPO
selectodude•7m ago
Okay? I didn’t sign a 10 year contract. We’re month to month and I use my own harness.

If they’re subsidizing my usage, that’s great.

persedes•16m ago
Not that anthropic models are very good at this, but due to the changes in tokenizers and thinking tokens: cost per token is not as helpful anymore as cost / task.
joshstrange•14m ago
As someone who has used Claude Code and Codex the prices don't matter in the same way but I found that I burned through my usage way faster on Codex even though I regularly hear that the Codex plans go further. That was not my experience and the intelligence was comparable to what I was getting in Claude.

If these price changes mean that coding plans have effectively more usage then that's great, but Codex is surviving on resets from my own experience using it. I was glad to go back to Claude.

hombre_fatal•10m ago
I mainly use Codex/Sol to review my plans drafted by Fable. But beyond that, Astra blows through usage limits too fast to be a daily driver and writes weird code despite what my "house style" is, and Codex is behind Claude Code in terms of critical features like seeing what's going on in subagents.

The parent + subagent workflow has become critical for keeping the reasoning agent (parent) context-lean while also letting me chat to the main agent while work is getting done.

My main process is to use Fable to reason and then spawn Opus subagents, and I get amazing results, and I'm always looking into what the subagents are doing.

ignoramous•7m ago
> 50% cheaper

Did cache read/write also decrease by 50% or similar? That's where most (95%+) of the cost is for agentic coding workloads.

mydreamof•31m ago
In the other hand the pricies dropped by a big margin
saadn92•30m ago
bots everywhere
civvv•29m ago
Welcome to the new internet. It was fun whilst it lasted. Next evolution will likely be closed, invite only forums.
ZeWaka•25m ago
Eternal September 2, I suppose.
LeBit•20m ago
They already exists. You haven’t been invited? Hmmmmm
qoez•27m ago
Fundamentally I feel like coders just doesn't even need to be that smart anymore given AI assistance. This place ten years ago used to be filled with some of the most interesting comments/takes around for that reason.
LeBit•22m ago
It’s not that bad: https://news.ycombinator.com/bestcomments
monkeydust•24m ago
Stick with it. The collapse of HN is a leading indicator to the fall of humanity.
Madmallard•24m ago
I remember how a few months ago Dang was criticizing people for making comments like this. Guess he just realized how stupid that was and stopped bothering eventually.
minimaxir•10m ago
The comment got flagkilled, I'm unsure what else dang would need to do.
woah•22m ago
Are the prices very nice or not?
sidrag22•21m ago
Ya all these articles lately about how everyone is sick of reading AI prose, and interacting with models in general. Tons of new model optimizations and workflow optimizations or whatever. I'm not really aware of any idea or product aimed at making the internet usable, and making it somewhat resistant to the generated noise. I think HN is a bit better than reddit for this type of example for floods of comments, first movers on reddit REALLY rise to the top and stay there.
cmrdporcupine•21m ago
Are you trying to imply that nothing OpenAI can release would justify that response and therefore the people must be bots?

Asking cuz I don't think I'm a bot [pats self], I legitimately prefer the GPT models to Anthropic's, don't like Anthropic's customer service/reliability story at all, and I welcome a massive price reduction. Seems like something I should be happy to get.

If you'd told me I'd be typing this a year ago I'd be skeptical though.

wyre•20m ago
Didn't SpaceX already set a precedence for garbage fire S-1s? I don't think OpenAI has to worry about that?
badatnames•14m ago
SpaceX is a different beast with extremely high friction to enter its market, a massive technology lead, and well developed preferential high level relationships with just about every country worth worrying about.

OpenAI/Anthropic meanwhile feel a bit like they're hoping to sell iPhones in a market about to be flooded by $20 flip phones, with almost no channel of their own to do it. And for whatever mad reason OpenAI are now signalling they will attempt to compete on price with flip phones despite their cost of labour, energy, and just about everything else being far higher

NorthSouthNorth•20m ago
Completely agree. I've been using 5.6 still even with Astra available to me for most tasks. It's funny how much of this is just "vibes" because I cannot quantify what it is. Astra is definitely better when I have an ambitious feature, but in like 9/10 tasks I prefer working with 5.6 Sol. A few weeks ago when the limits were seemingly higher, having 5.6 on fast mode was a good time.
nickreese•19m ago
This is 100% my experience. I rarely reach for Astra as we speak.
mcast•18m ago
It's a shame the labs don't open source their models after deprecating them. I get why, but, it's a piece of internet history I hope is preserved.
jdw64•15m ago
I agree. Sol followed my instructions well and wrote good code.
BowBun•13m ago
This has been my experience for a year. Same with Opus models. This is how I think this tech will be best used in the long term - finding the one you vibe with most. Much like IDEs!
bradly•12m ago
Not only was 6 worse the 5.6 Sol for my me, but it went through my Plus usage in minutes, while I could cruise for hours with 5.6. It would churn minutes and then just give up on usage limits.

Highlight and lowlight of my week was successfully convincing the OpenAI support chat robot to give me a refund for the month for my issues with 6 chewing threw my usage with no output.

AaronAPU•9m ago
I had this experience as well, but after rewriting my agent instructions it has been far better. I believe Astra’s “token efficiency” translates to “don’t research as much” which caused it to make poorly informed architectural decisions.
smith7018•16m ago
I read that there are rumors that they're getting rid of that tier. No idea where the rumor came from, though. This lends credence to it, I suppose.
yawnxyz•3m ago
Terra was always worse of both worlds (expensive and not that good) so I think they're just retiring it
sodacanner•7m ago
In my personal experience I currently get a lot, lot more usage on the 5x Claude plan than the 5x Codex plan.

Having limitless webUI ChatGPT usage is much better user experience, though. I'll give them that.

(edit: Sol-6 is half the price, so maybe the usage limits are going to be way better.)

basisword•4m ago
I've been using Claude Pro and recently gave Codex a try again. Both on the $20 plans. I get so much more usage with Claude. It's night and day for me. Codex runs out constantly, whereas Claude I hit limits very rarely.