frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Pi.dev: You Said No MCP

https://earendil.com/posts/you-said-no-mcp/
120•yarapavan•1h ago•43 comments

Livenerf: Has Opus 5.5 been nerfed yet?

https://github.com/ninjahawk/livenerf
673•bryan0•12h ago•266 comments

Show HN: JBR-001 – An open-source 3D printable desktop robot

https://projecthub.arduino.cc/syntheticaidata/jbr-001-a-desktop-companion-robot-powered-by-arduin...
14•gvuksic•1d ago•4 comments

Solving Factorio Quality

https://exyr.org/2026/solving-factorio-quality/
105•laurenth•1d ago•33 comments

September 2026: The world today, as seen by one Polish guy

https://tomwojcik.com/posts/2026-09-21/september-2026-the-world-today/
293•marjancek•4h ago•166 comments

Dots: Always-on agents

https://openai.com/index/introducing-dots/
644•alvis•18h ago•504 comments

Vermont replacing power plants with home batteries

https://www.bbc.com/future/article/20260928-a-virtual-power-plant-hidden-in-vermont-homes-is-keep...
226•devonnull•17h ago•173 comments

America.gov

https://america.gov/
594•plesiv•21h ago•501 comments

U.S. postal inspectors shut down website selling counterfeit postage labels

https://postalemployeenetwork.com/news/2026/09/26/u-s-postal-inspectors-shut-down-website-selling...
238•ilamont•16h ago•139 comments

NASA asked several former SR-71A staffers to help secret restart

https://aviationweek.com/defense/aircraft-propulsion/nasa-asked-several-former-sr-71a-staffers-he...
190•ilamont•1d ago•184 comments

GPT-6.1 Sol replaces GPT-6 Sol after just 7 days, with near-Astra intelligence

https://artificialanalysis.ai/articles/gpt-6-1-sol-replaces-gpt-6-sol-after-just-7-days-with-near...
48•theanonymousone•1h ago•50 comments

Testing WebGPU data layouts with Facet

https://www.mattkeeter.com/blog/2026-08-23-wgpu-facet/
51•luu•1d ago•3 comments

Show HN: Real-time Solar System with 526k asteroids and all tracked satellites

https://space.bl2.net/
273•wanick•16h ago•64 comments

Floppy Emu Hardware Failure Analysis Results

https://www.bigmessowires.com/2026/09/29/floppy-emu-hardware-failure-analysis-results/
14•zdw•12h ago•3 comments

How Delhi cut electricity loss from 50 to 5 percent

https://spectrum.ieee.org/delhi-electricity-loss
534•rbanffy•22h ago•296 comments

RSS Feeds for Last.fm

https://lfm.xiffy.nl/
85•Baljhin•8h ago•18 comments

Backblaze drive stats for Q2 2026

https://www.backblaze.com/blog/backblaze-drive-stats-for-q2-2026/
213•HieronymusBosch•21h ago•63 comments

Phyllotaxis: An audio-reactive LED display

https://jagi.studio/posts/phyllotaxis/
302•evakhoury•1d ago•48 comments

GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price

https://openai.com/index/introducing-gpt-6-1-sol/
972•crorella•18h ago•849 comments

Built Dental Scope

https://dental-scope.com/
41•Zeruxe•1d ago•9 comments

Show HN: Using 2D DFT, dithering, etc. to maximize eInk manga image quality

https://github.com/ciromattia/kcc
30•seam_carver•1d ago•4 comments

Needed 1+1, built a functional programming language

https://hereticpleb.vercel.app/blog/needed-one-plus-one/
114•birdculture•19h ago•37 comments

Language models for text classification: From bag-of-words to Jev

https://magazine.sebastianraschka.com/p/classifier-history-and-jev
136•Anon84•1d ago•7 comments

When oil prices spike, where does the money go?

https://theconversation.com/when-oil-prices-spike-where-does-the-money-go-280763
95•thelastgallon•1d ago•82 comments

Ask HN: What are you reading?

308•dan-bailey•21h ago•590 comments

A Staff Engineer's Guide to Inventing Work

https://sujithjay.com/inventing-work
281•amortize•1d ago•57 comments

PS5 Relapse Exploit

https://github.com/ntfargo/Relapse-Exploit
322•therepanic•19h ago•193 comments

Raleigh Vektar Restoration (2020)

https://retromash.com/2020/06/26/raleigh-vektar-restoration-part-1-the-arrival/
20•robin_reala•1d ago•3 comments

NAND-16: a computer built from 277,248 NAND gates

https://somethingbig.ai/computer
148•rossant•2d ago•87 comments

NRC issues first U.S. construction permit for a BWRX-300 small modular reactor

https://www.gevernova.com/news/press-releases/nrc-issues-first-us-construction-permit-bwrx-300-sm...
47•papa-whisky•12h ago•17 comments
Open in hackernews

GPT-6.1 Sol replaces GPT-6 Sol after just 7 days, with near-Astra intelligence

https://artificialanalysis.ai/articles/gpt-6-1-sol-replaces-gpt-6-sol-after-just-7-days-with-near-astra-intelligence
48•theanonymousone•1h ago

Comments

Dinuda•1h ago
After 5.5, I basically don't notice a jump in model performance, other than my usage ending sooner.
zero1009•46m ago
Felt the same until I started using Luna. I feel like I get similar performance, but faster, and my usage lasts so much longer.
tom1337•42m ago
Kinda same but I miss my 5.3 Codex. Thing lasted forever on my $20 subscription and with detailed prompts was able to pretty much implement everything I requested it to do with a acceptable quality.
ModernMech•28m ago
Same. 5.5 got work done then 5.6 was also fine then 6 was maybe not quite as good. Now with 6.1 they are cutting usage and raising prices and introducing ultra fast mode, but things were good enough 5 months ago.
user43928•25m ago
I do.

The results are less buggy, animations are much better.

It can work autonomously for hours and the result is decent most of the time.

That wasn't usually the case with 5.5, which needed more feedback and iterations to get things right.

baq•22m ago
Astra is noticeably smarter than any OpenAI model before it. Sol 6.1 is very noticeably smarter than sol 6 even after half a day of using it (sol 6 was actually terra 6 and opus 5.5 has taken them by a total complete surprise)
moomin•46m ago
This has got to be a panic move from OpenAI, right? They’ve had some bad press lately because from their billing changes, and Anthropic have finally released a fast, relatively cheap Opus with improved written English.
nba456_•44m ago
Didn't they announce the billing changes the same day?
egeozcan•42m ago
Worries of AI going rogue take so much attention that no governance body seems to care about the shady subscriptions and limits business.
semiquaver•33m ago
Do you mean the subsidized subscriptions which allow individuals to pay a tenth of the normal API cost for tokens?
muyuu•16m ago
Dumping is shady af.
entrope•21m ago
What about those do you think are "shady"? Price discrimination in favor of small customers at the expense of large customers is somewhat common; businesses want customers to buy more of their products, but customers are not obliged to buy more if they like their current deals. Having limits on using finite resources seems even easier to justify.
SkyBelow•9m ago
Humans see differences in prices as unfair. The greatest example of this is price gouging during an emergency, but also look at the level of hate that scalpers get.

Using AI to do this, if anything, given the common negative sentiment, is seen as even worse. People think of this as AI using information asymetry to squeeze more out of users, not to cut people a deal.

Taking advantage of information asymmetry is generally looked down upon as well. Look at all the laws we have protecting kids from this. Businesses often don't seem similar protections because those are businesses with big legal teams (and when it is a big legal team vs a small mom and pop store without a single lawyer on payroll, people do start taking issues with it). The power difference between the average company using AI pricing and the average consumer falls pretty solidly in the 'we don't accept this' side of taking advantage of information asymmetry.

I could keep going, but I think these are already plenty enough reasons to why people look at AI price discrimination as not just a bad business practice they don't like, but an immoral/unethical one.

Sure, one can make economical counter arguments, but that's arguing on an orthogonal dimension that simply isn't relevant to where these feelings/thoughts come from.

atkrista•41m ago
I would just LOVE to see all the behind-the-scenes shithousery both companies are employing to one-up the other in this, largely, 2-horse AGI race. Someone should make a mockumentary when all is said and done!
TeMPOraL•38m ago
AGI will make one, about humanity, after we're all gone - "They were so dumb, they just deserved to die".
f6v•31m ago
The sooner the better, brother.
petesergeant•32m ago
> largely 2-horse

The absolute frontier is largely 2-horse, but the rest of the pack is very close behind, which I'm grateful for. Grok, Facebook, and the Chinese vendors are producing excellent models.

Bluestein•26m ago
... and, must be said a plethora of largely unsung, small, unknown "labs", outfits, "researchers" and the like. There is a long tail of smart people having at this. I guess, sheer compute aside, I think much progress - or, at least, important pieces thereof, will come from there.-
bayindirh•25m ago
There are some niche research areas where bog standard machine learning algorithms make miracles. LLM is just the poster child. AI/ML is a much larger and wider research area.
madsgarff•36m ago
It would be very nice if artificialintelligence.ai actually, from a UX perspective, did the models in more than one thinking mode. I use claude, and I wanna build a feeling for what high, medium, etc. actually gives me. So far their comparisons, and having tried several different models for my work, has given me a feel of what 50 intelligence actually is. And I believe it would be be of even greater value to get a feel inside the single model I actually use, as most people do, because not many, I believe, switch heavily between models when working. I understand that the cost here is greater but the model provivders should obviously give you free access, because of the great work you are doing.
nextaccountic•33m ago
> It would be very nice if artificialintelligence.ai actually, from a UX perspective, did the models in more than one thinking mode

But they do. See for example the pareto curve they have, try to locate GPT-6 Luna (max), (xhigh), (high), (medium), (low)

girvo•25m ago
It does!

…but not for all models, which is pretty annoying.

rrr_oh_man•20m ago
> I wanna build a feeling for what high, medium, etc. actually gives me

Nothing, really. It's like oversampling your data set. You usually get a much better overall baseline performance if you use the default setting.

Pythagon•29m ago
Is this a duplicate thread of this? https://news.ycombinator.com/item?id=49896586
max979•29m ago
Wild pace. Guess they found a critical bug or a quick win to push it out so fast. Astra is a high bar.
sscaryterry•23m ago
Its nerfed as fuck. Unusable.
K0balt•11m ago
What is your use case?
sscaryterry•10m ago
Everything. It just stops working, clears goals. It doesn't listen. This concurs: https://marginlab.ai/trackers/codex/
bearjaws•11m ago
What do people get out of being so reductionist?

Every time a new model comes out, people come out in droves "oh I don't notice anything different".

People have been saying this about <currentModel-1> for 2 years now, and the entire state of AI has changed dramatically.

It cannot be that the next AI model isn't better, but also suddenly what they are capable of is on an entirely different level.

ulimn•1m ago
I suspect it's partly because people didn't jump from GPT-3 to GPT-6.1 Sol and partly because SOTA models from the last few(?) months have been able to tackle most of the regular tasks. It means this new model isn't different in that regard from Opus 4.8, if your mental benchmark is that they both are capable of implementing something like a CRUD app.
sosuke•1m ago
I like the fast releases. 6.1 coming so fast after 6.0 means they found some improvement solid enough for a new rollout.

The only time I remember the newest model making an obvious regression was when the first rolled out MoE. Super speed update but each request had less intelligence at hand. We’re way past that now

mapontosevenths•7m ago
Wow, and it only cost me half my usage!

Now I can finally build that half a thing I've had my eye on.

bayindirh•26m ago
Gemini is also pretty nice for researching things. It turns out that having the whole internet indexed and having unlimited access to YouTube is a force multiplier of some kind.

Since Google has their own TPUs, TPS is also pretty high w.r.t. Claude, for example.

thunfischtoast•17m ago
I think Gemini could be a great product, if they didn't decide to stuff it down my throat at every possible occasion.

Recent example: on my e-reader, tapping a word I don't now and clicking "Translate" pulls up the possible translations from a dictionary, a local file just a couple of megabytes big, near instantly, on this tiny processor.

Doing the same on my Android phone starts a Gemini-chat with the prompt "Translate the word x into y". Takes forever, internet access needed, results vary, burns who knows who much energy.

Why? Just why?

selestify•7m ago
So that some team at Google can hit their OKRs for user adoption and get promoted.
lxgr•7m ago
Gemini is unbelievably bad for research in my experience. It hallucinates like it's 2023, doesn't use its own search, makes up fake rationalizations for why it didn't need to etc.

It's baffling that OpenAI managed to get better at web search than the company literally synonymous with web search.

nbardy•28m ago
I think it's weirdly just a choice of deciding to cut releases.

We already know OpenAI has "bel" that is MUCH better than astra and is being used internally

mFixman•23m ago
Any strong enough model with weak enough safeguards can cause an AI Chernobyl event that will make people and governments against AI development and deployment, just like Chernobyl did for nuclear energy.
ChromeUltron•11m ago
tell me "I drank the kool aid" without telling me you drank the kool aid.
mFixman•9m ago
The US government and most large companies drank the kool aid, and they will be the ones blaming Big AI if things go very wrong.
iLoveOncall•21m ago
> We already know OpenAI has "bel" that is MUCH better than astra and is being used internally

You're just believing their own bullshit. There's no indication that this is true except from claims from people working at OpenAI.

If they really had a much more powerful model, it would make absolutely no sense to sit on it.

adamzenith•18m ago
You don't think having a more intelligent model they can use internally that others can't is an advantage?
iLoveOncall•12m ago
No? The top of human engineers are much better than any model would be, so AI models really aren't a big advantage when you're trying to develop anything that is SOTA.
ceejayoz•8m ago
Not every problem is best addressed by a top engineer.

Plenty of the tasks that keep a company running can benefit from good-enough (and better than the competition).

howdareme9•16m ago
its not done training, why would they release a model that hasn't finished training?

besides, we know anthropic are sitting on models too

iLoveOncall•14m ago
> its not done training, why would they release a model that hasn't finished training?

Because clearly they have no problem with releasing newer versions of models even just a week apart.

> besides, we know anthropic are sitting on models too

It is from your crystal ball or from other bullshit you heard from Anthropic employees on Twitter?

We all know Anthropic had Mythos and Fable, and they turned out to be completely normal models, entirely in line with the capability of their predecessors.

All they do is lie, and you're believing their lies.

meowface•9m ago
The poster is not claiming Bel is a secret AGI. Just that it exists and only exists internally at the moment.

It's rumored to be over 10T parameters. When released it'll probably be very good at certain tasks, albeit slow and expensive and not necessarily "wiser". You don't have to make this a binary.

meowface•12m ago
With all due respect, you do not have a single clue what you're talking about.
simonw•7m ago
It makes sense for them to sit on it until they've finished testing it. More powerful but also more likely to delete all your email by mistake = you shouldn't release it yet.
michelb•15m ago
I REALLY need new seasons of 'Silicon Valley'