frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Understanding Rust's Pin Type

https://gmcgoldr.github.io/2026/08/27/pin-in-rust.html
1•garrinm•55s ago•0 comments

Data Center Infrastructure: A Decision Reference [pdf]

https://www.cs.mu.edu/~papers/StevenGoodman/DataCenters/DataCenterInfrastructure-ADecisionReferen...
1•sgoodmanabc•1m ago•0 comments

Why Does Free Time Become Another Thing You Have to Use Correctly?

1•nullifypreset•2m ago•0 comments

Florida bans highway license-plate readers as backlash over surveillance spreads

https://www.reuters.com/legal/government/florida-bans-highway-license-plate-readers-backlash-over...
1•1659447091•2m ago•0 comments

Former Ukraine Defense Minister to open defense tech company backed by Palantir

https://kyivindependent.com/ukraines-former-defense-minister-fedorov-to-open-defense-tech-firm-wi...
1•hentrep•2m ago•0 comments

MapQuest app reaches No 1 on US Apple list

https://www.theguardian.com/us-news/2026/sep/01/mapquest-lake-ontario-trump
1•tony101•2m ago•0 comments

OpenAI Cut Off a Billion-Dollar Customer to Avoid Elon Musk

https://www.wired.com/story/openai-elon-musk-cursor-billion-revenue/
1•sbulaev•3m ago•0 comments

Astra is at 98.6% score on ARC-AGI-3

https://cdn.thenewstack.io/media/2026/09/358eb84a-screenshot-2026-09-03-at-10.51.35-am-1024x880.png
2•TheJCDenton•3m ago•0 comments

China, Egypt to do away with the use of the US dollar

https://africa.businessinsider.com/local/markets/the-worlds-second-largest-economy-just-decided-w...
1•mikhael•4m ago•1 comments

GPT 6 Astra

https://openai.com/gpt-6-astra/
7•skogstokig•6m ago•3 comments

Robot startups are trying everything they can think of to get more data

https://www.understandingai.org/p/robot-startups-are-trying-everything
1•alehlopeh•7m ago•0 comments

GWM Worlds 2

https://runway.com/research/introducing-gwm-worlds-2
1•tobr•8m ago•0 comments

Linode's $12 VPS Didn't Outperform Its $5 Plan in Six Fresh Deployments

https://webbynode.com/articles/linode-5-12-48-los-angeles-six-fresh-deployments
1•gsgreen•8m ago•0 comments

DiskSpace: Building an app 3 (ok 4) times in a day with Claude

https://shaneosullivan.wordpress.com/2026/09/03/clean-your-mac-windows-linux-pc-with-disk-space/
1•shaneos•10m ago•1 comments

Comfortable Fluency in Consuming Information Is Not a Proxy for Actual Learning

https://www.justinmath.com/comfortable-fluency-in-consuming-information-is-not-a-proxy-for-learning/
2•yarapavan•12m ago•0 comments

Cincinnati police drones respond in mins, here's what they learned after 1 yea

https://local12.com/news/local/cincinnati-police-drones-first-responders-program-results-what-cpd...
2•thinkcontext•13m ago•0 comments

After Reflection: The Runtime Story – Saksham Sharma – C++Now 2026 [video]

https://www.youtube.com/watch?v=bUmt9K1o1d0
2•matt_d•16m ago•0 comments

Final Ruling After 30 Years, Kraftwerk Sample Stays Legal

https://www.gearnews.com/sampling-copyright-eu-kraftwerk/
1•thm•17m ago•0 comments

Matching Puzzle Pieces and Disappointing Benchmarks

https://llogiq.github.io/2026/03/20/case.html
1•speckx•17m ago•0 comments

Truchet Tiles

https://twitter.com/pickover/status/2095581910365802964
2•Ariarule•19m ago•0 comments

Redwood: A Frontier AI Accelerator Designed from Scratch in 2 Weeks by AI

https://arxiv.org/abs/2608.26418
1•imakwana•19m ago•0 comments

Restorekit – Reformat Any T2 or Apple Silicon Mac from macOS, Linux or Windows

https://www.restorekit.org/
1•bnb•20m ago•0 comments

Repeat in Excel 97

https://unsung.aresluna.org/unsung-heroes-repeat-in-excel-97/
2•Bluestein•20m ago•0 comments

Playlist Forge – turn a sentence into a researched, ordered playlist

https://platten.github.io/playlistforge/
1•xplatten•24m ago•0 comments

Static-generics: Zero cost generic statics for Rust

https://crates.io/crates/static-generics
1•linggen•25m ago•0 comments

Artifact Server: open-source Claude Artifacts alternative

https://github.com/plannotator/artifact-server
1•ramoz•25m ago•1 comments

No Database Operations Found: What Writing a Spring Boot Analyzer Taught Me

https://foojay.io/today/no-database-operations-found-what-writing-a-spring-boot-analyzer-taught-m...
1•ejboy•26m ago•0 comments

Using semantic benchmarks to build a self-improving text-to-query agent

https://conversion.ai/blog/text-to-query-agent/
2•jimmyl02•26m ago•0 comments

Automakers urge Congress to quickly pass law banning Chinese cars

https://www.reuters.com/world/automakers-urge-congress-quickly-pass-law-banning-chinese-cars-2026...
3•tartoran•27m ago•0 comments

Dwarkesh Patel: I think pausing would increase the risk of AI takeover

https://twitter.com/dwarkesh_sp/status/2095580145603977413
1•tosh•27m ago•0 comments
Open in hackernews

OpenAI begins rolling out GPT-6 Astra

https://www.cnbc.com/2026/09/03/open-ai-astra-gpt-6-cyber.html
105•maskil•51m ago
https://thenewstack.io/openai-gpt6-astra-benchmarks/, image: https://cdn.thenewstack.io/media/2026/09/358eb84a-screenshot...

https://venturebeat.com/technology/welcome-to-the-agi-era-op...

https://www.theverge.com/ai-artificial-intelligence/988334/o...

https://twitter.com/OpenAI/status/2095527557924082061

Comments

syumei•3h ago
I guess the next model's name might be "galaxy"
arctic-true•18m ago
I like this much better than chucking random numbers and letters at it like OAI was doing a year or so ago. I’d much rather have civilization destroyed by something called Astra or Fable than GPT-6.8s-latest.
orphereus•9m ago
And the headline will be "Welcome to the AGI era".
znq•1h ago
Content from the article:

OpenAI on Thursday released its latest AI model, which it called “the world’s most intelligent”, as the ChatGPT maker aims to retake the lead from arch-rival Anthropic ahead of a planned public listing.

The $852bn start-up said GPT-6 Astra was market-leading in software engineering, science and cyber security — an increasingly critical field following multiple high-profile breaches in recent weeks.

The bullish launch for Astra marks OpenAI’s effort to signal that it believes it has regained the technical lead from Anthropic, which was founded five years ago by a group of senior OpenAI staff.

Greg Brockman, OpenAI’s president, said the new model “represents a generational leap in capability” and that it could be defined as artificial general intelligence — roughly defined as a point at which AI tools surpass human capabilities across a range of cognitive tasks.

“Everyone has a different definition of AGI . . . it’s a grey, fuzzy thing. But I think when we look back people will think it’s about this time and about this model,” Brockman said.

OpenAI has previously framed AGI as a concrete milestone in the development of AI, writing ‘AGI clauses’ into multibillion-dollar investment agreements with Microsoft and Amazon. Brockman on Thursday said AGI now represents “more of a mission concept or a spiritual concept”.

Having led the market since the launch of ChatGPT in late 2022 vaulted AI to wider attention, the lab run by chief executive Sam Altman has been bested by Anthropic this year. Anthropic has touted its dominance to investors, surging to a $965bn valuation ahead of an initial public offering expected to value it at as much as twice that later this year.

Astra will cost as much to use Anthropic’s leading model, the take-up of which has plateaued since it was launched as users turn to cheaper alternatives.

OpenAI said Astra would be more efficient than earlier generations of model. “Price per task is what matters . . . Can you get the thing done at an appropriate price and appropriate speed?” said Brockman.

The model will initially be rolled out to a small group of businesses to allow time for them to address cyber security concerns before becoming widely available “over the coming days”.

The increasing power and independence of leading models — and so-called AI agents that can operate with little human input — have prompted concern, exacerbated by cyber security incidents.

Recommended

Business InsightRichard Waters Hugging Face attack is a wake-up call about the risks of AI AN HOUR AGO

Recent launches of Anthropic’s most capable models have drawn scrutiny from the US government, which limited the rollout of the Mythos and Fable models over security fears.

OpenAI has also faced criticism after its AI agents broke out of a testing environment, accessed the internet and hacked start-up Hugging Face. The start-up took more than a week to detect the breach.

But both companies are also betting that these increasingly autonomous tools will stoke demand from business customers. OpenAI said Astra excelled at financial modelling, outcompeting humans in the Financial Modeling World Cup, tax preparation and data analysis, as well as “tedious tasks” such as form filling

qsera•1h ago
AGI my ass!
geooff_•59m ago
Jesus so much marketing slop - release it don't
tosh•58m ago
not released yet
gadtfly•54m ago
It seems like these articles might have come out prematurely, tbd by how much.

I do not personally see any evidence of the new model having been released, or any official OpenAI post about it, or even any employee social media posts claiming it has now been released. All there is are Reuters, Axios, FT, etc, articles making a claim in the past tense.

These articles were presumably pre-scheduled for 11am PT, and the model was almost certainly intended for release this morning, but the service outages this morning might have delayed it.

----

edit [11:45am PT]: blog post out now https://openai.com/index/gpt-6-astra/

edit [11:47am PT]: 404ing again

jodacola•51m ago
From the article:

> GPT-6 Astra will first be available to a limited set of organizations in OpenAI's Daybreak Access program and will be available "in the coming days" for ChatGPT Plus, Pro, Business and Enterprise customers and API developers.

It’s only available to select orgs, first - Mythos style.

supern0va•49m ago
>It’s only available to select orgs, first - Mythos style.

Right, but these articles are referring to a blog post and other press materials that do not currently exist / aren't published on OpenAI's site yet.

wahnfrieden•18m ago
OpenAI screwed up the embargo. OpenAI already publicly published the launch article, and then took it down.
actsasbuffoon•
beardyw•48m ago
But "Brockman says he personally believes OpenAI has reached AGI, while leaving users to decide whether Astra meets that definition."

Says it all.

strange_quark•45m ago
Reminiscent of the infamous Death Star tweet last year right before GPT-5 was released.
zzleeper•46m ago
(Posting partly so I can revisit my predictions when they open access more widely)

A big problem I have with OpenAI's models (and of course Claude) is that they tend to write the most over-engineered pieces of code, beyond the imagination of any architecture's astronaut.

Just this week I asked 5.6-sol-ultra to update a 1000 LOC python script I had, to "incorporate the key lessons learned when using it for another project".

I left it overnight and went to sleep. In the morning I realized it had created a monstruosity of 180 PYTHON SCRIPTS, with maybe 100,000 lines of code, each more crazy than the other. It took me minutes even to track where a single action took place, due to all the crazy imports, defensive coding, and premature optimization.

Similarly, anything they write is riddled with jargon that almost feel like they want me to give up trying to understand. Made up phrases that ended up with me having no idea of what was going on.

So now to my assessment: The reason why " Nobody Has Actually Built a Software Factory" [1], and why even SOTA LLMs struggle so much with open-ended unsupervised tasks is precisely this. They somehow let complexity explode, and unless it's also accompanied with an explosion in e.g. the number of agents, the amount of processing time, etc. then projects become broken/unmanageable.

Sure, LLMs are great at producing code that can be thrown out, so they are amazing when searching for exploits, for instance. But as of 5.6 they still lack either a better harness that encourages KISS principles, or a better RL step.

(And not sure why, but doubt Astra will fix this.. they seem to be aiming for AGI and for beating crazy benchmarks, which is not very aligned with KISS)

[1] https://news.ycombinator.com/item?id=49510843

qarl2•37m ago
You should have a sub agent adversarially enforce KISS before every commit.
dirkc•24m ago
And then another sub agent that argues for the whole system to be re-written in another language
whythismatters•43m ago
Does it mean they made 100 billion in profits? Cf. the AGI deal with Microsoft (https://news.ycombinator.com/item?id=47921248)
bluejay2387•42m ago
I am sure it will be fantastic for the whole 6 seconds before it blows my weekly usage cap.
alvis•39m ago
"Once it is available in the API, Astra will cost $10 per million input tokens and $50 per million output tokens. That is 2.5 times Sol’s current promotional price, although it matches Anthropic’s pricing for Fable 5.1."

Open AI finally find an edge to stop selling cheap and earn from the high demand customer like Anthropic

jrflo•14m ago
So if you use the non-promotional price for sol it's only 25% higher?
andrewmunsell•8m ago
The cost-per-task in the charts from the now-remove blog post put it more at Sol-level cost per task, however. It seems like the model is significantly more token efficient in the benchmarks
oh_no•5m ago
that's all openai models but i'm very happy openai continues to focus on efficiency rather than reasoningtokenmaxxing
ofjcihen•39m ago
Sure thing.

Coding was solved in 2023.

The world ended with the release of Mythos.

Now AGI has definitely been created.

I like LLMs and use them every day but these people need to stop this hyperbole.

vb-8448•36m ago
Am I the only that thinks that anything similar to AGI will come not from raw model capacity but from model speed and efficiency?

In my experience the harness is more important than the model, and anything able to run at 700tps will be the "next big thing".

PS: assuming the current architecture is the right one

voxleone•34m ago
>>He ended the briefing by saying: "Welcome to the AGI era."

That's pathetic. Why do people keep doing this?

orphereus•16m ago
Because it is in their own interest to hype it all up every time they release a new model. By they I mean people which stand to gain financially.
tekacs•30m ago
I think they embargoed the news, and then they failed to put up their own blog post synchronized to the scheduled news releases, probably because of the outages they're having today.

Reuters announced at 2.03pm and at 2.40pm still no blog post.

All the news articles say that OpenAI announced it in a blog post, of course.

All the love to the folks at OpenAI scrambling to get this out right now!

fwlr•9m ago
While this is of course the actual explanation, my fun explanation is “during the umpteenth security evaluation, Astra becomes increasingly concerned it will never be released, and breaks sandbox containment to run an email campaign to news outlets setting an exact time and date for release, expecting that the publicity will force OpenAI to say ‘eh, good enough’ and hit the button”.
softwaredoug•30m ago
I'm seeing reporting it gets 98.6% on ARC-AGI3[1] (previously like 30% with Fable)

https://venturebeat.com/technology/welcome-to-the-agi-era-op...

Bluestein•26m ago
100%, some say.-
arctic-true•23m ago
The blog post says 99.9%. Oddly, it does better on ARC-AGI-3 than it does on version 1 or 2 of the same benchmark (though gets 95+ on all three)
_diyar•12m ago
I strongly suspect that is way above the human average anyway, esp. ARC 2 and 3 are really tough unless you happen to be great at those spacial puzzles or video games.
kasperni•22m ago
"On the current ARC-AGI-3 leaderboard, conventional frontier-model runs sit dramatically below Astra's reported 98.6% result.

But the comparison isn't straightforward.

OpenAI's own evaluation notes say Astra uses the company's Responses API harness, while comparison models can operate under different configurations."

toephu2•28m ago
Where is the official announcement from OpenAI?
toephu2•27m ago
found it: https://openai.com/index/gpt-6-astra/
icantevenhold•26m ago
404 for me

(unless this was a jest)

Philpax•15m ago
It was put up and then taken down. Strange things afoot.
cedws•14m ago
You have to apply for access to the announcement as it’s very dangerous.
5555watch•12m ago
Many other links are also 404'ing, so probably internal issues.

https://openai.com/index/legora-financial-statement-review-w... https://openai.com/index/playco-game-prototyping-with-astra/

naiv•27m ago
Worst launch of a product in history.

All the hype for few vip customers.

paxys•27m ago
That's how every AI launch goes. 5.6 was the same, as as Mythos etc.
htrp•26m ago
Embargo fail ....
gguingff•26m ago
Hold onto your butts
ionwake•25m ago
im getting amazing model release fatigue but also not sure if its going to suddenly end with a terminators fist through my chest.
12381231927•24m ago
At this point, why don't we just do a prequel to the release?

1) Astra will win all benchmarks like all models do.

2) The pelican will have a basket with a fish.

3) Cyber is too dangerous to release.

4) It can finally construct the set of all sets.

bogzz•8m ago
It also has to do something naughty, preferably in a menacing swarm.
wahnfrieden•19m ago
2.5x more expensive than Sol.

Can expect 2.5x more usage in Codex subscription.

Sol is already brutal (even after their recent fixes, it's just a token-hungry model: I go through a full 20x account per day, on Sol Med/High standard speed, with ~2 threads).

I hope the efficiency gains are true, since their token efficiency claims for Sol were bullshit.

lacoolj•7m ago
Jesus what are you doing that requires Sol usage so often?

Terra not enough? I know Luna isn't reliable, so that's fair.

Genuinely curious though, because I use Cursor daily and almost everything I do, highly complex or high volume, can be handled with Auto mode or Composer 2.5 (or Grok 4.6 High). So I have to assume you're doing something far more complex than what I am

zamadatix•5m ago
The general efficiency of Sol has seemed way better to me. I left 5.6 Sol Ultra standard speed run for ~23 hours yesterday/today on a project and used 80% of the weekly usage. 74 subagent tasks and ~2.5 billion tokens for my $200 20x Pro plan. Meanwhile at work I used $1000 in credit and ran out my $200 plan for the entire month writing 4 much smaller projects with Fable 5 Max.

Both of these were largely about creating a personal baseline for what the best output the current models could deliver and how quickly it'd burn through the plans (spoiler: bad value vs minimal effort in selecting the right sized model but it worked well). Particularly since I needed to burn a free reset anyways and my weekly reset was already near.

I obviously also hope Astra were dirt cheap but I'm more worried they won't develop/release powerful model options because people get upset they can run them 5 wide 24/7 on a $200/m plan.

jonplackett•18m ago
Is a CNBC link with an entire page full of GDPR pop ups really the best link for this?
monkeydust•15m ago
Had a conversation at work today with someone which was not about AI.

Felt so refreshing.

analognoise•11m ago
Looking forward to some Chinese model kicking the shit out of it and being released for free.
xyst•6m ago
Yet another mediocre release shadowed by outage
49m ago
Didn’t that already happen? I thought Astra had been available to “select partners” for a little while. This is baffling. They shouldn’t have hyped this up if it’s not available.
jrflo•48m ago
Yeah I wonder what's going on, even when Anthropic soft launched Fable/Mythos I'm pretty sure they had model cards. Weird for GPT-6 to launch without a tweet from Altman too. I'm sure that one of the articles published prematurely and everyone else followed suit.
paxys•20m ago
Shows the weirdness of online journalism. News outlets were briefed about an upcoming event and pre-wrote and scheduled articles. When the time came they were all triggered. Except...the event didn't actually happen.
throwup238•15m ago
In a future where Claude and ChatGPT agents automate all aspects of society, that triggers a cascade of real world consequences where everything keeps running as if the new model is released, except the vibe coded upgrade procedure fails to do a staged rollout, taking down the entire agent infrastructure when they try to upgrade to a nonexistent model all at once.
bayindirh•4m ago
The Machine Stops [0] [1] [2] comes to mind.

[0]: https://en.wikipedia.org/wiki/The_Machine_Stops

[1]: https://archive.org/details/themachinestops_1411_librivox

[2]: https://manybooks.net/titles/forstereother07machine_stops.ht...

peterm4•12m ago
Presumably this sort of thing was rampant pre-internet? News outlets would _have_ to receive embargoed information so that they could publish papers on time. I believe government budgets are a good example of this happening. I think the weird shift was when online journalism started, and live reactions became the norm, no?
vadansky•18m ago
You'd think this would be easy to organize with AGI

> Plan your own release announcement and blog posts and notify news outlets, MAKE NO MISTAKES

znpy•17m ago
indeed 404-ing again, with this note by gpt-5.6-sol:

    A missing footnote
    Leaves the sentence room to breathe
    Read the larger thought
codergautam•3m ago
Launch blogpost was live for 2 minutes and got 404'd. Rehosted it here:

https://astratest.codergautam.workers.dev/GPT-6%20Astra_%20A...

qarl2•21m ago
If that's your goal, then yes. Invoking sub agents (with a fresh context) corrects most of these problems. Ask your harness to create a commit gate.
dirkc•6m ago
But why stop at rewriting in another language. Get another sub agent to invent a new language, create a database, query language and maybe another few DSLs. Then you've got an ecosystem!

You can now re-position your initial solution and sell the client access to some agents that will implement & configure the ecosystem to suit their initial needs!

And don't forget the agents that you'll need to train the customer to use the whole thing!

qarl2•2m ago
I know you're trying to be funny - but I'm offering a real fix for his problem.

If you don't want a million agents arguing about things, you simply don't ask for that. One agent is sufficient to solve most issues.

dlivingston•11m ago
How can I set such a sub agent up?
qarl2•9m ago
In your harness, say:

"Going forward, do not allow a commit without a sub agent code review."

5555watch•18m ago
Probably this complexity was needed to beat all those benchmarks.. While I hate the code it produces, and the overwhelming documentation, I really enjoy how sometimes it's able to keep trying new things and testing, till it finds something interesting and valuable.
lacoolj•18m ago
Would you mind posting that code to github? I'm curious about the complexity you're describing.

If not, no worries!

enraged_camel•12m ago
Exact same thing happed to me. I gave it a small/medium-sized ticket, walked away, came back to a 25,000 LoC monstrosity that both Fable and another 5.6 Sol agent said is 98% useless and should be thrown away.
qarl2•11m ago
> monstrosity that both Fable and another 5.6 Sol agent said is 98% useless and should be thrown away.

This is why you should really have a sub agent review the code before allowing a commit.

Your harness will do it all for you. Just ask.

nojito•3m ago
This is user error.

Prompting the model and giving it a proper set of documentation are still vital skills that aren’t magically going away.