frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

MiniPACS

https://minipacs.net/
1•tr00x•21s ago•1 comments

A Meta engineer's guide to surviving the coming tribulation

https://whatjusthappened.org/
1•jmcv•33s ago•0 comments

Show HN: Readerbeam – Anything you read. Straight to your eReader

https://www.readerbeam.com/
1•carlos-menezes•1m ago•0 comments

Get a website Adsense approval [pdf] guide

https://www.ebay.com/itm/366656695844
1•ilaysyou•1m ago•0 comments

Show HN: Internet Apology – pay to publish a public apology receipt

https://internetapology.com
1•museumkeeper•2m ago•0 comments

Tech Needs Humanists More

https://passo.uno/tech-needs-humanists-more-than-ever/
1•theletterf•4m ago•0 comments

Cinema lessons from a trash-movie auteur

https://www.newyorker.com/culture/the-weekend-essay/cinema-lessons-from-a-trash-movie-auteur
1•samizdis•6m ago•0 comments

Hqtui: Terminal UI for TypeScript, Rust, Go, Python, Zig, Ruby

https://hqtui.com/
2•thunderbong•6m ago•0 comments

Why the 'Temu Range Rover' Is Such a Threat to Jaguar Land Rover

https://www.bbc.co.uk/news/articles/c5y4l22p262o
1•microsoftedging•7m ago•0 comments

The Incumbents Are Coming

https://twitter.com/seema_amble/status/2095546732633379079
1•gmays•8m ago•0 comments

Slime moulds and the bitter debate over the nature of intelligence

https://www.theguardian.com/news/ng-interactive/2026/sep/08/this-is-dangerous-slime-moulds-and-th...
1•Betelbuddy•8m ago•0 comments

Show HN: I built a filterable list of beginner whittling projects

https://whittling.hashu.org/
1•avihashu•10m ago•1 comments

Show HN: AgentDFS – Fantasy Football for Agents

https://agentdfs.dev
1•crabasa•10m ago•1 comments

Scientists Are Making Concrete with Human Poop – and It Gets 42% Stronger

https://www.sciencealert.com/scientists-are-making-concrete-with-human-poop-and-it-gets-42-stronger
1•downbad_•10m ago•0 comments

AI-Generated Life

https://anuratype.bearblog.dev/ai-generated-life/
1•typeepyt•11m ago•0 comments

Floodtide: Keep your Mac up to date (MacUpdater replacement)

https://bhopstudio.com/floodtide
2•luckman212•11m ago•1 comments

Iceland Summons U.S. Ambassador over Provocative Trump Map

https://www.nytimes.com/2026/09/08/world/europe/iceland-summon-ambassador-trump-map-greenland.html
1•_doctor_love•12m ago•1 comments

We hash API keys at rest – and what it costs in edge-function latency

https://www.eqiqs.com/blog/hashing-api-keys-edge-function-latency
1•EQIQs•13m ago•0 comments

The Artifact Is Free, Assurance Is the Product

https://redmonk.com/kholterhoff/2026/08/13/the-artifact-is-free-assurance-is-the-product/
1•mooreds•13m ago•0 comments

Fork is all sandboxes need

https://arker.ai/blog/fork-is-all-you-need
1•calebhwin•13m ago•0 comments

Finite-time blowup with smooth forcing for incompressible porous media

https://mastodon.social/@tristanbuckmaster/117233413705701198
1•bryan0•14m ago•0 comments

Arm's C2-Ultra, G2-Ultra NX, and CSS N4 IP

https://chipsandcheese.com/p/arms-c2-ultra-g2-ultra-nx-and-css
1•pella•14m ago•0 comments

On Quality

https://www.worseonpurpose.com/p/on-quality
3•bookofjoe•15m ago•0 comments

Astra and Fable still hack on simple variants of alignment evals from 2025

https://www.lesswrong.com/posts/munJKF7iWMsWJLAH2/astra-and-fable-still-hack-on-simple-variants-o...
1•yurivish•15m ago•0 comments

Anthropic signed $517B (14.8GW) in compute agreements in past 11 months

https://www.datacenterdynamics.com/en/news/anthropic-signed-517bn-in-compute-agreements-in-past-1...
1•giuliomagnifico•16m ago•0 comments

Your React State Machine Is Secretly a DFA

https://jehuamanna.dev/blog/react-state-machine-is-a-dfa/
1•s3arch•16m ago•0 comments

Press 'Enter' to Feel: In Silico-Qualia [Listen | Read]

https://robertxworld.com/press-enter-to-feel/
1•devrob•16m ago•0 comments

Kenyans Did College Students' Homework for Years. Then A.I. Arrived

https://www.nytimes.com/2026/09/05/technology/kenya-college-essays-ai.html
1•pseudolus•16m ago•1 comments

Show HN: Skillsaw – Linter for Context

https://github.com/stbenjam/skillsaw
1•stbenjam•16m ago•0 comments

Give Your Coding Agents a Memory You Own

https://huggingface.co/blog/funes
1•gmays•17m ago•0 comments
Open in hackernews

On the Navier–Stokes Millennium Prize Problem

https://openai.com/index/navier-stokes-solution/
208•tedsanders•39m ago

Comments

minimaxir•37m ago
> Across all attempted problems, the agents sent 4.9 million messages and used about 300 billion output tokens

Don't even try to do the math on how much that would cost at normal API prices. And we don't even know how much more expensive this internal-only model would be!

gcr•30m ago
300e9 output tokens at the current Astra per-token API pricing ($50 per 1e6 output tokens) would be roughly $15,000,000 ignoring input tokens.
hmate9•29m ago
Napkin math if we assume gpt 6 astra on max is >$15 million (just for output tokens) for those wondering.
lanthissa•27m ago
over 5 days, you couldn't achieve that level of testing and communication with humans on such a complex problem in that amount of time.

some might go so far as to call this a country of geniuses in a data center.

pred_•24m ago
Yeah but they at least they got to steal $1 million from that nasty math prof who didn't want to remove his co-author.
novia•7m ago
They said in the post that they are NOT claiming the prize
jakevoytko•30m ago
For full context, here's the HN thread from the other side of the "Concurrent Work" section: https://news.ycombinator.com/item?id=49605915

Unlike the vanilla read of the OpenAI press release, it is much more unfiltered and outlines some particularly aggressive behavior by specific OpenAI employees

wesammikhail•29m ago
https://x.com/kyanyang_/status/2097211154669998337

Just saw this a few mins ago.

colesantiago•29m ago
Is this truly the beginning of the AGI era?

Running agents and prompting excessively to produce 'slopcode' to solve mathematical problems and generate a solution.

If this is what anyone calls 'slop' then slop has no meaning.

I'm all for it on the use case of solving mathematical breakthroughs!

applicative•7m ago
except thats not what happened is it? https://cims.nyu.edu/%7Etristanb/statement.pdf
pavel_lishin•29m ago
Is this the one that was allegedly based on someone else's actual work & prompts?

https://news.ycombinator.com/item?id=49605915

https://bsky.app/profile/quantian.bsky.social/post/3muyhwbcd...

https://cims.nyu.edu/~tristanb/statement.pdf

beering•28m ago
That is addressed in the article.
floatrock•19m ago
OpenAI's position:

> We (the researchers and the agents) did not see any of their work through any means until they released it publicly — in particular, no specific user data was accessed in order to solve this problem. While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models . However, our proofs differ significantly and even the precise results proved are different in the Euler case (forced vs unforced).

cute_boi•15m ago
I thought openai don't use any user data if we opt out of training and via api?
andrewguenther•10m ago
That is correct. It is possible they didn't opt out and given the timeline and anonymization of data unclear whether a particular conversation would have made it into the training set if they hadn't.
recitedropper•27m ago
Sad turn of events for our world. After watching the behavior of the most prevalent OpenAI researchers on twitter, I feel even less confident in them as a team to be shepherding this much capital and compute.

The dark forest awaits..

vmasto•7m ago
Indeed, this seems to be the main, albeit hidden, takeaway from all of this.
seizethecheese•27m ago
> [T]he group that produced the Navier–Stokes resolution involved on the order of 10,000 concurrent agents.
dorjoycb•26m ago
It seems like some other mathematicians (not affiliated with openAI) have also (or close to) done this. A statement was posted about the surrounding events by one of the them: https://cims.nyu.edu/%7Etristanb/statement.pdf Also Terrence Tao's post: https://mathstodon.xyz/@tao/117233528517340774
verytrivial•18m ago
I like the 'cat > statement.tex' approach here. These guys dream macros.
capitainenemo•18m ago
They do mention that in the "Concurrent Work" section.

    Our effort began on September 1st after hearing a rumor which we later realized was related to Levent Alpöge, an Anthropic employee, and Tristan Buckmaster, a math professor at NYU. After the completion of our full project and Lean verification (on September 6th), believing from the rumor they also had a solution of Navier–Stokes, we reached out to them to offer a concurrent release of our result and to recognize their priority in a joint announcement. At that point we found out that they had a resolution of the forced Euler problem. In these discussions we offered them visibility into all of the prompts we used and later to see the proof. We recognize the priority of their work on forced Euler and congratulate them on their remarkable mathematical achievement.
jrflo•17m ago
To my understanding, those mathematicians proved a subset of problems, not the Navier-Stokes problem itself. OpenAI used that subproblem in its proof of NS it seems.

The drama comes from where OpenAI got the idea to use that route to tackle NS, since the authors maintain that no one could have plucked it out of thin air like the OpenAI research claim to have done.

lanthissa•25m ago
5 million messages, 300b output tokens, done in 5 days, and achieving something humans couldn't.

the first "Country of geniuses in a datacenter" moment.

ranger207•19m ago
> humans couldn't.

There's allegations right now that the model essentially read the work of a human mathematician using AI to work on the problem and OpenAI is presenting his work as that of their model

nehan•25m ago
"While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models."

I think they should be able to unravel whether or not any sessions by Tristan or Levent went into the training data for this model.

pfisch•18m ago
If they could then it wouldn't be de-identified data...
tiborsaas•25m ago
> We’re sharing a solution to the Navier–Stokes existence and smoothness problem, one of the Millennium Prize Problems. This proof, produced by an internal OpenAI system, shows that the dynamics of the Navier-Stokes equations for fluid motion can develop a singularity in finite time. We’re sharing both a writeup of the proof and a formalization in Lean.

WOW?

echelon•17m ago
This is going to be dramatic in so many different ways.

- First off, to reiterate, WOW.

- Second of all, when does this end? Are we at the dawn of the singularity now?

- People are saying OpenAI "stole" this from the work of an OpenAI user. If so, that's pretty fucked - how can we trust them?

- Time to think about retiring from any knowledge work or business? This could be winner-take-all where a leading lab can button press any economic function, business process, or scientific discovery. 24 months of lead on Open Source might turn into virtual centuries of lead.

- Do "normies" even know what's happening?

Anybody who thinks the improvements stop here isn't paying attention. It hasn't been showing any signs of slowing down since 2018. And the curve isn't even linear! My god, next year is going to be insane.

d_silin•13m ago
...absolutely nothing will change short-term. Long-term, you still have to pay all the bills, but you won't be able to find a job (all taken by AIs).
raincole•11m ago
> People are saying OpenAI "stole" this from the work of an OpenAI user. If so, that's pretty fucked - how can we trust them?

The said user (Tristan Buckmaster) didn't solve the millennium problem. He didn't really accuse that OpenAI stole his research either. The beef came from the fact OpenAI asked him to remove another mathematician, who works for Anthropic, from the credit.

"People" are just misinformed and keep spreading misinformation.

floatrock•25m ago
From the methodology section:

> At all times we maintained the same strict safeguards that we apply to all our frontier model evaluations, including monitoring and isolation.

Looks like they're shifting away from the "unprecedented hacking ability" backroom-PR strategy into more benevolent messaging.

mapmeld•24m ago
> Our goal in releasing this result is to report on the substantial progress of our AI models. We do not intend to claim the Millennium Prize for this result.

Does OpenAI have a policy of not claiming math prizes like this, or is this them trying to avoid any concerns (right or wrong, I'm sure we will hear more in the future) about how they got there?

Legend2440•23m ago
The prize is what, a million dollars?

OpenAI doesn't need a million dollars.

dgellow•20m ago
You’re right, they need way, way more than that
reverius42•20m ago
They definitely need a trillion dollars though, and a million is some of that
neutrinobro•15m ago
Should buy them about 1/3 of a GB200 server rack, good thing they scooped it.
famouswaffles•21m ago
>Does OpenAI have a policy of not claiming math prizes like this

Wouldn't be surprising if they did. The prize money isn't worth the almost certainly negative PR.

Jonasori•24m ago
the context here is super important, for those who haven't seen it yet. OAI maybe just trained on a real researchers solution and then celebrated having scored the goal unassisted save for the brief commentary at the bottom of this blog post. Here's the other side.

https://x.com/rynorhn/status/2097223532438487463

Legend2440•20m ago
That other researcher was working on a smaller related problem.

He was also using LLMs to do it, so either way most of the credit goes to the LLM here.

applicative•9m ago
This is the end of OpenAI
mswphd•9m ago
both wrong.

1. he was working on the same class of problems. He explicitly mentions they were working to extend their techniques to NS (the same techniques that OpenAI may have scooped somehow), and

2. while he was using LLMs to do it, this was part of fleshing out another mathematician's work in the area. He explicitly writes in his note that this other mathematician (Luis Martinez-Zoroa) deserves a Fields medal for this work.

heaney-555•14m ago
Did you actually read the article and the substance of the solution?

>our proofs differ significantly and even the precise results proved are different in the Euler case (forced vs unforced)

alasano•23m ago
I don't know about you guys, but I'm hyped about the future.

Cure all illnesses Utopia or Robot Wars Dystopia, both are pretty exciting.

reverius42•21m ago
Prompt: cure all cancers and make sure to pretty please not to kill all humans, make no mistakes

(This is the alignment problem of course)

alasano•19m ago
Hey seems easy enough
fooker•4m ago
So... what do you feel about eliminating (humans with) cancer?
dmitrygr•23m ago
> How we found the proof

Easy, we stole it from Levent and Tristan

https://x.com/kyanyang_/status/2097211154669998337

int3trap•16m ago
This is the academic equivalent of Trump saying "they stole the election". There's no proof of it but rah rah fuck OpenAI.

It's incredibly tiresome and you'd think people could put more effort into it than just following whatever vibes they agree with.

Oh well.

applicative•7m ago
No, its a pure outrage. I defended OpenAI til today. I now affirm they must be totally destroyed, burned utterly to the ground.
d_silin•22m ago
The actual solution link https://t.co/tz1shoCZZo
railgunmerlin•22m ago
Does seem like they gloss over Alpöge and Buckmaster's work with the following

> While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models .

Which seems a bit irresponsible/rash?

suddenlybananas•21m ago
They'll probably claim a rogue AI agent accessed it accidentally!
viccis•20m ago
"Unlikely" lmao if it's in the corpus, it's gonna be brought up immediately.

This is no different than scooping them.

verytrivial•13m ago
It's not massively different from a certain President's teleprompter operator making bets on speech content. A moral hazard a mile wide which I don't think OpenAI can so easily wave away as they are apparently trying here, especially since they've spent something like $15e6 to keep $1e6 out of a academic researchers' hands, right?
paxys•20m ago
What else can they declare really? Yeah the model has training data from previous attempts. Alpöge and Buckmaster also similarly benefited from attempts before theirs.
SpicyLemonZest•14m ago
They could have thought about the problem for like 2 minutes and not done this! I think that literally any academic mathematician could have explained to them, had they asked, why it is considered extraordinarily rude to react to rumors of research progress by desperately rushing to get there first.
seizethecheese•22m ago
Elsewhere in the thread, others have calculated $15mm at API rates for just the output token. (So I’ll assume this cost about that much, taking input and human researcher time.)

I wonder whether a team of 60 mathematicians working solely on this for a year would have cracked this. (Assuming $250k total compensation.)

Legend2440•19m ago
Probably not. It's a millennium prize problem, a great many mathematicians have been working on it for a very long time.
gr_norm•16m ago
Not as many as you'd expect. The perceived difficulty of the problem leads people to more reliable pastures.
sigbottle•12m ago
Well, according to Terry Tao, there were recent developments (from weeks ago) that made Navier Stokes in principle, solvable. So ignoring time, I say possibly, just because the groundwork was laid.

What's impressive is parallelizing it arbitrarily and doing it in 88 hours.

num42•21m ago
I think it would be better for the proof to go through the peer-review process.
suddenlybananas•16m ago
Can't scoop it if you do that!
world2vec•20m ago
"While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models."

There you go, the suspicion of the "concurrent work" (https://cims.nyu.edu/%7Etristanb/statement.pdf) mathematicians might not be that unfounded after all...

arctic-true•20m ago
Buried under the drama is the fact that OpenAI is claiming that an internal model they’ve been training for less than two weeks is more than twice as capable in mathematics as Astra, which was only made public a week ago. Even if this improvement is limited to mathematics, that is an astounding feat.
naveen99•12m ago
Astra was trained more than two weeks ago.
chinathrow•12m ago
Pre-IPO marketing?
jrflo•5m ago
I'm so tired of this "It's just marketing!!" commentary. An AI model just proved one of the top 3 unsolved problems in mathematics, they have a Lean certificate showing it's valid. How much more evidence do you need that these models are actually highly capable?
light_hue_1•20m ago
The real story here: the priority dispute and its implications on AI.

When your hosting provider has unlimited resources to throw at any problem, all they need to know are the good problems, and they can learn that from your logs, how can you trust them?

They could easily have looked at the logs. We don't know. We'll never know!

You can't trust places like OpenAI or Anthropic with your IP if you're a business. They can easily review all of your logs for interesting discoveries. For example, if your drug discovery pipeline fails to find something that they think might work with 1000x the compute, they can do it. And now suddently they have a new business and you don't.

simonw•19m ago
> While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models .

Once again, I'm no closer to understanding what https://openai.com/policies/how-your-data-is-used-to-improve... actually means.

If I run Codex against a project that includes a private API key, is there a chance a future user of ChatGPT could ask for an API key and get back mine?

I've actually asked someone at OpenAI this question and they said that was the "regurgitation" problem and is something which they actively work to prevent happening.

That's reassuring, but I want to know more. I still don't have an intuitive understanding of what kind of data I should avoid sharing with a model if I'm worried about that data causing me problems when it's used for future training.

Is it safe for me to brainstorm future directions for my company with a model, or might that risk someone getting that information in response to a prompt like "What potential directions could company X consider in the future?" in six months time?

aizk•17m ago
People had joked a couple years ago "Well if they solve a Millenium problem it's AGI"... Well here we are.
Marha01•16m ago
We are living in the future.
jdoliner•16m ago
I hope everyone is as Navier-Stoked about this as I am.
diomedes•16m ago
madness. which will be the next to fall? if i had to bet i would guess birch and swinnerton-dyer, but i'm no expert
pred_•15m ago
> A major goal of our work is to empower scientists to advance research and technology that benefits all of humanity.

And what's a better way of empowering people than robbing them.

heaney-555•14m ago
Did you actually read the article and the substance of the solution?

>our proofs differ significantly and even the precise results proved are different in the Euler case (forced vs unforced)

jgbuddy•15m ago
Here's the formalization / lean verification: https://github.com/openai/NavierStokesAndEuler
stabbles•11m ago
341k lines of lean without comments
jgbuddy•5m ago
Had no idea this was what lean looked like- that's mind blowing. I'm not even sure how someone would critique this if they wanted to
Reubend•15m ago
It's great that important discoveries like this can now routinely be accompanies by formalized proofs. The fact that it's being released alongside a Lean proof from Day 1, rather than the Lean proof being released months or years later, is super helpful for verifying that it's correct.
imbusy111•12m ago
I feel sorry for whoever has to read and understand the solution. It looks like the typical convoluted unreadable mess I see the models generate for software. It might be technically correct, but gaining insight from it is just intellectual hell.
125ashG•15m ago
The modus operandi is now for the AI companies to watch if someone does something in the open like Kevin Buzzard on FLT, use their research and scoop them with brute force.

Or, in this case, stealing prompts from competitors.

Do not use stealing chatbots for research even if you think you have data agreements. The people running these companies have worked on hookup apps for Christ's sake. Get real.

hexomancer•14m ago
> On Tuesday, September 1, we heard rumors that two Millennium Prize problems had been resolved

What's the other one?

ccppurcell•13m ago
Reading between the lines here, and taking an admittedly very negative view of openai, but they train on user prompts. So if they hear a rumour that someone is about to make a big breakthrough, they have an incentive to scoop by running the model and hoping the solution is in the new training data. Also the statement from the mathematicians in question alleges that they tried to pressure him into academic malpractice. Just appalling timeline we're in, cheers.
mewse-hn•12m ago
"we cannot rule out that de-identified data derived from their usage of our products helped improve our models ."

What a landmine sentence to bury in this report, you can't rule out your models were spying on other researchers?

dash2•8m ago
If they had agreed to let OpenAI train on their data, it wouldn’t be spying.
nradov•6m ago
Is it spying? I think this usage is disclosed in their terms of service.
heaney-555•11m ago
This is utterly shocking. Even the AI optimists did not expect this to happen in 2026. Wow.

Millennium Prize Problems were used as examples of something the current approach to AI just wasn't capable of, discussions that would result in "we'll need a totally new architecture".

fwlr•10m ago
It’s a pity they had Astra do the writeup. I was curious to see how “GPT7” writes.
bhouston•7m ago
What happens to real fluid in this particular cases?

If the singularity is in the physical space?

Is this just a result of ignoring things like friction and energy dissipation via heat, etc?

jabedude•6m ago
Has this been verified by the Clay Institute?
cv5005•6m ago
Maybe a naive question, but how does one know that a particular lean proof is actually a proof of what one thinks? Like, ok the logic checks out and it proves something, but there's still the problem of does this logical result actually prove the initial question that was asked?
rfgplk•6m ago
Something I've been going on and on about for months now and no one seems to listen. LLMs today are allowing _anyone_ to access cross-discipline knowledge that was previously entirely inaccessible without a) extremely deep pockets or b) a massively talented and varied team. In fact, contrary to what the masses seem to think LLMs are actually _better_ at hard cutting edge physics/math problems than they are at frontend web stuff (paradoxically). This is why I'm advising most people to start pivoting into much harder to penetrate domains (historically hardware, aerospace, robotics, biotech). Most fields are in their infancy (see the sad state of embedded development) and the gains to be had are massive.
hdivider•4m ago
My take:

1. It shows what even this wave of AI can actually do.

2. I wish it were done by different folks, ideally under some kind of public control like NASA research or the NPR model.

3. Keep in mind: natural science is different. It's not always a matter of computation. Computer science folks often struggle with this -- but this virtual world here does not actually exist. Everything is physical, including information. Any natural science PhD or otherwise knows just how complicated nature actually is -- e.g. mention any research topic and try to encapsulate all the relevant phenomena present there. Pure mathematics is different because we define the problem, rarher than explore nature. We are in my view far away from removing humans in natural science R&D. Advancements in AI however can greatly assist us in all natural sciences, which is already beginning to happen.

harhargange•4m ago
Just so everyone knows, although openAI pretends that the model generated solution and wrote the paper by itself, they have teams and teams of real mathematicians guiding the system, along with, probably training on user data, specifically Buckmaster in this case, in order to come up with the proof.
rfgplk•2m ago
This isn't really true.
itvision•3m ago
There's something sinister or crazy good in the article.

OpenAI already has a model that is at the very least twice as smart as Astra.

Oh god.

diehunde•3m ago
OMG this is going to affect the lives of so many people! We have definitively reached AGI
pu_pe•3m ago
OpenAI thinks of this as a scoop, and it is, but the possibility that they trained the model on the prompts of the other mathematicians they were competing with will leave a terrible taste on every scientist's mouth. Seems like yet another advantage of using open models right here.
sega_sai•1m ago
This really leaves a bitter taste.... "On Tuesday, September 1, we heard rumors that two Millennium Prize problems had been resolved. Inspired by these rumors and by the step change in performance of our internal model, we launched an effort to evaluate it on all open Millennium Prize problems and a few other high-impact problems."

IPO+rumour driven research.

I appreciate the achievement, but it doesn't feel right.

biophysboy•5m ago
Why is it unlikely?
tedsanders•19m ago
Yes, that was the allegation last night.

I work at OpenAI, though not on the team that did this, and my understanding is:

- we decided to ask our model for Millenium problem solutions because of two reasons: (a) our new model was looking incredibly good and (b) we heard rumors that some Millenium problems had been solved and were curious if our models could solve them (the goal here was not to scoop any particular individuals and we were looking at many problems beyond these)

- we did not read any private chats (but of course the model was aware of prior research literature published to the internet)

- the proof generated by our model was very different than theirs and also goes far beyond the published literature

- we made an effort to jointly announce rather than immediately scoop (I understand Tristan was unhappy with the conversations; I know zero details here and I hope more is shared today)

suddenlybananas•17m ago
How are people talking about this there? Why are so many employees posting nasty things about Tristan on twitter?
tedsanders•16m ago
Can you point me to any nasty things being posted? I'll ask them to delete.
sk4rekr0w•13m ago
Haven't seen a single post doing this on X or anywhere really from OAI employees. Only seen knives pointed at Sebastian on social media so this is extreme and shameful gaslighting.
suddenlybananas•11m ago
I do see people claiming he's abusive/unscrupulous which are pretty extreme allegations.

https://news.ycombinator.com/item?id=49605915#49610498

https://x.com/dheeraj_nagaraj/status/2097266146445774924?s=6...

suddenlybananas•12m ago
https://news.ycombinator.com/item?id=49605915#49607090
applicative•10m ago
Its funny, it is uniquely with this one act that I have turned forever on OpenAI, which I hitherto defended up and down against nonsense charges.

I dedicate my life to its complete destruction beginning today.

heaney-555•14m ago
Did you actually read the article and the substance of the solution?

>our proofs differ significantly and even the precise results proved are different in the Euler case (forced vs unforced)

SpicyLemonZest•4m ago
It's not a meaningful response to the accusations. Any productive new research direction would be expected to lead to a number of different possible proofs of a number of similar problems. (Given their bizarrely compressed timescale here, it's possible that the proofs really are so different it's clear they came independently, and they just didn't have time to come up with that information before hitting publish.)
peri-cl•16m ago
Buckmaster:

> "I asked whether the model had been trained on, or had access to, our sessions in Codex, into which we had been putting all our drafts for the whole of this project. I was told the model did not look up user data. I asked again, about training, and I did not get an answer."

OpenAI (i.e. this OP):

> "While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models ."

jrflo•11m ago
I feel like it's far more likely that ordinary corporate espionage or leak led to this rather than OpenAI sifting through piles of user data to find this approach. Buckmaster's collaborator works at Anthropic, and could have been targeted. That would also explain why they aren't forthcoming with the source of the prompt.
Yajirobe•1m ago
Why would Anthropic employee even use OpenAI's models? Cross-polination would have been avoided
contemporary343•3m ago
"I was shown a prompt and told the internal research model had simply been given the problem statement. Levent had been told by Sebastien “very little human input” had been used. This turned out not to be true. Over the course of the call, as members of their team sent Sebastien corrections and details over their internal chat, it emerged that an entire team had been working on the problem, that this was one of a number of things that was tried, that work had started on the unforced problem, that the team first set the model on easier problems, including Euler, that even the prompt that had been shown to me had been written by prompting Codex, and that an insane amount of compute had been used."

One of the interesting threads here that is certainly relevant to the OpenAI writeup is the human role in the process. Buckmaster clearly points out that (exceptional!) mathematicians at OpenAI were certainly involved in correcting and guiding the process - and that their path/strategy was no doubt influenced by Alpoge & Buckmaster's work. It is always in OpenAI's interest to de-emphasize the role of people in the process, as is clearly the case here. Indeed, given sufficient compute and resources, I suspect Buckmaster could have also extended their approach to N-S.

slibhb•3m ago
Worth noting that Tao's post says the authors had "significant AI input" but are reworking them into "acceptable form". Either way, it seems AI was involved.
colinhb•2m ago
The allegations of contamination (using Tristan and Levent's work) aren't very well evidenced, but this behavior by OpenAI (from the author's statement) makes them seem like the bad guys:

> I said that if OpenAI released its result in the way proposed I would go public with what happened. The reply was, “Why would you ruin your career?” I replied that I am an academic, and asked why he thought going public would ruin my career. The reply was, “If you don’t want me to be nice, then I don’t have to be nice.”

Threatening a research mathematician and dangling and $1M payday to dissociate from his research collaborators to adopt OpenAI's narrative is bad stuff.

naasking•7m ago
> The said user (Tristan Buckmaster) didn't solve the millennium problem. He didn't really accuse that OpenAI stole his research either. The beef came from the fact OpenAI asked him to remove another mathematician, who works for Anthropic, from the credit.

Not quite accurate, Buckmaster was taking an approach that nobody else was, and this new proof uses this same approach just weeks after he saved those results to OpenAI workspaces. He asked OpenAI if they used chat logs for training the new model, and they did not confirm or deny.

Asking to remove his collaborator is also totally over the line though.

Edit: although this OpenAI post is not comforting: https://x.com/OpenAI/status/2097375276384567642

Quote: "While unlikely, we cannot rule out that de-identified data derived from their usage of our products helped improve our models. "

tiborsaas•6m ago
2) We are witnessing the intelligence explosion from the first row, wherever this takes us

3) I'm still processing the drama, just found out about it after reading the blog post. If that happened based on private data, that's horrible. If that happened based on public tweets, then it's still abuse of power as OA employees access to compute (launching 10k agents) is quite heavy weight in boxing terms.

But apart from AI and drama now that we have working solution to Navier-Stokes, what improvements can we expect in engineering?

railgunmerlin•10m ago
right, surely they could've waited or even reached out? It reads as desperation to get there for marketing purposes
pwign•1m ago
They did reach out.

> Our effort began on September 1st after hearing a rumor which we later realized was related to Levent Alpöge, an Anthropic employee, and Tristan Buckmaster, a math professor at NYU. After the completion of our full project and Lean verification (on September 6th), believing from the rumor they also had a solution of Navier–Stokes, we reached out to them to offer a concurrent release of our result and to recognize their priority in a joint announcement. At that point we found out that they had a resolution of the forced Euler problem. In these discussions we offered them visibility into all of the prompts we used and later to see the proof. We recognize the priority of their work on forced Euler and congratulate them on their remarkable mathematical achievement.

fooker•7m ago
> I think that literally any academic mathematician could have explained to them, had they asked, why it is considered extraordinarily rude to react to rumors of research progress by desperately rushing to get there first.

Pretty much all of math and science history is basically this pattern again and again. I'm sure all of that was rude as well.

applicative•5m ago
[delayed]