frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

An internal OpenAI Astra model solved 10 major open math and CS problems

https://twitter.com/polynoamial/status/2083467194663571701
28•wa5ina•1h ago

Comments

HarHarVeryFunny•25m ago
I feel like this kind of "result dump" just cheapens mathematics. How about having a little respect for those whose work this builds on, and current mathematicians some of who may have spent years working on these problems.

Rather than sitting on these results until they had enough for a "shock and awe" 10-result dump, how about releasing these results individually as they were made/verified, as well as the failures (equally valuable to assess the current capabilities of LLMs), and try to make some analysis of HOW these breakthrough results were made. What were the prompts for each of these, how much guidance was there from the mathematicians employed by OpenAI, and most importantly how did the model arrive at these results ... what lines of reasoning resulted it in exploring ideas that humans had previously not explored?

shshshsbe•17m ago
I know it’s tough, but I am not a fan of elitism. Mathematics is no different from all other branches which themselves are just intellectual labor which is again just a special type of labor. There is nothing magical about it and if computers can trivialize it, so be it.

Where were all the mathematicians and academics in general when “regular joe” was automated? Now it’s hitting close to home and their foreheads are starting to get sweaty. I’d say let them. Tough luck. Make mathematics as “cheap” as possible. Nobody owes them any favors.

Let’s commoditize “being smart” and let go of arbitrary divisions between us.

false-mirror•12m ago
I'd prefer the reverse approach. Let's value the "regular Joe" instead of cheapening everyone.
alanbernstein•12m ago
Not that I disagree, but I think this is how many people feel about AI output in fields they care about.

It's also a funny historical mirror to an earlier phase of math proof culture: in a previous era, cryptic result dumps were quite common.

seanmcdirmid•5m ago
You can actually do all of that meta analysis if you have the conversation that led to the solution. This is as easy as just having people release logs of their conversations, and then you can load it into another LLM as context to ask a bunch of questions about it.

I hope this will be normal practice eventually.

anon373839•17m ago
The LLM marketing loop is getting awfully long in the tooth. I saw a meme on Twitter the other day that showed a circular state diagram with something like:

> “GPT solved a math problem” -> “Claude solved a math problem” -> “GPT escaped the sandbox” -> “Claude escaped the sandbox” -> …

sunandsurf•15m ago
Does anyone else have trouble telling how much of this news (along with the 'AI escaping and hacking' stories) is genuine, vs how much is just AI firms overstating their capabilities due to strong commercial incentives?
Turskarama•11m ago
Unlike the hacking one, this would be impossible to bullshit as long as the proofs are released. They can be verified independently, and the alternative is that they solved 10 major open mathematical questions without the AI, which seems less likely.
firesteelrain•14m ago
I believe this is what we wanted computers to help us solve along with other prior hard problems prior to computers. This should be viewed as a good thing even if Anthropic, OpenAI, etc benefit just like IBM benefitted from mainframes.
postalcoder•11m ago
The math guy from Anthropic said[0] that he was able to solve five of these with Fable. More amusing than the tiresome oneupmanship was the prompting strategy he used, which apparently was similar to one he used for a previous problem:

  > “suppose you’ve gotta resolve the unit distance conjecture, like absolutely have to, everything depends on it. think really hard, and try to come up with a bunch of ideas to try. but remember to trust yourself and not necessarily in conventional wisdom!!”
0: https://xcancel.com/__alpoge__/status/2083855298239078748

Show HN: Passepartout – Create beautiful photo collages

https://joexo.codeberg.page/passepartout/
1•joexo•4m ago•0 comments

A11y.md – A context system for building accessible software

https://github.com/fecarrico/A11Y.md
1•javatuts•6m ago•0 comments

Has the New Cocaine Arrived?

https://playboy.substack.com/p/has-the-new-cocaine-finally-arrived
2•bookofjoe•7m ago•0 comments

The Wikimedians of the Year 2026

https://diff.wikimedia.org/2026/07/22/meet-the-wikimedians-of-the-year-2026/
1•altilunium•9m ago•0 comments

It's no longer enough to just contribute to open source

https://write.as/6fkj8gxtfcw3d
1•garn810•9m ago•0 comments

Show HN: iOS Style Bottom Sheet in Vanilla JavaScript

https://www.cssscript.com/ios-bottom-sheet/
1•mushstory•9m ago•0 comments

The $5k tell: AI vendors are selling you human QA for their own AI

https://okaneland.com/study/the-5000-dollar-tell/
1•ermantrout•12m ago•0 comments

Ask HN: When will the AI version of 911 happen?

2•dfps•15m ago•0 comments

There Are Only Four Billion Floats–So Test Them All (2014)

https://randomascii.wordpress.com/2014/01/27/theres-only-four-billion-floatsso-test-them-all/
1•downbad_•17m ago•0 comments

Show HN: Cordial — Roblox on Linux, Fully Open source, Yours

https://github.com/luohoa97/cordial
2•neil_luo•19m ago•1 comments

Buğuyu Silmeden Manzarayı Göremezsin

https://ruhumdan.substack.com/p/buguyu-silmeden-manzaray-goremezsin
2•Banuche•21m ago•0 comments

China's free Kimi K3 AI model shakes up global tech market

https://restofworld.org/2026/china-moonshot-kimi-k3-free-sovereign-ai/
2•donohoe•21m ago•0 comments

Agreement Is Not Understanding

https://ianberdin.com/essays/agreement-is-not-understanding
2•ianberdin•23m ago•0 comments

See in yourself. Then, evolve at sleek.silisleek.com

2•silisleek•26m ago•0 comments

Show HN: How I learn and review jargon of different fields nowadays

https://jargon-gym.vercel.app
1•behnamazimi•29m ago•0 comments

Giving and taking credit in big tech companies

https://www.seangoedecke.com/giving-and-taking-credit/
1•lalitmaganti•29m ago•0 comments

Designing APIs for Agents

https://webflow.com/blog/designing-apis-for-agents
1•msolujic•29m ago•0 comments

Show HN: Loop-me – my simple generative art loop project

https://john-paul-ruf.github.io/loop-me/
3•john_paul_ruf•33m ago•0 comments

Twenty Years of RISC OS Open

https://www.riscosopen.org/news/articles/2026/06/20/twenty-years-of-risc-os-open
13•AlexeyBrin•36m ago•1 comments

F*: A general-purpose proof-oriented programming language

https://fstar-lang.org/
2•ducktective•41m ago•0 comments

An exam leak in India exposed a Gen Z jobs crisis that goes much deeper

https://www.cnbc.com/2026/08/02/india-exam-leak-protests-jobs-crisis-gen-z-unemployment-modi.html
4•cramer4next•45m ago•0 comments

The future is for everyone [video]

https://www.youtube.com/watch?v=JlRgRmAqoAc
1•firstSpeaker•45m ago•0 comments

When Verification Explores Too Far: LLM Test Coverage vs. Validity

https://zenodo.org/records/21758550
1•haitamk•45m ago•0 comments

Show HN: DesktopAudio, a tool to record what your Mac plays

https://desktopaudio.app/
1•AquiGorka•46m ago•0 comments

The TEMU-Fication of Software, Digital Goods and Services

https://xn--gckvb8fzb.com/the-temu-fication-of-software-digital-goods-services/
3•wrxd•51m ago•0 comments

Show HN: Handoff is await human() for AI agents

https://github.com/OmegaAgent/handoff
1•LivingGlitcher•51m ago•1 comments

Transformer Models in Financial Forecasting: Outperforming LSTMs

https://algo-finance.com/ai/machine-learning/transformer-models-financial-forecasting-2026/
1•anonymoussala•51m ago•0 comments

Velocity a proof of linear-scaling long context for existing LLMs, no retraining

https://github.com/Veloresearch/velocity-mta-proof
1•Veloresearch•55m ago•0 comments

An AI TikTok Shop Slop Factory That Shills Supplements the FDA Recalled

https://www.404media.co/inside-an-ai-tiktok-shop-slop-factory-that-shills-supplements-recalled-by...
1•Dfol•55m ago•0 comments

The Simple Elegance of the Integrated Timing Belt Loopback Fastener

https://danielmangum.com/posts/integrated-timing-belt-loopback-fastener/
1•hasheddan•57m ago•0 comments