frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Learning more about Claude's mathematical capabilities

https://www.anthropic.com/research/riemann-zeta
40•tosh•55m ago

Comments

Philpax•49m ago
> Jarred Sumner, an Anthropic staff member (and non-mathematician) prompted Claude to “take a real stab” at the hypothesis itself, leaving the mathematical choices from there up to the model. Initially, Claude generated and tried 650 ideas, none of which worked. Jarred prompted Claude to try again, and it spent a day and a half coordinating about 60 Claude subagents, which this time went much deeper: between them, they ran 2,400 shell commands and wrote hundreds of Python scripts.1 The subagents ran thousands of numerical checks against known zeta zeros and refereed one another’s work. Throughout this process, Jarred's input was mostly limited to sending Claude messages of encouragement (mostly variants of “keep going” or “believe in yourself”).2 This seems to have helped Claude overcome some initial skepticism that it could make meaningful progress.

The world we live in is beyond parody.

astro1234•28m ago
Im curious if you find this to be a parody in a bad way or simply a “the state of the art in math research right now is telling a machine to believe in itself”. I am in the latter camp…
EMIRELADERO•12m ago
The former, because it's anthropomorphizing a model.

Anthropic is especially guilty of this, they have been using such language for a while, like when they analyze model weights for mechanistic interpretability and call it the model's "biology".

It's just distasteful.

Philpax•6m ago
For me, personally, it's that the Bun guy - specifically him, not a mathematician - indirectly progressed the Riemann hypothesis by repeatedly telling a model to ganbatte!

It's a ridiculous position we find ourselves in.

lorenzohess•35m ago
> An unreleased research version of Claude has improved on a longstanding lower bound for the fraction of zeros of the Riemann zeta function that satisfy the Riemann hypothesis. Drawing on extensive prior research by mathematicians over the past decades, it has increased this bound from 41.6% to 67.2%.
rvz•33m ago
Although it took an unsuccessful attempt at it, the progress is as follows:

"Claude found that combining the results from Baluyot, Goldston, Suriajaya, and Turnage-Butterbaugh with the work of Bombieri provides a way to surpass the previous state-of-the-art lower bound proportion of 41.6%, increasing it to 67.2%."

The transcripts, papers, and Claude's explanation are an interesting and a better read than this article, and this is exactly what Anthropic should continue to do and it helps other researchers outside the company as well.

  Claude's paper [0]

  Claude's Formalization [1]

  Anthropic's informal note stating the proof more concisely [2]

  Claude’s explanation of how it arrived at its result; [3]
    
  Detailed transcripts of Claude's process. [4]
[0] https://www-cdn.anthropic.com/564f962e60643842f5fcb4a17c9dbc...

[1] https://github.com/anthropics/zeta-23-lean

[2] https://www-cdn.anthropic.com/23455459f8832d06bb175cc0f88d01...

[3] https://www-cdn.anthropic.com/d7f3ecf1d01392d887f8bc974ca187...

[4] https://www-cdn.anthropic.com/8a0d1add3c637b858a9a181e98c40e...

bspammer•5m ago
[delayed]
tristanj•24m ago
> Throughout this process, Jarred's input was mostly limited to sending Claude messages of encouragement (mostly variants of “keep going” or “believe in yourself”)

He should consider using the PUA plugin. It detects when the AI is trying to give up on a problem and automatically harasses it with "encouragement" until it reaches a solution.

https://github.com/tanweai/pua

behnamoh•17m ago
Since they say that this is from an unreleased research version of Claude, I wonder if at some point Anthropic and OpenAI will start delaying the release of their models intentionally so they can reap the benefits from the models in, for example, mathematics, medicine, physics, and other fields. Just as an example, imagine if your model were capable of proving P = NP, or if your model could cure diseases. Would you release it for free, or would you try to make sure those benefits go directly to your company? From these companies' standpoint, I think they would choose the latter.
coffeeaddict1•16m ago
This is a beyond remarkable achievement. Finding this lower bound within a few days of prompting is absolutely crazy.
kingstnap•9m ago
Lets play over/under on an AI model proving (or counter exampling) the Riemann hypothesis?

I'm not sure what a good mark would be, but considering this result lets put it at 2027-08-10 (One year from today).

briansmith•5m ago
> Two mathematicians at Anthropic studied and validated Claude’s paper, and produced an informal note for experts stating Claude’s proof concisely.

Why hide the names of the people who wrote the second paper? To discourage it from being cited?

Sanders Calls on Tech Giants to Pause Development of Out-of-Control AI

https://www.sanders.senate.gov/press-releases/news-sanders-calls-on-tech-giants-to-pause-developm...
1•ChrisArchitect•51s ago•0 comments

Hk ask: why I canymt post here

1•Akaoutar•1m ago•0 comments

IPython is all you need

https://nathancooper.io/blog/2026-08-10-ipython-is-all-you-need
1•coop57•2m ago•0 comments

Theo Jaffee on Becoming AGI-Pilled

https://www.thenewcritic.com/p/present-at-the-creation
1•ekluger•2m ago•0 comments

I built a private travel journal for iPhone

https://kfragkoulis.com/blog/building-travel-log/
1•konstantinosfr•4m ago•0 comments

I thought an LLM gateway was unnecessary. Then we added a second model provider

https://github.com/maximhq/bifrost/
1•Swapnoneel•5m ago•0 comments

South Korea to launch $3.5B chip fund, speed development of semiconductor hubs

https://www.reuters.com/world/asia-pacific/south-koreas-lee-wants-military-airbase-relocated-by-m...
1•giuliomagnifico•5m ago•0 comments

US court rules Meta, others must face lawsuits over social media addiction

https://www.reuters.com/world/us-appeals-court-allows-thousands-lawsuits-against-social-media-com...
1•1659447091•5m ago•0 comments

Hackers Take over Inflight WiFi

https://bsky.app/profile/acarsdrama.bsky.social/post/3msqkuoawxx2v
1•jaredwiener•9m ago•0 comments

QaDiT: A 160M text-to-audio DiT trained on cosumer GPU's

https://huggingface.co/QuarkML/QaDiT
1•sidharthGN•13m ago•0 comments

Semiconductor Fabrication Basics – Home Chip Lab Tour [video]

https://www.youtube.com/watch?v=TrmqZ0hgAXk
1•binyu•13m ago•0 comments

Young Men Are Being Sold Fake Belonging and Fake Wealth Online

https://www.equimundo.org/the-grift-economy-press-release/
2•DeepLogin•14m ago•0 comments

We're not done with point clouds

https://claytonwramsey.com/blog/mvt/
1•claytonwramsey•14m ago•0 comments

A Mystery Avenger Escalates a War Against Scotland's Parking Police

https://www.wsj.com/world/europe/a-mystery-avenger-escalates-a-war-against-scotlands-parking-poli...
1•bookofjoe•15m ago•1 comments

KitBash Acquires ArtStation and Sketchfab

https://magazine.artstation.com/2026/08/kitbash-acquires-artstation-and-sketchfab/
1•A_D_E_P_T•15m ago•0 comments

A.I. Agents Are Taking Online Courses for Cheating Students

https://www.nytimes.com/2026/08/10/us/ai-cheating-online-degrees.html
1•vinni2•16m ago•0 comments

Deleted UK Government code repositories

https://github.com/uk-gov-mirror/deleted-repo-report
1•robin_reala•17m ago•0 comments

Why Are Rivers So Mathematical?

https://www.quantamagazine.org/why-are-rivers-so-mathematical-20260810/
1•pykello•17m ago•0 comments

Zuckerberg makes a case for unleashing Pandora's box

https://www.theverge.com/tech/977395/meta-mark-zuckerberg-superintelligent-ai-ramble
1•ashurandi•18m ago•2 comments

CVE-2026-65400: Apple macOS Screen Sharing authentication bypass

https://support.apple.com/en-us/148170
2•lasagnafan•19m ago•0 comments

New Projet Bck

1•dobibag•21m ago•0 comments

Kalshification: Gambling on Anything and Everything

https://illegal.solutions/posts/gambling_forever
1•totallygeeky•21m ago•1 comments

Geek Fighter – 2d fighter game

https://geek-fighter.vercel.app/
1•tareqismail•22m ago•0 comments

Rust SIMD on the GPU

https://www.vectorware.com/blog/simd-on-gpu/
2•sagacity•23m ago•0 comments

Over 415 Hours of annotated and segmented robotics data

https://demo.shotwell.ai/demo-padlock
2•aliabd•24m ago•0 comments

Security Vulnerability in Pioneer Rekordbox

https://alphatheta.com/en/information/important-notice-security-vulnerability-in-pro-dj-link/
2•butterknife•24m ago•0 comments

A pure CSS implementation of some sunlight streaming in through the window

https://github.com/jackyzha0/sunlit
2•birdculture•24m ago•0 comments

Flock camera map: How many are in your neighborhood?

https://thehill.com/policy/technology/6016998-flock-camera-map-how-many-are-in-your-neighborhood/
1•SubiculumCode•27m ago•0 comments

Earth beyond six of nine planetary boundaries

https://www.science.org/doi/10.1126/sciadv.adh2458
1•simonebrunozzi•29m ago•0 comments

A C++ toolchain from 357 bytes, in Bazel

https://fzakaria.com/2026/08/01/a-c++-toolchain-from-357-bytes-in-bazel
1•kaycebasques•30m ago•0 comments