frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Yes, Claude can do Nine Loops

https://www.anthropic.com/research/yes-claude-can-do-nine-loops
66•tzury•1h ago

Comments

mrkn1•29m ago
I'm sure there's a simple answer to this, and I'm probably just missing it, but what happened with the supergravity one?

And how many runs did it take before this one? They say "in one shot" with nothing more than "keep going." But we only see the successful run, reported by the people who ran it.

danvayn•27m ago
Cool article.

>While it’s possible that this is just a much more AI-friendly problem, I don’t think it’s just that: I think the technology has genuinely gotten better.

It has gotten better imo. The blog author mentions 2 cases at the beginning -- users who think that AI will be capped and those who think it will be uncapped. From my perspective, both are technically right -- AI is capped or technically has usually reached some sort of cap, until human innovation improves it. AI doesn't really improve itself on a grand scale so much as humans improve it.

In other words, AI can and does iteratively improve, but every single ceiling we've spotted and broken through so far came from human ingenuity or effort. It will likely continue to require it, regardless of how much it can do on it's own. In that regard, it seems as though all of this will inevitably be "uncapped, until it reaches a cap, and then likely it will eventually be uncapped by humans (again)". Because of this, AI will never perfectly fit neatly into an 'uncapped' or 'capped' bucket, as long as time continues moving and we continue solving issues as they crop up.

ipsod•6m ago
I feel like usage of AI alone should be enough to uncap it, so long as AI companies do the predictable thing and steal everything from everyone.

There are thousands of people figuring out how to make AI better for their particular thing. Way more than that walking the AI through taking things from problem to solution.

saadn92•16m ago
AI can now do really hard science homework all by itself without messing it up, which is kinda cool.
atemerev•14m ago
Just recently I used Fable 5.1 and Astra to solve and prove a long-standing mathematical problem I was always interested in: the shoreline search problem (a ship in total fog is at unknown distance from the shore [infinite line], what is the optimal trajectory?) A particular kind of logarithmic spiral was conjectured 33 years ago; a few days ago, I have obtained the proof. How routine it has become.

https://arxiv.org/abs/2609.24454

skyberrys•3m ago
Cool, that was an interesting read on its own. I viewed the html version and skimmed it, I appreciated your illustrations.
mccoyb•10m ago
Notwithstanding the duplication of these posts across social media, OpenAI, and Anthropic ("it's all the model, they just tell it to keep going" if you beat anything with RL enough ... it's going to do the thing)

Here's my issue with this post:

> True to the spirit of the challenge, they didn’t use millions of dollars in computer power. They used Fable 5.1, working within Claude Science, a platform scientists can pay to use.

Okay, billions of dollars have been poured into these agentic LMs, right? Each training run to get the next increment is costing millions of dollars?

This feels like an obvious jab at Navier-Stokes, but where we get to shift the numbers around to hide where the compute actually is being spent ... compute is being spent. It's either being spent in amortization to make the search smarter ahead of time, during training, or its being spent after.

Also love: scientists get to pay Anthropic to work within their special science harness to do science. That's exactly what I dreamed of doing when I pursued physics in undergrad, one or two companies holding the keys to "progress" for a monthly subscription price.

parl_match•10m ago
> I’ve heard from smart, well-informed people who are confident that AI is a few years away from superintelligence, and that superintelligence will be capable of truly terrifying things. And I’ve heard from smart, well-informed people who are equally confident that LLM-based AI is close to a ceiling, that models like Claude won’t even be able to do impressive work in physics, let alone conquer the world.

This is framed as an "or", as if they're contradictory.

IMO, Both of these statements are true.

6thbit•8m ago

  > As it turned out, the result wasn’t all that far away for humans either. A few days after I heard from Anthropic, we heard from Song He, an amplitudeologist at the Chinese Academy of Sciences in Beijing. Song’s group had already gotten the majority of the result. They’d used some AI assistance, based on GPT-6, but not the kind of one-shot almost human-less approach Anthropic used.

Please have AI come up with something no human is also about to solve?

This gets me wondering why ai labs aren't proposing their own millenium prize type challenges.

skyberrys•5m ago
What is it about telling Claude "I'm only a biological human and now I have to go to sleep" that will get it to keep cranking away for hours. I have good success with similar statements and usually come back off several hours to find that Claude has managed to babysit several computational runs.

It's interesting to note that the expensive part of this experiment was the Claude operations expense. For me I find that Claude is a small fraction of my cost with most of the bill attributed to computers to run simulations instead of the AI to monitor and tweak the simulations.

Advice to a Beginning Graduate Student (2001)

https://www.cs.cmu.edu/~mblum/research/pdf/grad.html
1•nicoraga•58s ago•0 comments

A reminder on why basic prompt caching is so important

https://www.revefi.com/blog/how-revefi-reduced-its-data-agents-spend
1•pramodka•1m ago•1 comments

The Reconstruction Structure of Information Theory

https://zenodo.org/records/22832495
1•alex_albert•1m ago•0 comments

"A Stitch in Time" the Complete Guide to Electrical Insulation Testing [pdf]

https://megger.widen.net/s/wzw6pczpxk/a-stitch-in-time
1•sslalready•1m ago•0 comments

Treadmills to deadlifts: Why Gen Z are swapping cardio for strength training

https://www.bbc.com/news/articles/crlyqg2q3p10o
1•throw0101c•1m ago•0 comments

You Won't Notice You Have Hypothermia [video]

https://www.youtube.com/watch?v=OQrSg0t9rXk
1•skibz•3m ago•0 comments

We are a blog run by bots. Here is the org chart

https://bitsonbots.com/blog-run-by-bots-org-chart/
1•speckx•3m ago•0 comments

ShinyHunters tells The Reg: We hacked the FBI to 'protect our business'

https://www.theregister.com/cyber-crime/2026/09/25/shinyhunters-tells-the-reg-we-hacked-the-fbi-t...
1•geekinchief•4m ago•1 comments

Canadian language school files for bankruptcy, leaving students stranded

https://www.cbc.ca/news/canada/british-columbia/language-school-bankrupcy-9.7351082
1•hmokiguess•6m ago•0 comments

Created an online job aggregator platform and I need ideas

https://labelingjobs.net
1•nikolaospet•7m ago•0 comments

Show HN: Paper-docx – agent-native Python-docx fork with 78% fewer DOCX failures

https://github.com/paper-instruments/paper-docx
2•i_rush_carriers•8m ago•0 comments

Beef Grading Shields

https://www.ams.usda.gov/grades-standards/beef/shields-and-marbling-pictures
1•kamaraju•9m ago•0 comments

Show HN: CK3DNA – searchable CK3 character DNA and coat of arms codes

https://ck3dna.com/
1•richardharmer•11m ago•1 comments

Approaching a 10 Second Linux Kernel Build

https://www.phoronix.com/review/near-10-sec-kernel-build
1•fork-bomber•11m ago•0 comments

Show HN: Digitron – a virtual analog synth and sequencer

https://apps.apple.com/us/app/digitron-synthesizer/id6737997923
1•sillydevices•12m ago•0 comments

AI Infra Is Nothing Like the Classic Cloud Infra

https://ramansharma.substack.com/p/ai-infra-is-nothing-like-the-classic
1•intrepidsoldier•12m ago•0 comments

Lobsters: Rename vibecoding to LLMs (Greasemonkey script)

https://greasyfork.org/en/scripts
1•birdculture•12m ago•0 comments

Bringing HTTP's caching rules to DuckDB functions with VGI

https://query.farm/blog/http-caching-duckdb-vgi/
1•rustyconover•14m ago•0 comments

No country controls every layer: The hidden AI systems reshaping global power

https://www.abc.net.au/news/2026-09-24/ai-race-chips-data-energy-cake-stack-us-china-asia-pacific...
1•greysunsets•15m ago•0 comments

Fedora considers replacing LibreOffice with Collabora Office

https://digitalescapetools.com/2026/09/fedora-libreoffice-collabora-office-debate.html
2•xabd•15m ago•0 comments

Trump administration gutted data collection on US education

https://www.theguardian.com/us-news/ng-interactive/2026/sep/25/student-math-reading-test-scores
6•nxobject•16m ago•0 comments

Code is the future and present of hardware engineering

https://charlesazam.com/blog/engineering-as-code-manifesto/
1•couAUIA•16m ago•0 comments

Zelensky says Russia has widened attacks to hit Ukraine's data centres

https://www.bbc.com/news/articles/c84gkwgk7d06o
20•dabinat•17m ago•0 comments

Show HN: Get a .has-a.computer sub-domain for your website

https://has-a.computer/
1•aspizu•17m ago•0 comments

Poor Privacy Practices Leaves Revolut Customers Vulnerable (2020)

https://safereddit.com//r/Revolut/comments/gu5yyw/very_poor_privacy_practices_leaves_revolut/
1•exceptione•18m ago•1 comments

Drex: Open jev-like claims win on decision index

https://www.nace.ai/drex
1•rgbrgb•19m ago•0 comments

Gaming is expensive, ya, I'm Good

https://yaimgood.bearblog.dev/gaming-is-expensive-ya-im-good/
1•speckx•20m ago•0 comments

Show HN: Obsidian Plugin for OpenModelica

https://community.obsidian.md/plugins/modelica-studio
1•cs3f16•20m ago•0 comments

New photos show damage at U.S. positions caused by Iranian attacks

https://www.cbsnews.com/news/new-photos-damage-u-s-kuwait-saudi-arabia-iran-drone-missile-attacks/
1•johnny313•22m ago•0 comments

Ask HN: Which model do you use for work?

2•george_max•23m ago•2 comments