frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Spin audit of SQD/QSCI quantum-chemistry benchmarks on iron–sulfur clusters

https://zenodo.org/records/21359923
3•purestatelabs•47m ago
Author here.

Background, for anyone who hasn't followed this fight: IBM's iron–sulfur SQD results (Sci. Adv. 2025) are one of the flagship "quantum computers are useful for chemistry now" claims, and a published critique (arXiv:2501.07231) argues the quantum samples never beat classical selected-CI at matched cost. That argument is still live.

Both sides have been arguing about energies. Neither measured which electronic state these calculations actually converge to.

So I measured it. ⟨S²⟩ comes out between 4.7 and 7.0 depending on the system and the subspace size, where the target these papers name is a singlet at ⟨S²⟩ = 0. Every starting guess I tried lands in the same place, and the error at convergence is about the size of the spin-state ladder itself, which is the physics under dispute.

IBM ships a mitigation for this. Run exactly as shipped, using their driver, their recovery loop and their own spin_square() diagnostic, it moves the ground energy by under a nanohartree while making the subspace 4.00x bigger: 194,481 determinants against 48,600. It makes a singlet representable. It never produces one. Their solver also takes a spin_sq argument that would target the singlet directly, and the default pipeline never sets it.

I filed that narrow part on their tracker yesterday, before posting this: https://github.com/Qiskit/qiskit-addon-sqd/issues/337. A maintainer answered and closed it the same day, and his answer is the useful part. The flag, in his words, augments the sampled subspace by using alpha CI strings as beta strings and vice versa. That is a statement about which determinants span the space, not about what spin the returned state comes out in, which is exactly what the measurement says. He did not contest the numbers. The second question, whether the default path is meant to reach spin_sq at all, is still unanswered.

The obvious objection is that this is all my own reimplementation, so here's the part that isn't. IBM's data-availability archive for the flagship paper contains the raw hardware measurement records: 2,457,600 shots on [2Fe-2S] from December 2023, 3,163,742 outcomes on [4Fe-4S] from April 2024, plus the integrals and the optimized circuit's parameters. Running their shots through their own pipeline, [2Fe-2S] reaches low ⟨S²⟩ but sits 248 mHa off their own reference, and [4Fe-4S] converges to a spin-pure triplet, Var(S²) = 3e-6, a genuine S=1 eigenstate: a clean state, and the wrong one, 1,438 mHa from their reference.

Two things in that archive need no analysis from me at all. Their uniform-random null control matches or beats the hardware samples in every published [2Fe-2S] comparison. And their largest [4Fe-4S] runs, at subspace dimension 10⁸, trail their own classical HCI file by 149 mHa.

On the AI angle, since that's half of why this is on HN: Claude did the initial audit end to end in about 72 hours under my direction, and it's been through many rounds of adversarial review since. The part I'd defend as actually interesting isn't the speed. It caught five defects in its own work through pre-registered validation gates, one of them by cross-checking against IBM's own published energy tables, and it retracted its own strongest pro-quantum finding when the new instrument showed that result was a spin-sector artifact. AUDIT_TRAIL.md has the timeline. REPRO_MAP.md maps every claim to a file and a command that regenerates it.

If you want to kill this, and I mean that, here's how. Exhibit any state in a spin-completed ground manifold of these benchmarks with ⟨S²⟩ < 1. Or show me a quantum-sampled subspace at matched determinant count whose spin-identified energy beats HCI or CIPSI. Or get any shipped-pipeline run on IBM's archived samples to land ⟨S²⟩ < 1 and error under 50 mHa at once, with no spin penalty. The harness is in the archive and I'll publish whatever comes back, including if it's me who's wrong.

Preprint: https://doi.org/10.26434/chemrxiv.15006382/v1

The limitations section is real: comparison dimensions are fixed, the [4Fe-4S] reference is approximate, and the 58.6M "samples" file is deduplicated, so the shot-level distribution of their optimal-circuit numerics isn't publicly auditable. It's short. Worth reading before the hot take.

Reasons I [Gary Marcus] wouldn't count Google out

https://garymarcus.substack.com/p/seven-reasons-i-wouldnt-count-google
1•mdp2021•28s ago•0 comments

Gnome Shell Design Dreams – Tobias Bernard and Jakub Steiner Guadec 2026 [video]

https://www.youtube.com/watch?v=GcFMpc9WMdY
1•pizzaiolo•1m ago•0 comments

Show HN: Claude Code that renders Hebrew/Arabic/Persian in the terminal

https://github.com/noambrand/kivun-terminal-wsl
1•Noambbb•2m ago•0 comments

Fediverse Sovereignty and Patentable Experiences

https://future-secured.com/40904
1•adrianwaj•3m ago•0 comments

Quake: Dawn of the Machine – Launch Trailer [video]

https://www.youtube.com/watch?v=2Uf4PcgE8gA
1•boredemployee•5m ago•0 comments

Hackers Stalked Me by Hijacking a Smartwatch for Kids

https://www.wired.com/story/hackers-stalked-me-by-hijacking-a-smartwatch-for-kids/
2•worik•7m ago•0 comments

Trump again tries to limit US birthright citizenship with new executive orders

https://www.bbc.com/news/articles/cj63966j95yo
1•samaysharma•8m ago•0 comments

US agency ends 39% household cap on local TV station owners

https://www.reuters.com/business/media-telecom/us-agency-votes-end-39-local-tv-station-ownership-...
2•voxadam•8m ago•0 comments

Show HN: Developer Inbox

2•ricky300300•8m ago•0 comments

Show HN: ARF – a record format for AI evaluation runs, with reproducible digests

https://www.korvo.xyz/arf
2•akshay_bhardwaj•12m ago•0 comments

Node Banana

https://github.com/shrimbly/node-banana
2•falloon•12m ago•0 comments

No one believes a True Believer

https://pavlemiha.substack.com/p/no-one-believes-a-true-believer
1•PavleMiha•13m ago•0 comments

3D printable partial Antikythera mechanism

https://www.printables.com/model/284372-antikythera-mechanism
1•jbaber•15m ago•1 comments

Show HN: I rewrote Decap CMS with < $1K in Claude tokens

https://github.com/laikacms/decap-cms
1•afirus•16m ago•0 comments

Starbucks new powder making people sick [video]

https://www.youtube.com/watch?v=8kML1mzUZqQ
1•Bender•23m ago•0 comments

To Understand Language Is to Understand Generalization (2021)

https://evjang.com/2021/12/17/lang-generalization.html
1•rzk•24m ago•0 comments

List of Edible Salts

https://en.wikipedia.org/wiki/List_of_edible_salts
1•Chaseraph•26m ago•0 comments

Anthony Kenny [Philosopher] (1931-2026)

https://dailynous.com/2026/08/06/anthony-kenny-1931-2026/
1•mdp2021•26m ago•0 comments

New TONTOU CPU attack bypasses Spectre v2 fixes, leaks Linux password hashes

https://www.bleepingcomputer.com/news/security/new-tontou-cpu-attack-bypasses-spectre-v2-fixes-le...
4•sbulaev•30m ago•0 comments

Airlines Push Back Against ICE Enforcement at Airports

https://www.wsj.com/business/airlines/airlines-push-back-against-ice-enforcement-at-airports-50a4...
1•petethomas•31m ago•0 comments

Someone Stole $15,000 And Nobody Cared

https://michelesteelenow.substack.com/p/the-mahomes-heist-you-dont-know-about
2•jawns•31m ago•0 comments

Show HN: Compare frontier speech-to-speech models for free

https://voiceduel.com/
2•armaanp23•32m ago•0 comments

Prompt Injection Vulnerability in Ollama, Gemma4 and HuggingFace's Transformers

https://www.reddit.com/r/LocalLLaMA/comments/1vhivsb/prompt_injection_in_ollama_gemma4_and/
2•pavelai•34m ago•1 comments

The Challenger Launch Decision

https://press.uchicago.edu/ucp/books/book/chicago/C/bo22781921.html
2•gregsadetsky•35m ago•1 comments

Intuitize – The Thoughtful AI

https://intuitize.pages.dev/
1•telui•36m ago•0 comments

California Man Is Kidnapped Then Executed in Front of Police

https://www.nytimes.com/2026/07/31/us/kidnapping-shooting-chino-hills.html
4•Michelangelo11•37m ago•0 comments

The energy use of agentic AI

https://www.theclimatebrink.com/p/the-real-energy-use-of-agentic-ai
1•alphabetatango•38m ago•0 comments

An Agentic IDE That Builds Itself

https://www.sawyerhood.com/blog/an-agentic-ide-that-builds-itself
5•sawyerjhood•38m ago•0 comments

Attacks your login screen should be terrified of

https://fusionauth.io/lp/10-attacks-mfa
1•mooreds•38m ago•0 comments

The Pentagram Map

https://e-infinity.space/pentagram-map/
1•mathgenius•38m ago•0 comments