frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Elevated errors on Claude Opus 5

https://status.claude.com/incidents/mfdtrknpxghq
20•croemer•1h ago

Comments

croemer•1h ago
Error message: API Error: 529 Overloaded. This is a server-side issue

Related: https://news.ycombinator.com/item?id=49066591 https://news.ycombinator.com/item?id=49056194 https://news.ycombinator.com/item?id=49067964

Aldipower•48m ago
I would say "Elected errors _in_ Claude Opus 5" wouldn't be incorrect either.. Opus 5 isn't very reliable for coding and introduces a lot of regressions every single time I use it. Do you have the same experiences?
cbg0•42m ago
Do you not have unit tests, or how does it introduce regressions? You can tell it how to run the test suite in CLAUDE.md
Aldipower•40m ago
I detect the regression already in planning with Opus 5, so I do not let Opus 5 implement anything. But it is a waste of time and tokens! Does planning with Opus 5 works out for you?
cbg0•3m ago
It sounds like you just need to correct the plan it lays out to avoid the regression? I'm just looking to debug with you, not defending the model. I've mostly used Opus 5 for code reviews & bugfixes.
quaheezle•39m ago
Opus 5 tries to modify the unit tests as a cover to its own regressions - thinking its own logic is correct and the test must be wrongly specified
simiones•39m ago
It is a well known fact that projects with unit tests never have regressions.
cbg0•6m ago
I don't understand the need for the snarky comment, LLMs can run the test suite and avoid regressions.
benjiro29•32m ago
Opus 5 isn't very reliable for coding and introduces a lot of regressions every single time I use it.

And GPT 5.6 Sol over engineers just about everything. No LLM is perfect, its about learning the issues with each LLM and figuring out if you can live with it. Knowledge means that you can anticipate if it tries to pull something funny, and harness it against that behavior.

cyanydeez•29m ago
this would be _Great_ advice if you owned your own LLM and your knowledge was trapped in Amber because you were satisfied.

It's horrible advice given what we've seen consistent: changing alignments, changing guardrails, changing system prompts, changing inference priorities, etc.

Anyone who relies on these for their work product is chaining themselves to a matrix multiple of indetermintism.

Aldipower•28m ago
Sure, but I am a long time Opus user 4.5,4.6,4.7,4.8 and I wonder what's wrong with 5?
trentor•30m ago
Maybe it's my harness but I haven't seen it introducing regressions.
egeozcan•25m ago
I don't know what I could be doing differently to you but I found Opus 5 to be more reliable than even myself at times. Maybe your stack is unusual or you have conflicting commands in your prompts vs CLAUDE.md (that really confuses it)? It could be anything but this huge error bar in delivered quality is one of the biggest issues with LLMs.
bearjaws•23m ago
I do not experience any regressions, I don't really notice much difference either.
flaburgan•8m ago
I also found it to forget some obvious cases in a quite simple flow (validate the email address of a user who register), that surprised me a lot. Maybe it's because I got used to Fable? But I am quite sure Opus 4.8 wouldn't have make this mistake. If I had time I would try the same prompt with it to see. Anyway, back on 100% Fable for me.
jarym•3m ago
Getting to grips with each new model does require some tweaking and experimentation. So far I've found Opus 5 to repeatedly pause its work and give me some seemingly randomly invented decisions to make.
ectoloph•6m ago
These bursts of downtime are one of the reasons I end up with multiple smaller subscriptions between providers.

I'd just end up being really annoyed about the downtime if it lands in the middle of a working day.

How is the Bun Rewrite in Rust going?

https://lockwood.dev/ai/2026/07/27/how-is-the-bun-rewrite-in-rust-going.html
135•tomlockwood•1h ago•78 comments

Kimi-K3 Releases on HuggingFace 7/27

https://huggingface.co/moonshotai/Kimi-K3
535•nateb2022•6h ago•246 comments

Elevated errors on Claude Opus 5

https://status.claude.com/incidents/mfdtrknpxghq
22•croemer•1h ago•18 comments

AI companies are shredding rare books

https://xcancel.com/HedgieMarkets/status/2081534588485296565
68•anon373839•29m ago•21 comments

PGSimCity - How PostgreSQL Works

https://nikolays.github.io/PGSimCity/
720•jonbaer•12h ago•68 comments

Libsm64: Mario 64 as a library for use in external game engines

https://github.com/libsm64/libsm64
40•klaussilveira•2h ago•8 comments

Magnolias Are So Old That They're Pollinated by Beetles, Not Bees

https://mymodernmet.com/magnolia-ancient-flowers-beetles/
116•speckx•4d ago•47 comments

Removing React.js from the codebase and adapting Htmx for UI interactivity

https://misago-project.org/t/removing-reactjs-from-the-codebase-and-adapting-htmx-for-ui-interact...
30•Ralfp•3h ago•12 comments

Building a Fast Lock-Free Queue in Modern C++ from Scratch

https://blog.jaysmito.dev/blog/04-fast-lockfree-queues/
23•ibobev•4d ago•3 comments

The Birth of the American 12-string Guitar

https://www.harpguitars.net/history/grunewald/12-string.htm
28•bilegeek•2h ago•12 comments

Show HN: Physically accurate black hole you can put in your room

https://blackhole.plav.in
376•aplavin•3d ago•124 comments

Towards a Theory of Bugs: The Ruliology of the Unexpected

https://writings.stephenwolfram.com/2026/07/towards-a-theory-of-bugs-the-ruliology-of-the-unexpec...
10•nsoonhui•3d ago•0 comments

The Proof Machine (2016)

https://incredible.pm/
4•BenoitP•32m ago•0 comments

Shay Locomotives

https://www.shaylocomotives.com/
23•Rygian•3h ago•4 comments

VLC for Unity now supported on Linux

https://code.videolan.org/videolan/vlc-unity
50•martz•3h ago•14 comments

Elevated errors on Claude Opus 5

https://status.claude.com/incidents/lhqp09kxq7pb
31•flyaway123•4h ago•15 comments

Scriptc by Vercel: TypeScript-to-Native compiler, no JavaScript engine in binary

https://github.com/vercel-labs/scriptc
221•maxloh•14h ago•109 comments

Modern email can be built from borrowed parts

https://en.andros.dev/blog/d7ed8b07/modern-email-can-be-built-from-borrowed-parts/
40•andros•4h ago•10 comments

Worse on Purpose

https://ledger.worseonpurpose.com/brands
14•bookofjoe•32m ago•1 comments

Decker, a platform that builds on the legacy of Hypercard and classic macOS

https://beyondloom.com/decker/
334•tosh•18h ago•76 comments

US citizen charged after GrapheneOS phone wipes during airport search

https://www.techspot.com/news/113236-us-prosecutors-charge-atlanta-man-after-grapheneos-phone.html
929•eecc•14h ago•720 comments

We have proof automation now

https://www.imperialviolet.org/2026/07/26/zstd-lean.html
193•zdw•16h ago•75 comments

Chinese chipmaker shares surge 470%

https://www.bbc.com/news/articles/c9q9w3x9qn2o
111•pingou•3h ago•98 comments

I wanted a clock that never needed setting. Things escalated

https://arstechnica.com/gadgets/2026/07/i-wanted-a-clock-that-never-needed-setting-things-escalated/
139•lee_ars•4d ago•116 comments

Google Chrome Arrives on ARM64 Linux, Widevine DRM Included

https://www.omgubuntu.co.uk/2026/07/chrome-arm64-linux-available
14•twapi•1h ago•5 comments

Introduction to Data-Oriented Design [pdf]

https://www.gamedevs.org/uploads/introduction-to-data-oriented-design.pdf
197•tosh•18h ago•53 comments

Measuring developer productivity with the DX Core 4

https://getdx.com/research/measuring-developer-productivity-with-the-dx-core-4/
30•saikatsg•3d ago•25 comments

How Unix spell ran in 64 kB of RAM

https://blog.codingconfessions.com/p/how-unix-spell-ran-in-64kb-ram
58•donw•4h ago•4 comments

Fonts In Use – Find out where a font is used

https://fontsinuse.com/
87•open_•13h ago•10 comments

Simulate cassette tape audio profiles using FFmpeg

https://github.com/AARomanov1985/Audio-Cassette-Simulation
148•xterminal•16h ago•65 comments