frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Ask HN: Is anybody producing good code with coding agents?

5•ruffrey•59m ago
This is a genuine problem that I hear from senior engineers. I'm looking for a solution.

--

The quality of ai-generated code is <censored> (claude, agy, copilot, codex a little better).

It's exhausting to read.

I used to love learning from my experienced colleagues and taking pride in what we made. We spent time on elegance and craftsmanship.

Now I spend nearly the whole workday slogging through convoluted code riddled with footguns. I ride on hopes and dreams I might understand a changeset.

It takes 5x longer to review Claude merge requests and I barely understand what I approve.

Each day I drift further from understanding as I shovel the same <censored> into the codebase.

The solution seems to be, reach "level 4 autonomy" and don't read or write code anymore.

--

WHO HAS SOLVED THIS PROBLEM? PLEASE HELP!

Comments

verdverm•47m ago
I'm really happy with my opencode + open weight setup, the code is generally pretty good, but I do spend tokens having agents go look for common ai slop patterns.

It's heavily customized, replaced most internal systems via plugins, a set of custom agent instead of builtin ones, different model families for different sub tasks. (don't have claude review its own code)

I'm working on polishing them up and porting a few more from my own harness, then will be open sourcing. Keep your eye out for a "better-opencode" plugin suite, I'll be sure to share it with HN :]

In the near-term, I really like GLM 5.3 prose for code explore / review, give it a shot, it's cheaper (flash model) and catches all sorts of mistakes from the Big Ai models. We fully rolled out our custom pr-review on glm-5.3-flash last week. Most devs are still on claude, moving them towards fireworks and opencode.

OpenCode Go is a great way to try out open weight models for $10/month

drewg123•41m ago
Its a spectrum. For ai-maintained code (like a gui to visualize performance data), IDGAF what the code looks like. I just let claude or codex go nuts and 100% vibe code.

For code I care about, I audit every single hunk as its produced. I give it extensive style guidelines, and crack down on things like a 20-line essay in a comment.

For mission critical code, I write the code myself and have an agent review it.

drgo•38m ago
The way I have been doing it is to use LLMs to generate the code that I don't want to write: prototypes, tests, benchmarks, isolated,straightforward almost copy-paste code. I still write my own code as before because I enjoy doing that and because trying to understand and fix what an LLM generates and regenerates is harder and more tedious and time consuming than writing the code the way I want to do it in the first place.
tmarice•35m ago
After vibing myself into a corner multiple times on important projects, I now have only two modes: clankermaxx for code I don't really care about (mostly frontend react), and write by hand everything else. Using Django for backend already removes the most of the cruft, and writing by hand also means I actually understand what's going on. Works fine for now.

I'm still on the fence for tests: I don't really like to write them, but the LLM-generated tests are pretty bad, even the frontier models on xhigh thinking. I usually generate them, but I don't really have the confidence they test anything except 1==1. Unfortunately it's hard to justify the time spent on writing them manually.

Readable Regular Expressions for JavaScript/TypeScript, Inspired by Emacs' Rx

https://rahuljuliato.com/posts/emacs-rx-in-typescript
1•ibobev•49s ago•0 comments

Show HN: Outis – Fight AI spam by sending fake "user unknown" bounce emails

https://github.com/dtonon/outis
1•dtonon•1m ago•0 comments

Fulcrum – Writing model built on top of Kimi K3

https://echo.fulcrum.inc/
1•gapo•2m ago•0 comments

Prediction Markets a Destination for Gambling Addicts

https://www.npr.org/2026/10/02/nx-s1-5981420/kalshi-betting-prediction-markets-gambling-addiction
1•waffletower•2m ago•0 comments

Humans Too Nuanced to Predict?

1•Sol_Trevino•4m ago•1 comments

Best app stack of 2026, according to LLMs: Next.js, Supabase, Vercel, Stripe

https://openllmrank.io/blog/ai-app-stack-2026
1•zipvoila•5m ago•0 comments

GitHub repository landing pages now show an accessibility tab, if provided

https://ericwbailey.website/published/github-repository-landing-pages-now-show-an-accessibility-t...
1•ibobev•5m ago•0 comments

A 20-year-long permanent cookie: America.gov and tracking

https://www.biometricupdate.com/202610/america-gov-launches-with-privacy-pledge-as-login-gov-code...
1•paimapi•7m ago•0 comments

App's front end UI is now optional

https://seroter.com/2026/09/30/your-apps-frontend-ui-is-now-optional/
1•richards•8m ago•0 comments

One month coding with GLM 5.3 Flash

https://wagtail.org/blog/one-month-on-glm-53-flash/
2•ThibWeb•8m ago•1 comments

A Eulogy for the Software Engineer

https://davekiss.com/blog/eulogy-for-the-software-engineer/
1•freediver•8m ago•0 comments

OpenAI's wandering AI agents earn it a California subpoena

https://www.theregister.com/ai-and-ml/2026/10/02/openais-wandering-ai-agents-earn-it-a-california...
1•seanhunter•9m ago•0 comments

Show HN: Status lines in the Claude desktop app

https://github.com/NathanAB/statusline-anywhere
1•nastynate•9m ago•0 comments

Announce: Mpx, a cross platform multi-attach syncing terminal multiplexer

https://forum.nim-lang.org/t/14202#86214
1•capocasa•9m ago•0 comments

PewDiePie uncensors Qwen3.5-9B with Heretic

https://www.tomshardware.com/tech-industry/artificial-intelligence/pewdiepie-unveils-uncensored-a...
1•alanwreath•9m ago•0 comments

Americans are rushing to buy hybrid cars. They're turning to Asian brands

https://www.cnn.com/2026/10/02/business/q3-us-car-sales-hybrid
2•sleepyguy•10m ago•0 comments

SQLite converts integers to text two digits at a time

https://twitter.com/TrisH0x2A/status/2105761445095108705
1•tosh•10m ago•0 comments

Muon Tomography: Inspection and monitoring of assets without internal access

https://inspenet.com/en/articles/muon-tomography-inspection-and-monitoring/
1•bookofjoe•11m ago•0 comments

Benchmarks Are Free Now

https://code.dblock.org/2026/08/15/benchmarks-are-free-now.html
1•bucket2015•11m ago•0 comments

Power approval set to delay Oracle's Wisconsin AI datacenter

https://www.theregister.com/on-prem/2026/10/02/power-approval-set-to-delay-oracles-wisconsin-ai-d...
3•Betelbuddy•13m ago•0 comments

America's Largest Producer of Coal Is Operating Under an Expired Permit

https://insideclimatenews.org/news/01102026/abc-coke-operating-under-expired-permit-in-alabama/
2•paimapi•16m ago•0 comments

Nikon 'Re-Reviewing' Winner of Small World in Motion Contest After AI Accusation

https://petapixel.com/2026/10/01/nikon-re-reviewing-winner-of-small-world-in-motion-contest-after...
2•cainxinth•17m ago•0 comments

The Four Horsemen of Agentic Coding

https://distantprovince.substack.com/p/the-four-horsemen-of-agentic-coding
3•haute_cuisine•18m ago•1 comments

Mozilla shutting down Solo AI website creator

https://support.soloist.ai/doc/solo-shutdown-faq
3•nedrylandJP•18m ago•0 comments

Dutch women, girls filmed with Meta glasses on creep sites

https://nltimes.nl/2026/10/01/hundreds-dutch-women-girls-filmed-meta-glasses-creep-sites-lawsuit-...
1•Betelbuddy•19m ago•0 comments

The 7-year-old Nvidia Shield TV is now $100 more expensive thanks to AI

https://arstechnica.com/gadgets/2026/10/the-7-year-old-nvidia-shield-tv-is-now-100-more-expensive...
1•rbanffy•19m ago•0 comments

Ask HN: Why do all LLM interfaces look the same?

2•grendelt•19m ago•2 comments

A functional taxonomy for LLM inference in agentic tasks

https://jeffauriemma.leaflet.pub/3mv6jnffo6k24
1•jdauriemma•19m ago•0 comments

AI Makes Me Sad

https://mondobe.com/ai-makes-me-sad
25•mondobe•19m ago•2 comments

Knowledge platform powered by agents demonstration

https://github.com/rajatrao/ai-knowledge-platform
1•raorajat007•20m ago•0 comments