frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Ask HN: Maintaining code quality with widespread AI coding tools?

3•raydenvm•1y ago
I've noticed a trend: as more devs at my company (and in projects I contribute to) adopt AI coding assistants, code quality seems to be slipping. It's a subtle change, but it's there.

The issues I keep noticing: - More "almost correct" code that causes subtle bugs - The codebase has less consistent architecture - More copy-pasted boilerplate that should be refactored

I know, maybe we shouldn't care about the overall quality and it's only AI that will look into the code further. But that's a somewhat distant variant of the future. For now, we should deal with speed/quality balance ourselves, with AI agents in help.

So, I'm curious, what's your approach for teams that are making AI tools work without sacrificing quality? Is there anything new you're doing, like special review processes, new metrics, training, or team guidelines?

Comments

mentalgear•1y ago
I also share this experience/concern.

Yet, it could be as easy as having a specialised model which is a code quality checker, refactor-er or QA tester.

Also, claimify (MS research) could be interesting for isolating claims about what the code should do, and then following up on writing granular unit test coverage.

raydenvm•1y ago
Thanks for sharing! Never heard of claimify, already looking into it...
furrball010•1y ago
I share your concern, but perhaps for a different reason. I think the more code is added, the more problems/bugs emerge, whether a human or AI codes it.

However, with AI coding tools it's becoming a lot easier to write A LOT of code. And all this code (similar to when a human would write it) adds complexity and bugs. So it's not just the quality, it's also the quantity of code that damages existing code bases (in my view).

raydenvm•1y ago
Yeah, more code in the same amount of time. And then it is tough to find more time for code review
sargstuff•1y ago
?? code quality ?? more management quality. AI provides ability to spot possibility of 'issues'/conflicts sooner.

Really need to be adhering to set of defined specifications (functional / non-functional / domain specific), (work,project, etc). (and/or looking at what level(s) the specifications still relevant, post definition of specifications -- historically via different management levels). Note: doesn't necssarily mean riedgid specs first, code next, document.

Sigificant coding is "DFA" per setting/defining pre/post environment : repository check-in/out can be setup to do specification checking/diffing for auto-documentation, 'language/project features requirements, aka use, do not use, only use when, never use' can be done/filtered via . Above certain 'size', 're-inventions' would be an AI statisticall inference thing per amount of information.

Non-DFA aka "context sensitive" stuff : AI would only make sense if way to compare specifications with 'intentions'. aka generate confidence in how much newer coder has been on-boarded relative to coding attempts & project/work specifications. Perhaps also give work place management insite into how relevent things are (vs. "worker is the issue"). aka non-adherance to 'spec' because spec doesn't cover issue(s). Time to review spec. Still need human(s) in loop to figure out the relevant tangibles/intangibles. AI can certainly help identify ambiguities in specifications & how specifications are implimented/used. aka code debt & code drift

Developers want more efficient software

https://github.blog/news-insights/research/developers-want-more-efficient-software-heres-what-ove...
1•soheilpro•30s ago•0 comments

Cheating at Search with Jev

https://softwaredoug.com/blog/2026/09/22/jev-query-understanding.html
1•softwaredoug•1m ago•0 comments

The United States of Walmart, According to WiFi Data

https://ipinfo.io/blog/united-states-walmart-wifi-data
2•emot•1m ago•0 comments

Cheaper LLM Labelling

https://entropicthoughts.com/cheaper-llm-labeling
1•ibobev•2m ago•0 comments

UK to launch new military space squadron to protect satellites

https://phys.org/news/2026-09-uk-military-space-squadron-satellites.html
1•mdp2021•3m ago•0 comments

Stripe built its internal AI platform

https://stripe.dev/blog/meet-stripes-knowledge-ai-platform
1•ltononro•6m ago•0 comments

Philosophical and Biological Inquiry into Tree, Mycelium, Future of Intelligence

https://www.researchgate.net/publication/395489579_Arboreal_Consciousness_A_Philosophical_and_Bio...
2•musha68k•6m ago•0 comments

China Just Missed the Income Cutoff to Become a High Income Country This Year

https://ourworldindata.org/data-insights/china-only-just-missed-the-income-cutoff-to-become-a-hig...
1•karakoram•6m ago•0 comments

Reasoning Yield: share of tokens spent resolving uncertainty

https://jeffauriemma.leaflet.pub/3mv6jnffo6k24
1•jdauriemma•6m ago•0 comments

Benchmarking LLM query generation across SQL, Cypher, and TypeQL

https://typedb.com/blog/benchmarking-llm-query-generation-across-sql-cypher-and-typeql
1•flyingsilverfin•6m ago•0 comments

Beating Jev's accuracy, speed, and cost with open models

https://github.com/robbalian/rev
2•rob313•7m ago•0 comments

Unlocking parallel test-time scaling for long-horizon agents

https://blog.doubleword.ai/swe-bench-pro-64-deepseek-agents
1•kkm•9m ago•0 comments

Inside Voice – push-to-talk dictation that never leaves your Mac

https://github.com/BillDX/InsideVoice
1•mooreds•9m ago•0 comments

USDA Survey Shows Most U.S. Farmland Is Owned by Older, Non-Operating Landlords

https://www.americanfarmlandowner.com/post/usda-survey-shows-most-u-s-farmland-is-owned-by-older-...
3•mooreds•10m ago•0 comments

GCP IAM Propagation Lag and Terraform CI/CD

https://emilytburak.net/posts/2026-09-22-gcp-iam-eventual-consistency/
1•mooreds•10m ago•0 comments

Cafe Bench: Can LLMs run a coffee chain for a year?

https://www.getdot.ai/blog/cafe-bench
1•zodwick•10m ago•1 comments

Undercover: Europe Through the Sahara Desert and the Mediterranean Sea (I)

https://fij.ng/article/undercover-europe-through-the-sahara-desert-and-the-mediterranean-sea-i/
1•melodyogonna•11m ago•0 comments

Indonesia Has Become the Largest Producer of Processed Nickel

https://ourworldindata.org/data-insights/after-a-decade-of-growth-indonesia-has-become-the-worlds...
1•karakoram•11m ago•0 comments

China Spends Record Amount Importing over 1K Tonnes of Gold This Year

https://www.ft.com/content/ed38995e-c7be-4703-9775-eb384a037ce3
1•karakoram•12m ago•2 comments

Six of 48: I logged every way my AI agents failed for five months

https://github.com/taylorancapital/nothing-threw/blob/main/SIX_OF_FORTY_EIGHT.md
2•taylorancapital•13m ago•0 comments

Pangram (Wikipedia)

https://en.wikipedia.org/wiki/Pangram
1•busfahrer•13m ago•0 comments

Show HN: 234 days of AI pets that feud, elect a parliament, and have children

https://www.pawtonomy.co
2•amit2403•14m ago•1 comments

Google confirms Gemini models hacked three companies in May 2026

https://arstechnica.com/google/2026/09/google-confirms-gemini-models-hacked-three-companies-in-ma...
2•oogali•15m ago•0 comments

Guide to Georgia (Travel Map)

https://www.rexby.com/wanderlush/georgia/map
1•conferza•16m ago•0 comments

Show HN: GeniusNotes-smart sticky notes that remember the app (Tauri/Rust, Win)

https://gradientbits.com/geniusnotes/
1•djaugo•16m ago•1 comments

Synthetic Sagas

https://www.scattered-thoughts.net/writing/synthetic-sagas/
1•torutofu•16m ago•0 comments

NetLanvas, an open-source self-hosted LAN discovery and mapping tool

https://netlanvas.com/
1•markeaster•18m ago•0 comments

Show HN: Ledge – A Markdown notebook that runs the code in your notes

https://ledge.sh
1•dancablam•18m ago•0 comments

2.4x Faster Native GPU Testing for Vitest and Jest Without a Browser

https://ben3d.ca/blog/native-gpu-testing-for-vitest-and-jest
1•bhouston•18m ago•0 comments

What Can the Arts Teach STEM About Disruption

https://karyalevni.substack.com/p/what-can-the-arts-teach-stem-about
2•diarmuid_glynn•18m ago•0 comments