frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Coding AI without deterministic outcomes

https://www.aha.io/engineering/articles/coding-ai-without-deterministic-outcomes
6•FigurativeVoid•55m ago

Comments

justinweiss•28m ago
> Let's say it kind of does what you want, but not exactly. I now have a problem. How do I fix what I told it to do so it does exactly what I want? I do not know. Do I need to be more persuasive? Should I be argumentative? Do I need to be rude? Was I ambiguous in a way that I did not understand as I wrote it?

Yes! This is one of the major challenges I still have. It's easy to say "fix model errors in AGENTS.md / the system prompt." But finding a way to prove the effect of that change is the hard part.

It's so frustrating when a model says "no, your guidance was fine, I just didn't follow it, I'll do better next time..." It just makes me want to yell "No! You won't do better next time! This is the problem!"

> In that respect, working with an LLM is much more like working with another person than working with traditional software.

Yes, but if you work with another person, at least they'll have a chance of remembering suggestions you make.

The simple answer is "Have the AI do a refinement -> eval -> refinement loop." That pushes the problem somewhere else, into the "what goes in the eval" question ("and make sure you don't overfit" -- something AI-driven prompt refinements are not good at).

joshum97•21m ago
> Coding AI is different. By that I mean building features that use an LLM, where much of the programming involves writing prompts to get the output you want.

As soon as you add LLM fuzziness to a codebase, it seems to propagate like a virus. You can’t get anything reliably useful for machines downstream of a prompt—even with structured output & friends there’s always the possibility it’ll be flat-out wrong. Reminds me of the old “function coloring” issue except it’s probabilistic instead of async. Obviously it opens up incredible possibilities but I do miss the days where you could predict exactly what would happen by reading code.

It's not just the f*cking sandbox

https://x.com/joedaroo/article/2104335929293127851
1•bananaflag•1m ago•0 comments

Jeff – Jev-compatible 0.8B decision models, trained at home, ~30 ms

https://github.com/firelex/jeff
2•firelex•2m ago•0 comments

What Improves Developer Productivity at Google? Code Quality [pdf]

https://dl.acm.org/doi/pdf/10.1145/3540250.3558940
1•ibobev•4m ago•0 comments

Best of British Design

https://best-of-british-design.vercel.app/
1•ArisC•4m ago•0 comments

Where's the "Intelligence Explosion"?

https://www.noahpinion.blog/p/wheres-the-intelligence-explosion
1•pretext•4m ago•0 comments

Did Anthropic's A.I. Make a Scientific Discovery on Its Own?

https://www.nytimes.com/2026/09/27/science/anthropic-biology-enzyme-mestre.html
1•socializer•5m ago•0 comments

Deshittification Part 2: Bypassing the App Store Gatekeeper

https://fabricati-diem.inform.social/post/deshittification-as-a-service-part2-bypassing-the-app-s...
1•ibobev•5m ago•0 comments

Space Jam (1996)

https://www.spacejam.com/1996/
1•rizsyed1•5m ago•0 comments

How I made $10k in two days writing ebooks with ChatGPT

https://notes.hella.cheap/how-i-made-10k-in-two-days-writing-ebooks-with-chatgpt.html
1•ibobev•6m ago•0 comments

World Labs Is Joining AMD

https://www.worldlabs.ai/blog/amd-announcement
1•mfiguiere•7m ago•1 comments

AMD to Buy Fei-Fei Li's World Labs AI Startup for $8.2B

https://www.bloomberg.com/news/articles/2026-09-28/amd-to-buy-fei-fei-li-s-world-labs-ai-startup-...
3•forthwall•8m ago•3 comments

World Labs is joining AMD

https://drfeifei.substack.com/p/worldlabs-joining-amd
2•tomrod•9m ago•1 comments

Your SBOM Is Fan Fiction

https://yeet.cx/blog/your-sbom-is-fan-fiction
1•zasc•9m ago•0 comments

Rewriting Tailwind CSS in OCaml: Is It (Pixel-)Correct?

https://gazagnaire.org/blog/2026-09-23-tailwind-parity.html
1•beckford•10m ago•0 comments

A mysterious organ in the sperm whale's head gives it the power to sink ships

https://www.bbc.com/future/article/20260923-the-science-behind-moby-dick-the-prized-head-of-the-s...
1•rmason•12m ago•1 comments

Unsurprisingly, Meta's new Muse AI agent blatantly ignores users permissions

https://appleinsider.com/articles/26/09/28/metas-new-ai-agent-blatantly-ignores-users-permissions
2•cdrnsf•12m ago•0 comments

AMD acquiring Fei-Fei Li's World Labs AI firm in deal worth $8.2B

https://www.cnbc.com/2026/09/28/amd-fei-fei-li-world-labs.html
5•cramer4next•12m ago•1 comments

FBI Hackers Say They Won't Publish Trove of FBI Employee Data

https://www.404media.co/fbi-hackers-say-they-wont-publish-massive-trove-of-fbi-employee-data/
1•cdrnsf•12m ago•0 comments

Developer Coma Time

https://dct.novian.works
1•reverseblade2•15m ago•1 comments

Nowhere in Particular

https://nowhere-in-particular.pages.dev/
1•srid•15m ago•0 comments

How to Know When the AI Boom Is About to Go Bust

https://www.wsj.com/finance/stocks/how-to-know-when-the-ai-boom-is-about-to-go-bust-61af3d26
2•tolugenius•15m ago•1 comments

Ask HN: Feedback on GitHub-compatible Git hosting

1•rohityadavcloud•16m ago•0 comments

AMD is acquiring Fei-Fei Li's World Labs in an $8.2B, all-stock deal

https://twitter.com/EdLudlow/status/2104664216393420843
1•htrp•16m ago•1 comments

Data Parallel Pretty-Printing

https://futhark-lang.org/blog/2026-06-18-data-parallel-pretty-printing.html
1•LiljeHDue•16m ago•0 comments

Nvidia launched a tool designed to stop AI agents from going rogue. How it works

https://www.businessinsider.com/nvidia-launches-open-agent-safety-platform-ai-going-rogue-2026-9
2•rmason•18m ago•1 comments

Virus Stole a Human Gene and Won't Let Go of It

https://www.nytimes.com/2026/09/28/science/virus-molluscum-human-gene.html
1•gumby•18m ago•0 comments

The Netherlands is rolling its own NixOS-based software stack after U.S.

https://www.tomshardware.com/software/the-netherlands-is-rolling-alternative-nixos-based-software...
1•sbulaev•18m ago•0 comments

Artificial Intelligence Incident Database

https://incidentdatabase.ai/
1•gurjeet•19m ago•0 comments

Electric Cheaper to Operate Than Diesel for Almost Half of Trucks Sold in EU

https://cleantechnica.com/2026/09/28/electric-trucks-cheaper-to-operate-than-diesel-for-almost-ha...
2•gumby•21m ago•1 comments

China's AI Ecosystem: A Background Explainer

https://oxfordchinapolicylab.org/research/china-s-ai-ecosystem-a-background-explainer
1•herbertl•22m ago•0 comments