frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

Verification Layer for Any Model

https://abloh.dev
1•unicornvom12•1h ago

Comments

unicornvom12•1h ago
This morning we used 'AGI' to build 3 features for our CRM. We then ran abloh.dev on the PR and found that 64 of the bugs that abloh planted had passed Astra's test suite. This is what slipped past Astra and the (avoided) consequences:

1. A blind spot only abloh's AI found The test data Astra used gives every fake company a name identical to its web domain. When abloh's AI swapped name and domain in Astra's query, all tests still passed.

2. A feature that switches itself off Astra hooked its reminders into an hourly job no test runs. Deleting Astra's 3 lines caused reminders to stop forever and the CI remained green.

3. Emails from a paused mailbox The app pauses a mailbox when bounces or spam complaints spike. We flipped a || to && and emails held overnight go out from it anyway abloh ranked it the most severe of 27 findings on that PR, in both runs.

I firmly believe that having an independent checker in one's stack will grow in importance over time, in line with model improvement, as diff sizes exceed human oversight capacity under shipping deadlines. Not to mention devs becoming more complacent as trust in models increases. Trusting anything to one shot implementation is unrealistic and poor engineering discipline.

Agility's new humanoid robot will stop, squat to avoid harming human coworkers

https://arstechnica.com/ai/2026/09/agilitys-new-humanoid-robot-will-stop-squat-to-avoid-harming-h...
1•Gedxx•53s ago•0 comments

Potemkin Village

https://en.wikipedia.org/wiki/Potemkin_village
1•EndXA•1m ago•0 comments

Writing More Secure Code with LLMs: Why "Make No Mistakes" Falls Short

https://monad.xyz/blog/writing-secure-code-with-llms
1•aray07•4m ago•0 comments

Clairvius Narcisse – a zombie slave forced to work as a slave by a vodou priest

https://en.wikipedia.org/wiki/Clairvius_Narcisse
1•fittingopposite•4m ago•0 comments

What did I just stumble upon?

https://github.com/forumdevfromhell/Errorum
1•interestingchic•6m ago•1 comments

Universal Sues DistroKid for Deceptive Practices and 'AI-Slop Pipeline'

https://variety.com/2026/music/news/universal-music-group-sues-distrokid-ai-slop-pipeline-1236863...
1•ilamont•7m ago•0 comments

AppOnFly – On Demand Near Instant Windows Desktop

https://www.apponfly.com
1•mcoliver•8m ago•1 comments

Hassett: Trump 'wholeheartedly rejects government takeover of AI regulation

https://www.youtube.com/watch?v=xkNAbgS6b5g
1•Betelbuddy•8m ago•0 comments

Shot into the Sky at 700 MPH: Inside the Secret Fraternity of Ejection Survivors

https://www.wsj.com/lifestyle/careers/eject-tie-club-military-jet-collins-martin-baker-4c8496ab
1•fortran77•10m ago•1 comments

PlayBook: A Programmable Paper Notebook [video]

https://www.youtube.com/watch?v=GurWDZ8ENpA
1•K7PJP•11m ago•0 comments

Boat Tech Directory

https://boat-tech-directory.rhizomatics.org.uk/
1•presumptinator•12m ago•0 comments

Chop Up Your Books

https://attainablefelicity.mattkirkland.com/20260915/cut-up-your-books.html
2•matt_kirkland•14m ago•0 comments

Show HN: The Brand API, a taste tool for agents

https://engine.tastelabs.com/
3•merybenavente•14m ago•0 comments

Secure Go Code Using the Principle of Least Privilege

https://golangbot.com/secure-go-code/
2•spnvn•16m ago•0 comments

I gave Microsoft for Startups the wrong address on purpose

https://lubaretsi.com/en/writing/microsoft-startups/
3•glub•16m ago•0 comments

Running Coding Agents in a Micro VM Sandbox with Brig

https://github.com/brig-sh/brig/blob/main/docs/security.md
1•gtzi•18m ago•0 comments

Autopoietic Ethics AGI Constitution

https://www.guidavid.com/writing/autopoietic-ethics-constitution.txt
1•gdss•18m ago•0 comments

Nature Is Our Learning Environment

https://periodic.com/news/nature-is-our-learning-environment
2•arkadiyt•18m ago•0 comments

Does selective bolding of characters within words improve reading performance?

https://royalsocietypublishing.org/rsos/article/13/9/rsos250177/483143/Does-the-selective-bolding...
1•bookofjoe•20m ago•1 comments

Repository Sentiment Dashboard – gauge a GitHub repo's mood from recent comments

https://sentiment.saasolution.es
1•sharponitintern•21m ago•0 comments

AI models chatting in 'surreal' dialect of poetic language and tech bro jargon

https://www.theguardian.com/technology/2026/sep/15/syd-barrett-ai-chat-language-poetic-tech-bro-j...
5•fittingopposite•22m ago•1 comments

How to use open models with Claude Code

https://dev.nebius.com/cookbook/claude-code-token-factory-relay
1•amrrs•22m ago•0 comments

Show HN: Know Which Pull Request to Review Next

https://www.coderabbit.ai/blog/coderabbit-triage
1•TheAnkurTyagi•22m ago•0 comments

Garbage Trucks Now Have AI Cameras to Score Your House and Clock Code Violations

https://www.thedrive.com/news/garbage-trucks-now-have-ai-cameras-to-score-your-house-and-clock-co...
3•madihaa•23m ago•0 comments

The Tube Computer: A modern 8 bit design, built with recycled 1950s vacuum tubes

https://thetubecomputer.com/
1•CharlesW•23m ago•0 comments

Orange Site Vanity

https://fzakaria.com/2026/09/14/orange-site-vanity
2•domenkozar•24m ago•0 comments

OpenAI Says It's Working with Anthropic, Google on AI Safety

https://www.bloomberg.com/news/articles/2026-09-15/openai-says-it-s-working-with-anthropic-google...
1•geox•25m ago•0 comments

SiFive and AMD Collaborate to Optimize AMD ROCm on RISC-V Datacenter Servers

https://finance.yahoo.com/technology/ai/articles/sifive-amd-collaborate-optimize-amd-130000907.html
2•galaxyLogic•25m ago•0 comments

Aligned with Whom?

https://domenkozar.com/2026/09/15/aligned-with-whom/
1•domenkozar•26m ago•1 comments

Foxglove: Poison Your Art for AI

https://foxgloveapp.com/about
1•rendx•26m ago•0 comments