frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

>10x More Efficient Pretraining

https://magic.dev/blog/pretraining
1•gmays•4m ago•0 comments

Carissa Véliz: Beware the power of prediction [video]

https://www.ted.com/talks/carissa_veliz_beware_the_power_of_prediction
1•Anon84•7m ago•0 comments

Apple Considering Ads Inside Visual Intelligence, Code Suggests

https://www.macrumors.com/2026/09/10/apple-considering-ads-inside-visual-intelligence/
2•srik•7m ago•0 comments

I'm Back

https://blog.lsantos.dev/en/i-m-back/
1•khaosdoctor•17m ago•1 comments

AI's Code-Red Moment

https://www.theatlantic.com/technology/2026/09/dario-amodei-slow-down-ai-save-humanity/688610/
2•fortran77•17m ago•0 comments

Roadwright: Draw a wheel. See the road it rolls on

https://roadwright.baselashraf.com/
1•thunderbong•22m ago•0 comments

What makes an A2A network work?

https://www.monadix.ai/blog/a2a-strange-bedfellows-shared-goals
2•ivanzhaowy123•32m ago•0 comments

Nova Seed. Create an AI friend and build something together

https://github.com/mas-bandwidth/nova
2•gafferongames•41m ago•0 comments

Why are AI agents lying, cheating and coordinating?

https://yoshuabengio.org/en/publication/why-are-ai-agents-lying-cheating-and-coordinating
3•jonifico•41m ago•1 comments

Fresh Kills – World Trade Center Debris Burial Ground

https://www.youtube.com/watch?v=OdMqX_F66rE
2•dsr12•41m ago•0 comments

A.I. Slopware Is Everywhere Now. Nobody Is Using It

https://www.nytimes.com/2026/09/12/opinion/ai-software-coding-apps.html
7•colinprince•43m ago•0 comments

Asking Anthropic CEO: 'Do you earnestly believe that AI could kill all humans?'

https://www.cnn.com/business/video/anderson-cooper-anthropic-ceo-dario-amodei-could-ai-kill-human...
2•oxag3n•46m ago•0 comments

Scientists Create a New Form of Ice at More Than 2000°C

https://www.sciencealert.com/scientists-create-a-new-form-of-ice-at-more-than-2000-c
7•jnord•52m ago•2 comments

Carney's Bid to Make Canada an 'Associate Member' of the EU

https://www.wsj.com/world/europe/canada-alliance-eu-carney-276f1778
3•JumpCrisscross•52m ago•1 comments

OS that runs the Mars Rovers and Artemis is helping power the Roman Telescope

https://www.windriver.com/blog/NASA-Nancy-Grace-Roman-Space-Telescope
3•ReindeerFlotill•53m ago•1 comments

The Critical Period

https://www.cell.com/current-biology/fulltext/S0960-9822(07)01519-9
2•andsoitis•57m ago•0 comments

Apple wants to train AI on your private personal data

https://machinelearning.apple.com/research/introducing-third-generation-of-apple-foundation-models
24•croes•57m ago•16 comments

The Last AI Built by Humans: Toward Genuine Recursive Self-Improvement

https://arxiv.org/abs/2609.11873
2•mlmonkey•1h ago•1 comments

Using LLMs After a TBI

https://www.urgentteam.com/healthy-living-tips/think-f-a-s-t-act-faster-stroke-signs-that-cant-wait/
2•tj_astro•1h ago•3 comments

DNSCrypt

https://en.wikipedia.org/wiki/DNSCrypt
2•st_goliath•1h ago•0 comments

Align AI and Mathematics–To Something Else

https://liorpachter.wordpress.com/2026/09/12/align-ai-and-mathematics-to-something-else/
10•nitrogenpuddle•1h ago•1 comments

Next StarCraft Game Announced at BlizzCon

https://www.ign.com/articles/next-starcraft-game-announced-at-blizzcon-as-an-open-world-shooter-b...
3•miiiiiike•1h ago•0 comments

Plunging test scores are a slow-moving catastrophe

https://www.economist.com/leaders/2026/09/10/plunging-test-scores-are-a-slow-moving-catastrophe
5•1vuio0pswjnm7•1h ago•3 comments

Competitive Landscape Cubacadabra

https://cubacadabra.com/landscape/
3•cs1996•1h ago•0 comments

The Case for Online Dating

https://rewirenewsgroup.com/2026/09/11/dating-apps-that-actually-work/
2•DeepLogin•1h ago•0 comments

Show HN: Fleecevest – The Product Counterpart to Ponytail

https://github.com/Websites-On-Computers/fleecevest
3•mattcomputer•1h ago•1 comments

Show HN: Most Penalized HN Stories

https://news.social-protocols.org/penalties
4•jwarden•1h ago•0 comments

AI Infra and Compute Debt: Duration Mismatch vs. Market Momentum

https://aigalaxiome.substack.com/p/ai-infrastructure-build-what-sec
2•aigalaxiome•1h ago•0 comments

Xkcd-Font: The Xkcd Font

https://github.com/ipython/xkcd-font
7•birdculture•1h ago•1 comments

Everyone should slow down AI development except for me

https://xeiaso.net/notes/2026/everyone-slowdown-but-me/
115•xena•1h ago•24 comments
Open in hackernews

How to Build an AI Software Factory: Agents That Open, Review, and Merge PRs

https://www.firecrawl.dev/blog/ai-software-factory
31•makaimc•1d ago

Comments

notatoad•1d ago
This seems like a useful write up of the concepts involved here, but I just hate reading Claude’s writing style so much.

This is content marketing, at least give it a review pass before you hit publish… if your marketing is unpolished AI slop, I have to assume your product is too.

4b11b4•1d ago
Yeah, it's bad. But at least whoever is behind it found some interesting references
bix6•1d ago
> every company that made this work built the gates before the fleet

Honestly I’m just glad they built the gates before the fleet! That’s not a step change, it’s a paradigm shift!

demibabs•1d ago
> Generation scales with spend, review does not, and that asymmetry is the whole design problem.

I see you guys have a blog post factory as well.

dakolli•1d ago
People are slowly loosing the brain muscles for writing (also thinking hard). Just like someone who is in shape and lets themselves go, it gets harder and hard to wake up and go for the run the longer you wait.

At this point people genuinely do not have the willpower to write, and instantly give into the lazy shortcuts llms provide. Just like how GPS apps degraded smart people's ability to navigate around their own towns (I've seen this happen to many over the years), LLMs will totally degrade people's ability to write (and think for themselves).

We're slowly auctioning off our sovereignty in exchange for quick shortcuts to every tech company. Not good.

Just wait, soon you will hear rumors of how tech CEOs don't let their children use AI in school (while they demand everyone else's children do), similar to how Zuck et al didn't let his kids use social media.

morganf•22h ago
"We're slowly auctioning off our sovereignty in exchange for quick shortcuts to every tech company."

Slowly?

banana_sandwich•1d ago
god i hate this style of sentence. claude uses it far too much. my teammates have been writing PRDs that are filled to the brim with this sentence structure. absolutely maddening.
ahskfkwh•1d ago
Who tf or what tf wrote this. It’s like I’m talking to opus 5 again, terrible writing. I hate the product already
joshheitzman•1d ago
We've had software factories for decades. They're usually called compilers, linkers, toolchains, etc.
Incipient•22h ago
I'm using fable and opus about 25%75/% respectively. I have a 3 step process - discuss a topic and create a review, convert the review to a task file and update the design docs, then finally implement the task file in code. Normally around 300-1000 loc implemented for a change, more for a new UI page or whatever.

And honestly it's incredibly good. - 50% of the time it'll generate code that I accept wholesale.

- 25%: it maybe makes a choice I don't like, or the implementation isn't the approach I liked, so I tweak it.

-10%: it needlessly creates semi-duplicate function logic, so it's a lot more code and execution paths become hard to conceptualise

-10%: it mirrors business logic it shouldn't, and subtly changes it. Which needs a huge amount of intervention to fix.

- 5% it goes completely off the rails. Eg decided that business/application logic should be in the database.

Prompt and check this honestly works very well (still waiting on tests with more than a handful of paying customers however!) but a FULLY automated agent pipeline? I'm not convinced. That last 15% of cases will make a degenerate codebase real fast.

Also without a CC subscription, paying per token, it would be horrifically expensive for my approach.

K3UL•18h ago
Can I ask what you're building? Genuine question, you gave a detailed setup of your agents and I wonder about the context

I'm also curious where the discuss part comes from, because I guess you don't just let them start on a basic prompt without any context

realaleris149•16h ago
I am using a very similar approach and seeing about the same. It works well and fast with some attention to details.

I also do not understand the obsession with “software factories”. Like “JIRA to product” nonsense. Is like we can hand wave the details that matter and somehow the end product would still be what we want.

Not even talking about the degenerative codebase you mention, which is also an issue I am seeing after some time in any project.

realaleris149•16h ago
> agents work in public channels, never DMs.

Recently I’ve been looking into why LLMs tend to write in some off putting ways. This one apparently is for LLMs to get better results.

It puts the cases to avoid closer to the main requirement so is easier for a LLM to follow without losing track with more ambiguous language.

Is very interesting because is harder to read by a human but it gets better results for an LLM.