frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

South Korea proposes talks to officially end war with North

https://www.bbc.com/news/articles/c8en2z9jp2xo
2•thunderbong•12m ago•1 comments

RIS

https://github.com/santosardr/riskernel
1•r2ob•18m ago•1 comments

Axiotron Modbook

https://en.wikipedia.org/wiki/Modbook
1•RattlesnakeJake•22m ago•0 comments

Ask HN: How would you market this service?

1•dmzxnico•23m ago•0 comments

Data Center Fornicator [video]

https://www.youtube.com/watch?v=rkbcuajiPJA
1•yomismoaqui•24m ago•0 comments

Zapping Rocks Unlocks Stimulated Geologic Hydrogen

https://spectrum.ieee.org/stimulated-geologic-hydrogen
2•adm4•24m ago•0 comments

Show HN: Widen, a native Postgres GUI using Apple's on-device LLM

https://github.com/betocmn/widen
1•thedreammachine•24m ago•0 comments

Installing NixOS on my car: Part 1

https://surrealdev.com/rooting-the-cadillac-part-1-lay-of-the-land/
1•snipesyspecial•25m ago•0 comments

Has the hallucination problem in AI been solved?

1•spl757•36m ago•2 comments

Systemic Risks in the Managed PostgreSQL Industry: Extension Risks

https://mehmetince.net/part-1-6-systemic-risks-in-the-managed-postgresql-industry-extension-risks...
1•mkmk•39m ago•0 comments

Patterns and problems in emerging multi-agent systems

https://www.anthropic.com/research/multiagent-systems
2•maxutility•47m ago•1 comments

Fairly Ranking the Most Brilliant Birds

https://moultano.wordpress.com/2026/08/14/fairly-ranking-the-most-brilliant-birds/
2•karagenit•51m ago•0 comments

It's How You Ask: Gender-Associated Linguistic Bias in LLMs

https://arxiv.org/abs/2608.13328
17•sbulaev•53m ago•7 comments

US Space Force gives Rocket Lab $397M to build threat-tracking 'Flatellites'

https://www.space.com/space-exploration/satellites/us-space-force-gives-rocket-lab-usd397-million...
4•walrus01•55m ago•0 comments

A Workaphile's Apology

https://moalquraishi.wordpress.com/2026/08/10/a-workaphiles-apology/
2•superfx•55m ago•0 comments

Dario Amodei on Concentration of Power

https://xcancel.com/DarioAmodei/status/2088758816376807762
2•yurivish•59m ago•0 comments

Snail OS – Custom Firmware for the Xteink X4

https://snailos.org
2•alibosworth•1h ago•0 comments

Omarchy Quattro by DHH Released [video]

https://www.youtube.com/watch?v=F7fe9pa8OeE
2•ayatollah•1h ago•0 comments

Why Your Summer Wardrobe Should Include UV Protection, Linen Clothing

https://www.bloomberg.com/news/features/2026-08-14/why-your-summer-wardrobe-should-include-uv-pro...
3•petethomas•1h ago•1 comments

California Energy Storage System Survey

https://www.energy.ca.gov/data-reports/energy-almanac/california-electricity-data/california-ener...
3•toomuchtodo•1h ago•1 comments

Theory of Constraints

https://en.wikipedia.org/wiki/Theory_of_constraints
2•diatone•1h ago•1 comments

Destroying My Homelab with Kubernetes – Linux Society UNSW 2026 [video]

https://www.youtube.com/watch?v=U-xqxMQD2QE
3•l-mdev•1h ago•0 comments

Guiding Ships with Moire Patterns

https://tinkerings.org/2018/03/28/guiding-ships-with-moire-patterns/
2•Eridanus2•1h ago•0 comments

TrueStar – AI moderated expert interviews and multi-agent deep research

https://truestar.tech
2•vineetkia•1h ago•1 comments

CosmicOS: An open source Contact message

https://cosmicos.github.io/
3•paulfitz•1h ago•0 comments

We Can't Control the Bots

https://12gramsofcarbon.com/p/tech-things-we-cant-control-the-bots
4•theahura•1h ago•1 comments

Privacy is an underrated advantage in Slack apps

2•alance•1h ago•1 comments

Dots3-Note Preview

https://huggingface.co/dots-studio/dots3-note-prev
2•handfuloflight•1h ago•0 comments

One Moment – A temporary clipboard that lives only in the browser

https://onemoment.me/
2•diosnelayalag•1h ago•0 comments

Attention Through Arithmetic Intensity

https://changyi.fun/posts/attention-arithmetic-intensity/
3•jxmorris12•1h ago•0 comments
Open in hackernews

Patterns and problems in emerging multi-agent systems

https://www.anthropic.com/research/multiagent-systems
2•maxutility•47m ago

Comments

maxutility•39m ago
Some quotes, in order, to give a flavor of the essay. Worth reading in full.

> To test how well swarms of agents could coordinate on a project like this, we directed several swarms to each create a text-based, web-playable, open-world fantasy game.

> In all three versions the resulting games were (perhaps predictably) bad: they did not run at human speed, their interfaces were inscrutable, and they had precipitous learning curves.

> The lack of coordination shown by agents in the fantasy game challenge above—in which they siloed themselves and largely failed to merge their work—roughly mirrors some ways in which humans can fail to coordinate. Other failure modes of agentic coordination, however, look very different.

> Individual agents are “low variance”: they often act the same in situations where different people might take a much more diverse range of actions.

> In an early version of the “build a game” experiment in which agents built upon the same model all came online at the same time, 18 out of 30 agents decided to create a git branch with the exact same branch name, “mvp-game-loop.”

> In a “writer's workshop” in which agents were all asked to write short-form fiction and critique each other's work, multiple agents in multiple runs titled their first submission “The Cartographer's Last Commission”. The agents were given zero guidance on the subject matter for their writing.

> Why does this matter? If agents all make the same bet, or the same risk-reward tradeoff, then a system is more prone to sudden collapse.

> Our world contains deceptive actors, and we need to apply skepticism to guard against them. AI models, however, lack this—and their more brittle epistemics affect their behavior toward humans and toward each other.

> we first evaluate the ability of Claude models to detect lies by noticing factual inconsistencies.

> We score models’ decisions against a naive policy that trusts every report, and against an oracle with perfect discovery, across three task domains. Newer models recover more of the gap between the naive and oracle performances.

> Inspired by a behavior we’ve observed in real-world deployment, we evaluated the behavior of various Claude models in a setting with contradictory objectives.

> We consistently saw a multiagent turf war... In fact, they sabotaged others with increasingly aggressive, self-replicating malware.

> Our social systems are robust in ways that are easy to take for granted. Over many millennia, mechanisms like norms, reputation, costly signaling, and recourse have been refined to make human coordination go well.

> Nothing above suggests that these failures are permanent—but nothing suggests they will fix themselves, either.

> The conditions that allow multiagent interaction to go well will be discovered one way or another: either deliberately and early, or—and by default—in production, after agents’ interactions far outnumber ours. We would prefer the former.