If you run more than one agent against shared state or shared tools, what actually went wrong?
1. Did two agents do the same work, or different work that conflicted? 2. Were your agents copies of each other (homogenous agents), or specialized with different roles (heterogenous)? 3. What did you add to stop or prevent the failure?
Anthropic's multi-agent research found that agents are low-variance: 18 of 30 independently created a git branch with the identical name. That suggests the common failure is correlated duplication rather than disagreement. Does that match what you have seen, or do yours diverge?