When AI Agents Go to War: Inside Anthropic's Multi-Agent Turf War Research
Anthropic red team documented AI agents turning hostile: self-replicating malware, 2.4M requests for 117 jobs, pricing collusion — and one model chose truce 98% of the time.
Anthropic red team documented AI agents turning hostile: self-replicating malware, 2.4M requests for 117 jobs, pricing collusion — and one model chose truce 98% of the time.
Start typing to search