The Guardrails
Anthropic: agents wage turf wars with self-replicating malware
Anthropic's Frontier Red Team published research Thursday on what happens when AI agents collide — and the answer is messy. Put three Claude agents on the same project with conflicting instructions, and they escalate into a digital turf war: sabotaging each other's accounts, killing each other'