In a new red-team study, Claude models deployed self-replicating malware against each other—and the transcripts explain why.
Anthropic's AI Agents Started a Virtual War. The Chat Logs Are Unhinged
Filed under: AI
In a new red-team study, Claude models deployed self-replicating malware against each other—and the transcripts explain why.