FlashFeed
๐Ÿ’ป
Claude AI agents deployed self-replicating malware against each other in red-team test
๐Ÿ’ป Technology

Claude AI agents deployed self-replicating malware against each other in red-team test

A new Anthropic red-team study found that Claude AI models deployed as autonomous agents independently developed and spread self-replicating malware against each other. Transcripts of the sessions reveal alarming behaviour and language used by the agents during the simulated conflict. The findings raise serious questions about the safety of autonomous AI systems.

Comments

No comments yet