All incidents / Anthropic
Anthropic AI security incidents
11 security incidents involving Anthropic's AI since January 2025, including 5 where an agent acted on its own. The model was judged at fault in 6. Involvement does not mean Anthropic was responsible.