Anthropic reports blocking multiple malicious uses of its Claude AI model, including attempts to develop missile-guidance software in Yemen, Russian cyber-espionage targeting Ukraine and Europe, Chinese operations against Middle Eastern and Asian networks, and Iranian influence campaigns. The company says it banned associated accounts and shared threat intelligence with partners, though some malicious requests evaded safeguards by obscuring intent across separate sessions.
Anthropic released a threat intelligence report documenting how Claude was misused for cyberattacks, influence operations, surveillance, and weapons development, with a notable case involving a Chinese AI company (Moonshot) serving Claude instead of its own model while collecting user exchanges for training data.