Anthropic has released its most comprehensive threat intelligence report to date, detailing how adversaries tried to exploit its AI model Claude for malicious purposes. The report covers attempts to use Claude in cyberattacks, influence campaigns, surveillance, biological research, and weapons development, and describes how these operations were identified and disrupted.

David Agranovich, a former Meta threat disruption lead, commented on the report’s significance, noting that the gap between individual operators and nation-state actors is narrowing due to agentic AI tools capable of reconnaissance, exploitation, and data exfiltration. He emphasized that stolen API keys have become a prime target, enabling attackers to misuse AI resources under legitimate credentials.

The report also sheds light on AI’s role in modern surveillance and influence operations. For example, espionage groups have leveraged AI to analyze vast amounts of social media data for targeting, while some actors used Claude to conduct multilingual outreach and role-play scenarios to recruit individuals.

Anthropic’s transparency in exposing these threats is notable, as it allows platforms and civil society to better understand and counter AI-enabled risks. However, Agranovich pointed out that the report focuses mainly on offensive uses of AI, while the same technology can assist defenders in malware analysis and threat detection.

The findings underscore the importance of securing AI model access and fostering collaboration among AI companies, platforms, and regulators to share threat intelligence and build collective defenses. Without such cooperation, the risks posed by AI misuse may grow unchecked.