Skip to content
Unlisted Report logoUnlisted ReportSubscribe
AI Security

Anthropic report details how hackers misuse AI models

Anthropic says it disrupted Russian espionage, Chinese model-copying, scam apps and surveillance tools that tried to abuse its Claude models.

By · Published · Updated · 4 min read

Anthropic report details how hackers misuse AI models

Anthropic has published a report describing how threat actors tried to misuse its Claude AI models between December 2025 and August 2026, according to the company.

What the report covers

Anthropic says it disrupted activity across seven areas:

  • Cyber operations
  • Influence operations
  • Surveillance
  • Scams and fraud
  • Biological misuse
  • Conventional weapons development
  • "Distillation" — copying a model's abilities

The actors included suspected state-sponsored groups, criminals, commercial spyware vendors and state propaganda bodies. Cases ranged from a network of fake dating apps built to defraud users to surveillance systems designed to identify and monitor dissidents.

Headline cases

  • A suspected Russia-linked espionage campaign used Claude to automatically rebuild its malware whenever security tools detected it, The Hacker News reports.
  • Anthropic accused Chinese AI firms of trying to extract and replicate Claude's abilities, Reuters reports.

Why it matters

AI tools speed up attackers just as they speed up everyone else. Reports like this show defenders what real-world abuse looks like, and they put pressure on AI companies to detect and share misuse.

What organizations should take from it

  • Assume attackers can update their malware quickly; rely on behaviour-based detection, not only known signatures.
  • Treat AI tools your staff use as part of your security policy.

Sources

Read next