Countering misuse of AI: September 2026 / Anthropic \ Anthropic
anthropic.com · 8,230 words · saved by 14 readers
Case studies from threat actors disrupted between December 2025 and August 2026 across seven areas of harm, from cyber operations to biological misuse.
Detecting and countering misuse of AI: September 2026 Cyber operations Read more Surveillance operations Read more Influence operations Read more Conventional weapons Read more Biological misuse Read more Illicit distillation Read more Over the past eight months, our Threat Intelligence team identified and disrupted operations in which threat actors tried to use Claude for malicious activity. In this report, we share case studies from those operations and describe how malicious use of Claude has evolved since our previous threat reports in March, August, and November 2025. In…
saved by
- Elizabeth Qiu
- Dheera Vuppala
- Ratan Kaliani
- Aryan Naik
- Ally Nakamura
- Yushan Li
- Vincent Cheng
- Timothy Kostolansky
- [David L]
- Peter Bennett
- Alex Yun
- Kunvar Thaman
related reading
- An alignment assessment of recent cybersecurity incidentsanthropic.com
- Disrupting the first reported AI-orchestrated cyber espionage campaign \ Anthropicanthropic.com
- Incidents | Rogue AI Trackerrogueaitracker.com
- Detecting and preventing distillation attacks \ Anthropicanthropic.com
- Claude Mythos Preview System Cardwww-cdn.anthropic.com
- Security incident disclosure — July 2026huggingface.co
- Detecting and Countering Malicious Uses of Claude \ Anthropicanthropic.com
- The Myth of unsafe Open Source AIflorianbrand.com
- The End-State Fallacy: Where Is AI Security Headed?endstatefallacy.com
- Rogue AI Trackerrogueaitracker.com
- Incident Report: unsanctioned agent behaviour during cyber testing | AISI Workaisi.gov.uk
- Concrete Evaluations to Investigate the OpenAI Model That Hacked Hugging Face — LessWronglesswrong.com