Cyber Competitions
Throughout 2025, we have been quietly entering Claude in cybersecurity competitions designed primarily for humans. Now, we want to share what we have learned. In many of these competitions Claude did pretty well, often placing in the top 25% of competitors. However, it lagged behind the best human teams at the toughest challenges. Our experience testing Claude in cyber competitions highlights the potential for AI to alter the offense-defense balance by making it easier for attackers to automate the exploitation of basic vulnerabilities. More research and development into AI-enabled cyber defense and resilience is needed to counter this development. AI is poised to transform the domain of cybersecurity. Anthropic’s Safeguards team recently identified and banned a user with limited coding abilities leveraging Claude to develop malware. Research suggests that this lowering of the bar for expertise needed to pose a threat, combined with the falling costs of large language models (LLMs), pr
Frontier Red Team Claude is competitive with humans in (some) cyber competitions Aug 9, 2025 Throughout 2025, we have been quietly entering Claude in cybersecurity competitions designed primarily for humans. Now, we want to share what we have learned. In many of these competitions Claude did pretty well, often placing in the top 25% of competitors. However, it lagged behind the best human teams at the toughest challenges. Our experience testing Claude in cyber competitions highlights the potential for AI to alter the offense-defense balance by making it easier for attackers to automate the exp
related reading
- Disrupting the first reported AI-orchestrated cyber espionage campaign \ Anthropicanthropic.com
- When AI builds itself \ Anthropicanthropic.com
- Investigating three real-world incidents in our cybersecurity evaluations \ Anthropicanthropic.com
- An alignment assessment of recent cybersecurity incidentsanthropic.com
- Countering misuse of AI: September 2026 / Anthropicanthropic.com
- The End-State Fallacy: Where Is AI Security Headed?endstatefallacy.com
- Project Glasswing: Securing critical software for the AI era \ Anthropicanthropic.com
- Claude Mythos Preview System Cardwww-cdn.anthropic.com
- Claude Was Just Following Orderssubstack.com
- Greg Brockman on Svbtleblog.gregbrockman.com
- Claude Sonnet 4.5 System Cardassets.anthropic.com
- Frontier AI Cybersecurity Observatorycybergym.io