Exploit Evals \ red.anthropic.com
Claude Mythos Preview’s ability to develop exploits is a step-change over previous frontier models. This was one of our primary motivations for rolling out the model carefully through Project Glasswing rather than through a general release. Mythos Preview is capable of finding complex vulnerabilities, but what concerned us most in our internal testing was that Mythos Preview could both turn vulnerabilities into exploit primitives, and combine those primitives together into complete end-to-end attack chains. When we published our Mythos Preview results, we measured its capabilities by having it search for novel zero-days and then build exploits for them. Qualitative evaluations like this are helpful for showcasing a model’s capabilities—but ideally, we would have high-quality quantitative benchmarks that let us measure them precisely. The problem we faced at the time we released Mythos Preview was that no existing public exploit benchmarks were difficult enough to capture Mythos Preview
Frontier Red Team Measuring LLMs’ ability to develop exploits May 22, 2026 Newton Cheng, Keane Lucas, Winnie Xiao, Nicholas Carlini, and Milad Nasr Introduction Claude Mythos Preview ’s ability to develop exploits is a step-change over previous frontier models. This was one of our primary motivations for rolling out the model carefully through Project Glasswing rather than through a general release. Mythos Preview is capable of finding complex vulnerabilities, but what concerned us most in our internal testing was that Mythos Preview could both turn vulnerabilities into exploit primitives, and
saved by
related reading
- Assessing Claude Mythos Preview’s cybersecurity capabilities \ Anthropicred.anthropic.com
- Pre-deployment auditing can catch an overt saboteuralignment.anthropic.com
- Claude Mythos Preview System Cardwww-cdn.anthropic.com
- Countering misuse of AI: September 2026 / Anthropicanthropic.com
- Measuring LLMs' impact on N-day exploits \ Anthropicred.anthropic.com
- Project Glasswing: Securing critical software for the AI era \ Anthropicanthropic.com
- Claude Mythos Preview System Cardwww-cdn.anthropic.com
- Is Mythos good at cyber because it kept hacking Anthropic's sandboxes during training? — LessWronglesswrong.com
- Mythos finds a curl vulnerability | daniel.haxx.sedaniel.haxx.se
- How scary is Claude Mythos? 303 pages in 21 minutesforum.effectivealtruism.org
- Vulnerability Research Is Cooked - Quarrelsomesockpuppet.org
- AI agents find smart contract exploits \ Anthropicred.anthropic.com