Measuring AI capabilities in intelligence targeting and conventional weapons \ Anthropic
Anthropic’s Frontier Red Team developed new evaluations to measure AI capabilities in tactical intelligence targeting and conventional weapons development.
Anthropic’s Frontier Red Team developed new evaluations to measure AI capabilities in tactical intelligence targeting (like finding where people are based on fragmentary information) and conventional weapons development (like engineering drones to strike a moving target). For some tasks in military and intelligence domains, models could do things that, historically, only a set of scarce, highly-trained human experts could do. These evaluations show how models have become useful to actors seeking to misuse our platform for surveillance and conventional weapons development. They also show…
saved by
related reading
- Research — Paradigm 3paradigm3.org
- Countering misuse of AI: September 2026 / Anthropicanthropic.com
- GPT-6 Astra System Card - OpenAI Deployment Safety Hubdeploymentsafety.openai.com
- Claude Mythos Preview System Cardwww-cdn.anthropic.com
- Claude Mythos Preview System Cardwww-cdn.anthropic.com
- Off Target | CNAScnas.org
- CAIS AI Dashboarddashboard.safe.ai
- Governing Automated Strategic Intelligencearxiv.org
- AI in 2025: gestalt — LessWronglesswrong.com
- Frontier Risk Report (February to March 2026) - METRmetr.org
- [2403.13793] Evaluating Frontier Models for Dangerous Capabilitiesarxiv.org
- My picture of the present in AI — LessWronglesswrong.com