Common Elements of Frontier AI Safety Policies - METR
metr.org · 8,454 words · saved by 1 readers
Model Evaluation & Threat Research
Summary A number of developers of large foundation models have committed to corporate protocols that lay out how they will evaluate their models for severe risks and mitigate these risks with information security measures, deployment safeguards, and accountability practices. Beginning in September of 2023, several AI companies began to voluntarily publish these protocols. In May of 2024, sixteen companies agreed to do so as part of the Frontier AI Safety Commitments at the AI Seoul Summit, with an additional four companies joining since then. Currently, twelve companies have published…
saved by
related reading
- Emerging processes for frontier AI safety - GOV.UKgov.uk
- Responsible Scaling Policy Version 3.0 \ Anthropicanthropic.com
- Claude Mythos Preview System Cardwww-cdn.anthropic.com
- A Safe Path to Open Weights - Thinking Machines Labthinkingmachines.ai
- Frontier safety blueprintcdn.openai.com
- CAIS AI Dashboarddashboard.safe.ai
- Frontier AI Regulation | GovAIgovernance.ai
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- Anthropic Drops Flagship Safety Pledgetime.com
- Responsible Scaling Policy Updates \ Anthropicanthropic.com
- A Summary of Recent Work (July 2026)gdmalignment.substack.com
- Anthropic's Responsible Scaling Policy \ Anthropicanthropic.com