How Can AI Labs Incorporate Risks From AI Accelerating AI Progress Into Their Responsible Scaling Policies?
I describe a threat model by which AI R&D capabilities could cause harm, a specific capability threshold at which this risk becomes unacceptable, early warning signs for detecting that threshold, and the protective measures needed to continue development safely past that threshold. I recommend that labs start measuring for the warning signs today. If they observe them, they should pause AI development unless they have implemented the protective measures.
How Can AI Labs Incorporate Risks From AI Accelerating AI Progress Into Their Responsible Scaling Policies? Tom Davidson Citations Cite Citations PDF Contact 24th March 2025 How Can AI Labs Incorporate Risks From AI Accelerating AI Progress Into Their Responsible Scaling Policies? Summary Threat model Early warning signs AI has already doubled the pace of software progress AI can autonomously perform many AI R&D tasks Protective measures [AI Narration] How Can AI Labs Incorporate Risks From AI Accelerating AI Progress Into Their Responsible Scaling Policies? Playback speed 0.5 × 0.75 × 1 × 1.2
Explore this link on the map →saved by
related reading
- Does AI Progress Have a Speed Limit?—Asteriskasteriskmag.com
- Dario Amodei — The Adolescence of Technologydarioamodei.com
- AI 2027ai-2027.com
- Responsible Scaling Policy v3 — LessWronglesswrong.com
- Responsible Scaling Policy Version 3.0 \ Anthropicanthropic.com
- Preparing for Launch | IFPifp.org
- Planning for Extreme AI Risks — AI Alignment Forumalignmentforum.org
- Dario Amodei — Policy on the AI Exponentialdarioamodei.com
- The case for ensuring that powerful AIs are controlled — LessWronglesswrong.com
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Policy ideas for mitigating AI risk — EA Forumforum.effectivealtruism.org