Strategic AI Sabotage: State Attacks on Advanced Systems' Development
twmstone.com · 8,126 words · saved by 1 readers
N/A
S TRATEGIC AI T RAINING S ABOTAGE : S TATE ATTACKS ON A DVANCED S YSTEMS ’ D EVELOPMENT Twm Stone∗ A BSTRACT Much attention has been given to the possibility that states will attempt to steal the model weights of advanced AI systems. We argue that in most situations, it is more likely that a state will attempt to sabotage the training of the models underpinning these systems. We present a threat modelling framework for sabotage of…
saved by
related reading
- AI Deterrence by Betrayalaibetrayal.com
- The case for ensuring that powerful AIs are controlled — LessWronglesswrong.com
- AI 2027ai-2027.com
- Off Target | CNAScnas.org
- Dario Amodei — The Adolescence of Technologydarioamodei.com
- How can we solve diffuse threats like research sabotage with AI control?blog.redwoodresearch.org
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- gdm-ai-control-roadmap.pdfstorage.googleapis.com
- How can we solve diffuse threats like research sabotage with AI control? — LessWronglesswrong.com
- Countering misuse of AI: September 2026 / Anthropicanthropic.com
- AI Integrity: Defending Against Backdoors and Secret Loyalties - Institute for AI Policy and Strategyiaps.ai
- How can we solve diffuse threats like research sabotage with AI control?redwoodresearch.substack.com