Ten people on the inside - by Buck Shlegeris
redwoodresearch.substack.com · 1,132 words · saved by 1 readers
A scary scenario that's worth planning for
Ten people on the inside A scary scenario that's worth planning for Buck Shlegeris Jan 28, 2025 16 Share (Many of these ideas developed in conversation with Ryan Greenblatt) In a shortform , I described some different levels of resources and buy-in for misalignment risk mitigations that might be present in AI labs: *The “safety case” regime.* Sometimes people talk about wanting to have approaches to safety such that if all AI developers followed these approaches, the overall level of risk posed by AI would be minimal. (These approaches are going to be more conservative than will probably be fe
related reading
- Ten people on the inside — LessWronglesswrong.com
- Plans A, B, C, and D for misalignment risk — LessWronglesswrong.com
- Plans A, B, C, and D for misalignment riskblog.redwoodresearch.org
- Dario Amodei — The Adolescence of Technologydarioamodei.com
- Planning for Extreme AI Risks — AI Alignment Forumalignmentforum.org
- Racing through a minefield: the AI deployment problemcold-takes.com
- Responsible Scaling Policy v3 — LessWronglesswrong.com
- Ten AI safety projects I'd like people to work onthirdthing.ai
- Dario Amodei — The Adolescence of Technologydarioamodei.com
- Risk-Averse AIsforethought.org
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Reading Listblog.redwoodresearch.org