✳flâneur — a map of the web's best reading
Ten people on the inside - by Buck Shlegeris
redwoodresearch.substack.com · 1,132 words · saved by 1 readers
A scary scenario that's worth planning for
Ten people on the inside A scary scenario that's worth planning for Buck Shlegeris Jan 28, 2025 16 Share (Many of these ideas developed in conversation with Ryan Greenblatt) In a shortform , I described some different levels of resources and buy-in for misalignment risk mitigations that might be present in AI labs: *The “safety case” regime.* Sometimes people talk about wanting to have approaches to safety such that if all AI developers followed these approaches, the overall level of risk posed by AI would be minimal. (These approaches are going to be more conservative than will probably be fe
Explore this link on the map →related reading
- Ten people on the inside — LessWronglesswrong.com
- Plans A, B, C, and D for misalignment risk — LessWronglesswrong.com
- Dario Amodei — The Adolescence of Technologydarioamodei.com
- Planning for Extreme AI Risks — AI Alignment Forumalignmentforum.org
- Racing through a minefield: the AI deployment problemcold-takes.com
- Responsible Scaling Policy v3 — LessWronglesswrong.com
- Efficient tradeoffs and the safety-usefulness tradeoff model — LessWronglesswrong.com
- Where I agree and disagree with Eliezer — LessWronglesswrong.com
- Ten AI safety projects I'd like people to work on — LessWronglesswrong.com
- Rohin Shah on what it's really like to run AGI safety at Google DeepMind (and where I disagree with 'doomers') | 80,000 Hours80000hours.org
- Dario Amodei — The Adolescence of Technologydarioamodei.com
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com