✳flâneur — a map of the web's best reading
Ajeya Cotra on accidentally teaching AI models to deceive us - 80,000 Hours
80000hours.org · 38,258 words · saved by 1 readers
Imagine you are an orphaned eight-year-old whose parents left you a $1 trillion company, and no trusted adult to serve as your guide to the world...
Ajeya Cotra on accidentally teaching AI models to deceive us | 80,000 Hours Search for: On this page: Introduction 1 Highlights 2 Articles, books, and other media discussed in the show 3 Transcript 3.1 Rob's intro [00:00:00] 3.2 The interview begins [00:02:38] 3.3 How Ajeya's views have changed since 2020 [00:05:09] 3.4 Are neural networks more like a sped-up version of evolution, or a slower version of human learning? [00:17:42] 3.5 Situational awareness [00:26:10] 3.6 Misalignment stories Ajeya doesn't buy [00:42:03] 3.7 The orphan heir with a trillion-dollar fortune [00:59:14] 3.8 Saints, S
Explore this link on the map →related reading
- Dario Amodei — The Adolescence of Technologydarioamodei.com
- AI 2027ai-2027.com
- Alignment remains a hard, unsolved problem — LessWronglesswrong.com
- Current AIs seem pretty misaligned to me — LessWronglesswrong.com
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- AI 2027ai-2027.com
- AI #24: Week of the Podcast — LessWronglesswrong.com
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Andrej Karpathy — AGI is still a decade awaydwarkesh.com
- Current AIs seem pretty misaligned to meblog.redwoodresearch.org
- A positive case for how we might succeed at prosaic AI alignment — AI Alignment Forumalignmentforum.org
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com