Jan Leike on OpenAI's massive push to make superintelligence safe in 4 years or less - 80,000 Hours
Enjoyed the episode? Want to listen later? Subscribe by searching “80,000 Hours” wherever you get your podcasts, or click one of the buttons below: If you’re thinking about how do you align the superintelligence — how do you align the system that’s vastly smarter than humans? — I don’t know. I don’t have an answer. I don’t think anyone really has an answer. But it’s also not the problem that we fundamentally need to solve. Maybe this problem isn’t even solvable by humans who live today. But there’s this easier problem, which is how do you align the system that is the next generation? How do you align GPT-N+1? And that is a substantially easier problem. Jan Leike In July, OpenAI announced a new team and project: Superalignment. The goal is to figure out how to make superintelligent AI systems aligned and safe to use within four years, and the lab is putting a massive 20% of its computational resources behind the effort. Today’s guest, Jan Leike, is Head of Alignment at OpenAI and will b
Jan Leike on OpenAI's massive push to make superintelligence safe in 4 years or less | 80,000 Hours Search for: On this page: Introduction 1 Highlights 2 Articles, books, and other media discussed in the show 3 Transcript 3.1 Cold Open [00:00:00] 3.2 Rob's intro [00:00:47] 3.3 The interview begins [00:03:30] 3.4 The Superalignment project [00:04:04] 3.5 Why 4 years? And why does the project need so much compute? [00:13:07] 3.6 State of the art in alignment methods [00:22:56] 3.7 The Superalignment team isn't trying to train a really good ML researcher [00:35:07] 3.8 Reactions from the bro
Explore this link on the map →related reading
- OpenAI Launches Superalignment Taskforcethezvi.substack.com
- Alignment remains a hard, unsolved problem — LessWronglesswrong.com
- ALIGNMENT - by vincent huang - a slice of my mindmindslice.substack.com
- AI 2027ai-2027.com
- IIIc. Superalignment - SITUATIONAL AWARENESSsituational-awareness.ai
- AGI Ruin: A List of Lethalities — LessWronglesswrong.com
- Quotes from Leopold Aschenbrenner's Situational Awareness Paperthezvi.substack.com
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- Product Alignment is not Superintelligence Alignment (and we need the latter to survive) — LessWronglesswrong.com
- AI 2027ai-2027.com
- PSA: Almost nobody is directly working on superintelligent alignment — LessWronglesswrong.com
- Sequent: Scale and Automation for Higher Confidence in Alignment — Sequentsequent.org