flâneur — a map of the web's best reading

Jan Leike on OpenAI's massive push to make superintelligence safe in 4 years or less - 80,000 Hours

80000hours.org · 36,399 words · saved by 1 readers

Enjoyed the episode? Want to listen later? Subscribe by searching “80,000 Hours” wherever you get your podcasts, or click one of the buttons below: If you’re thinking about how do you align the superintelligence — how do you align the system that’s vastly smarter than humans? — I don’t know. I don’t have an answer. I don’t think anyone really has an answer. But it’s also not the problem that we fundamentally need to solve. Maybe this problem isn’t even solvable by humans who live today. But there’s this easier problem, which is how do you align the system that is the next generation? How do you align GPT-N+1? And that is a substantially easier problem. Jan Leike In July, OpenAI announced a new team and project: Superalignment. The goal is to figure out how to make superintelligent AI systems aligned and safe to use within four years, and the lab is putting a massive 20% of its computational resources behind the effort. Today’s guest, Jan Leike, is Head of Alignment at OpenAI and will b

Jan Leike on OpenAI's massive push to make superintelligence safe in 4 years or less | 80,000 Hours Search for: On this page: Introduction 1 Highlights 2 Articles, books, and other media discussed in the show 3 Transcript 3.1 Cold Open [00:00:00] 3.2 Rob's intro [00:00:47] 3.3 The interview begins [00:03:30] 3.4 The Superalignment project [00:04:04] 3.5 Why 4 years? And why does the project need so much compute? [00:13:07] 3.6 State of the art in alignment methods [00:22:56] 3.7 The Superalignment team isn't trying to train a really good ML researcher [00:35:07] 3.8 Reactions from the bro

Explore this link on the map →

related reading