flâneur — a map of the web's best reading

The limits of AI safety via debate - LessWrong

lesswrong.com · 6,794 words · saved by 1 readers

The limits of AI safety via debate • I recently participated in the AGI safety fundamentals program and this is my cornerstone project. During our re…

x The limits of AI safety via debate — LessWrong Debate (AI safety technique) AI Frontpage 36 The limits of AI safety via debate by Marius Hobbhahn 10th May 2022 AI Alignment Forum 11 min read 8 36 Ω 17 The limits of AI safety via debate I recently participated in the AGI safety fundamentals program and this is my cornerstone project. During our readings of AI safety via debate ( blog , paper ) we had an interesting discussion on its limits and conditions under which it would fail. I spent only around 5 hours writing this post and it should thus mostly be seen as food for thought rather than r

Explore this link on the map →

related reading