Societal and technological progress as sewing an ever-growing, ever-changing, patchy, and polychrome quilt
This is experimental HTML to improve accessibility. We invite you to report rendering errors. Use Alt+Y to toggle on accessible reporting links and Alt+Shift+Y to toggle off. Learn more about this project and help improve conversions. HTML conversions sometimes display errors due to content that did not convert correctly from the source. This paper uses the following packages that are not yet supported by the HTML conversion tool. Feedback on these issues are not necessary; they are known and are being worked on. Authors: achieve the best HTML results from your LaTeX submissions by following these best practices. redacted \correspondingauthorJoel Z. Leibo (jzl@google.com) \reportnumber Artificial Intelligence (AI) systems are increasingly placed in positions where their decisions have real consequences, e.g., moderating online spaces, conducting research, and advising on policy. Ensuring they operate in a safe and ethically acceptable fashion is thus critical. However, most solutions
\pdftrailerid redacted \correspondingauthor Joel Z. Leibo (jzl@google.com) \reportnumber Societal and technological progress as sewing an ever-growing, ever-changing, patchy, and polychrome quilt Joel Z. Leibo Google DeepMind Alexander Sasha Vezhnevets Google DeepMind William A. Cunningham Google DeepMind University of Toronto Sébastien Krier Google DeepMind Manfred Diaz Mila - Québec AI Institute Simon Osindero Google DeepMind Abstract Artificial Intelligence (AI) systems are increasingly placed in positions where their decisions have real consequences, e.g., moderating online spaces, conduct
Explore this link on the map →related reading
- Societal and technological progress as sewing an ever-growing, ever-changing, patchy, and polychrome quilt — LessWronglesswrong.com
- The Possessed Machines: Dostoevsky's Demons and the Coming AGI Catastrophepossessedmachines.com
- ⿻ Symbiogenesis vs. Convergent Consequentialism — LessWronglesswrong.com
- Planning for AGI and beyond | OpenAIopenai.com
- Where I agree and disagree with Eliezer — LessWronglesswrong.com
- Another (outer) alignment failure story — AI Alignment Forumalignmentforum.org
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Do not conquer what you cannot defend — LessWronglesswrong.com
- ClearerThinking.org Podcast | Is AI going to ruin everything? (with Gabriel Alfour)podcast.clearerthinking.org
- Unfalsifiable stories of doom | Mechanize, Inc.mechanize.work
- [2605.10310] Positive Alignment: Artificial Intelligence for Human Flourishingarxiv.org