TechnicalAgenda.pdf
intelligence.org · 7,589 words · saved by 1 readers
N/A
Agent Foundations for Aligning Machine Intelligence with Human Interests: A Technical Research Agenda In The Technological Singularity: Managing the Journey. Springer. 2017 Nate Soares and Benya Fallenstein Machine Intelligence Research Institute {nate,benya}@intelligence.org Contents 1 Introduction…
saved by
related reading
- Embedded Agency (full-text version) — LessWronglesswrong.com
- Approval-directed agency and the decision theory of Newcomb-like problemslink.springer.com
- AGI Ruin: A List of Lethalities — LessWronglesswrong.com
- The hot mess theory of AI misalignment: More intelligent agents behave less coherently | Jascha’s blogsohl-dickstein.github.io
- VingeanReflection.pdfintelligence.org
- Four Background Claims - Machine Intelligence Research Instituteintelligence.org
- Optimality is the tiger, and agents are its teeth — LessWronglesswrong.com
- Towards a scale-free theory of intelligent agencymindthefuture.info
- [1902.09469] Embedded Agencyarxiv.org
- Towards a scale-free theory of intelligent agency — AI Alignment Forumalignmentforum.org
- What Does It Mean to Align AI With Human Values? | Quanta Magazinequantamagazine.org
- Toward Idealized Decision Theoryarxiv.org