FAQ - Textbook from the Future
Trial and error works when failure is cheap. For the alignment of superintelligent AI we may only have one try. Only mathematical theory can provide assurance when there is no second try. The Textbook from the Future is a concept introduced by Yudkowsky. While the specific contents are obviously beyond our current knowledge, we can already guess at its overall gestalt; the table of contents. In other words, we are not certain of the answers contained in the textbook but we do know the questions it must answer. By committing the table of contents to paper we will have a skeleton that the community can flesh out. Importantly we can already start now! Even unfinished the textbook will enable scholars to draw from a coherent canon. The textbook is a coordination instrument. In some ways it is more accurate to think of it as the shared context of a research programme, and as a tool for establishing coherence in the field. It is the blueprint of the solution. Simultaneously it is also an on
FAQ - Textbook from the Future FAQ Why theory? Trial and error works when failure is cheap. For the alignment of superintelligent AI we may only have one try. Only mathematical theory can provide assurance when there is no second try. Isn't Alignment unsolved, how can you write a textbook on it? The Textbook from the Future is a concept introduced by Yudkowsky. “So, if we had the textbook from the future , like we have the textbook from 100 years in the future, which contains all the simple tricks that actually work robustly, then it would be much easier. But right now, we don't have that. We'
Explore this link on the map →related reading
- LessWronglesswrong.com
- Alignment remains a hard, unsolved problem — LessWronglesswrong.com
- Study Guide — LessWronglesswrong.com
- PSA: Almost nobody is directly working on superintelligent alignment — LessWronglesswrong.com
- AGI Ruin: A List of Lethalities — LessWronglesswrong.com
- Sequent: Scale and Automation for Higher Confidence in Alignment — Sequentsequent.org
- Shah and Yudkowsky on alignment failures — LessWronglesswrong.com
- Sequent: scale and automation for higher confidence in alignment — AI Alignment Forumalignmentforum.org
- OpenAI Launches Superalignment Taskforcethezvi.substack.com
- Critical review of Christiano's disagreements with Yudkowsky — LessWronglesswrong.com
- The Field of AI Alignment: A Postmortem, and What To Do About It — LessWronglesswrong.com
- Shtetl-Optimized >> Blog Archive >> Theory and AI Alignmentscottaaronson.blog