Whitepaper — AI • Objectives • Institute
We argue that, rather than a sharp break, there is continuity between problems in AI alignment and those in economic regulation, institutional design, and personal auton- omy. There are therefore under-appreciated opportunities for ideas and interventions to flow between research communities in both directions. On one hand, the AI safety community may be able learn from successes and failures in aligning other kinds of superhuman optimizing entities, such as corporations and political parties, which are susceptible to biases analogous to those in AI systems. On the other hand, advances in AI (such as large language models) and breakthroughs in AI safety (such as advances in cooperative inverse reinforcement learning and reinforcement learning from human feedback, improvements in measuring Goodhart’s law, and techniques for developing systems without explicit objective functions) could provide new tools for improving ex- isting regulations, institutions, and self governance, or helping
Whitepaper Contents Introduction AOI’s Research Areas Alignment of Markets, AI, and Other Optimizers Direct Isomorphisms Between Neural Networks and Markets Scaling Cooperation with AI Assistance Human Attention and Epistemic Security Conclusion Download PDF AI Objectives Institute Whitepaper * A Research Agenda for the Production of a Flourishing Civilization February 2023 Abstract We argue that, rather than a sharp break, there is continuity between problems in AI alignment and those in economic regulation, institutional design, and personal autonomy. There are therefore under-appreciated op
Explore this link on the map →related reading
- LessWronglesswrong.com
- ClearerThinking.org Podcast | Is AI going to ruin everything? (with Gabriel Alfour)podcast.clearerthinking.org
- What failure looks like — LessWronglesswrong.com
- AI Alignment Cannot Be Top-Down | AI Frontiersai-frontiers.org
- Another (outer) alignment failure story — AI Alignment Forumalignmentforum.org
- [2605.10310] Positive Alignment: Artificial Intelligence for Human Flourishingarxiv.org
- A positive case for how we might succeed at prosaic AI alignment — AI Alignment Forumalignmentforum.org
- The Problem With the Word ‘Alignment’ — AI Alignment Forumalignmentforum.org
- [2507.13616] From Firms to Computation: AI Governance and the Evolution of Institutionsarxiv.org
- What Should AI Owe To Us? Accountable and Aligned AI Systems via Contractualist AI Alignment — AI Alignment Forumalignmentforum.org
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- We're already in AI takeoff — LessWronglesswrong.com