Research Areas in AI Control (The Alignment Project by UK AISI) — AI Alignment Forum
The field of AI Control aims to reduce this risk through techniques such as monitoring of untrusted AI systems by other AI models and restricting the affordances of untrusted AI systems.
x Research Areas in AI Control (The Alignment Project by UK AISI) — AI Alignment Forum The Alignment Project Research Agenda AI Frontpage 16 Research Areas in AI Control (The Alignment Project by UK AISI) by Julian Stastny , Tomek Korbak , Mojmir , Buck , Alan Cooney 1st Aug 2025 Linkpost for alignmentproject.aisi.gov.uk 21 min read 0 16 The Alignment Project is a global fund of over £15 million, dedicated to accelerating progress in AI control and alignment research. It is backed by an international coalition of governments, industry, venture capital and philanthropic funders. This post is pa
Explore this link on the map →saved by
related reading
- Recent Redwood Research project proposals — AI Alignment Forumalignmentforum.org
- 7+ tractable directions in AI control — AI Alignment Forumalignmentforum.org
- Research Areas in Benchmark Design and Evaluation (The Alignment Project by UK AISI) — AI Alignment Forumalignmentforum.org
- Research Areas in Methods for Post-training and Elicitation (The Alignment Project by UK AISI) — AI Alignment Forumalignmentforum.org
- Research Areas in Evaluation and Guarantees in Reinforcement Learning (The Alignment Project by UK AISI) — AI Alignment Forumalignmentforum.org
- Research Areas in Interpretability (The Alignment Project by UK AISI) — AI Alignment Forumalignmentforum.org
- The case for ensuring that powerful AIs are controlled — LessWronglesswrong.com
- AI Control: Improving Safety Despite Intentional Subversion — LessWronglesswrong.com
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- AI Control: Improving Safety Despite Intentional Subversion — AI Alignment Forumalignmentforum.org
- The case for ensuring that powerful AIs are controlledblog.redwoodresearch.org
- The Case Against AI Control Research — LessWronglesswrong.com