Research Areas in AI Control (The Alignment Project by UK AISI) — AI Alignment Forum
The field of AI Control aims to reduce this risk through techniques such as monitoring of untrusted AI systems by other AI models and restricting the affordances of untrusted AI systems.
x Research Areas in AI Control (The Alignment Project by UK AISI) — AI Alignment Forum The Alignment Project Research Agenda AI Frontpage 16 Research Areas in AI Control (The Alignment Project by UK AISI) by Julian Stastny , Tomek Korbak , Mojmir , Buck , Alan Cooney 1st Aug 2025 Linkpost for alignmentproject.aisi.gov.uk 21 min read 0 16 The Alignment Project is a global fund of over £15 million, dedicated to accelerating progress in AI control and alignment research. It is backed by an international coalition of governments, industry, venture capital and philanthropic funders. This post is pa
saved by
related reading
- Recent Redwood Research project proposals — AI Alignment Forumalignmentforum.org
- 7+ tractable directions in AI control — AI Alignment Forumalignmentforum.org
- Research Areas in Benchmark Design and Evaluation (The Alignment Project by UK AISI) — AI Alignment Forumalignmentforum.org
- Research Areas in Methods for Post-training and Elicitation (The Alignment Project by UK AISI) — AI Alignment Forumalignmentforum.org
- Research Areas in Evaluation and Guarantees in Reinforcement Learning (The Alignment Project by UK AISI) — AI Alignment Forumalignmentforum.org
- Research Areas in Interpretability (The Alignment Project by UK AISI) — AI Alignment Forumalignmentforum.org
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- The case for ensuring that powerful AIs are controlled — LessWronglesswrong.com
- Reading Listblog.redwoodresearch.org
- AI Control: Improving Safety Despite Intentional Subversion — LessWronglesswrong.com
- AI Control: Improving Safety Despite Intentional Subversion — AI Alignment Forumalignmentforum.org
- Introduction to AI Control - by Sarah - BlueDot Impactblog.bluedot.org