Research Mission | Oxford Witt Lab
OWL’s mission is to de-risk the large-scale deployment of advanced AI by building trust through technical assurance. Technical assurance is the disciplined, auditable evidence that an AI system will behave within specified safety and security bounds - by design, under test, and in operation - even in adversarial and multi-agent settings. We work from first principles to practice: from the mathematical foundations of multi-agent security and undetectable threats to next-generation evaluations and real-world demonstrations in domains such as open-source intelligence, biology, and climate science. We are goal-driven and method-agnostic. Depending on the question, we draw on high-dimensional anomaly detection, game theory, RL/MARL, cryptography and steganography, and interpretability (theory and practice). We publish across leading AI, security, and interdisciplinary venues. OWL is pioneering the field of multi-agent security, addressing a critical gap in current AI safety and security by
Research Mission OWL’s mission is to de-risk the large-scale deployment of advanced AI by building trust through technical assurance . Technical assurance is the disciplined, auditable evidence that an AI system will behave within specified safety and security bounds - by design, under test, and in operation - even in adversarial and multi-agent settings. We work from first principles to practice: from the mathematical foundations of multi-agent security and undetectable threats to next-generation evaluations and real-world demonstrations in domains such as open-source intelligence, biology, a
Explore this link on the map →related reading
- Oxford Witt Labwittlab.ai
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- Student Projects / Supervision | Oxford Witt Labwittlab.ai
- Off Target | CNAScnas.org
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- AI Control: Improving Safety Despite Intentional Subversion — LessWronglesswrong.com
- AI Safety | Arkosevictoriabrook.github.io
- 2312.06942arxiv.org
- [2312.06942] AI Control: Improving Safety Despite Intentional Subversionarxiv.org
- AI safety - Wikipediaen.wikipedia.org
- gdm-ai-control-roadmap.pdfstorage.googleapis.com