Owain Evans
Owain Rhys Evans is a British artificial intelligence researcher who works on AI alignment and machine learning safety. He founded Truthful AI, a research group based in Berkeley, California, and is an affiliate of the Center for Human Compatible AI (CHAI) at the University of California, Berkeley. His research addresses AI truthfulness, emergent behaviors in large language models, and the alignment of AI systems with human values.
Owain Evans - Wikipedia Jump to content From Wikipedia, the free encyclopedia This biographical article is written like a résumé . Please help improve it by revising it to be neutral and encyclopedic . ( May 2026 ) Not to be confused with Owain Wyn Evans . Owain Evans Alma mater Massachusetts Institute of Technology (PhD) Columbia University (BA) Known for AI alignment research TruthfulQA benchmark Reversal curse Emergent misalignment Scientific career Fields Artificial intelligence , AI safety , machine learning Institutions Truthful AI Center for Human Compatible AI , UC Berkeley Future of H
Explore this link on the map →related reading
- Owain Evans, AI Alignment researcherowainevans.github.io
- Current AIs seem pretty misaligned to me — LessWronglesswrong.com
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- Current AIs seem pretty misaligned to meblog.redwoodresearch.org
- Automated Alignment Researchers: Using large language models to scale scalable oversight \ Anthropicanthropic.com
- Import AIjack-clark.net
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Discovering Language Model Behaviors with Model-Written Evaluations — LessWronglesswrong.com
- Model Organisms of Misalignment: The Case for a New Pillar of Alignment Research — AI Alignment Forumalignmentforum.org
- Tim Hua Personal Websitetimhua.me
- AI Safety | Arkosevictoriabrook.github.io
- Group | Sherry Tongshuang Wucs.cmu.edu