flâneur — a map of the web's best reading

Owain Evans, AI Alignment researcher

owainevans.github.io · 2,977 words · saved by 1 readers

Owain Evans is an AI Alignment researcher leading a new research group in Berkeley and affiliated with Oxford University. Discover his publications, blog posts, and collaborative opportunities on AI alignment, AGI risk, and related topics.

Owain Evans, AI Alignment researcher Blog posts Papers Video List of Mentees Owain Evans Director at Truthful AI (research group in Berkeley) Affiliate Researcher at CHAI, UC Berkeley Recent papers (May 2026): Negation Neglect: When models fail to learn negations in training . ( tweet , blog ) Conditional misalignment: common interventions can hide emergent misalignment behind contextual triggers . ( tweet ) The Consciousness Cluster: Emergent preferences of Models that Claim to be Conscious . ( blog ) Activation Oracles: Training and Evaluating LLMs as General-Purpose Activation Explainers .

Explore this link on the map →

related reading