The Scalable Formal Oversight Research Program — LessWrong
Introduction Every serious person who thinks about AI safety observes the fundamental asymmetry between the ease of AI content generation and difficu…
x The Scalable Formal Oversight Research Program — LessWrong AI Frontpage 36 The Scalable Formal Oversight Research Program by Max von Hippel 22nd Feb 2026 11 min read 5 36 Introduction Every serious person who thinks about AI safety observes the fundamental asymmetry between the ease of AI content generation and difficulty of audits thereof. In the codegen context, this leads unserious people to talk about the need for AI-driven unit test generation or LLM-as-a-judge , and serious people to talk about concepts like property based testing or refinement testing, fuzzing, interactive theorem pro
related reading
- AI in 2025: gestalt — LessWronglesswrong.com
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- A shallow dive into formal verificationvitalik.eth.limo
- Verified Machine Learning Infrastructure: Formal Methods for Trustworthy Artificial Intelligence Deployment | RANDrand.org
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Intent Formalization: A Grand Challenge for Reliable Coding in the Age of AI Agentsalphaxiv.org
- Automatic Formal Verification for Code Generationlogicalintelligence.com
- When AI Writes the World's Software, Who Verifies It? — Leonardo de Mouraleodemoura.github.io
- Human Judgment as a Specificationblog.brownplt.org
- Ten AI safety projects I'd like people to work onthirdthing.ai
- AI Control: Improving Safety Despite Intentional Subversion — LessWronglesswrong.com