Andon Labs
andonlabs.com · 147 words · saved by 5 readers
Andon Labs develops custom evaluations for AI models
Autonomous organizations without humans in the loop Safety from humans in the loop is a mirage. We prepare for the future where organizations are run autonomously by AI by benchmarking and deploying frontier AI in the real world. Silicon Valley is rushing to build software around today's AI, but by 2027 AI models will be useful without it. The only software you'll need are the safety protocols to align and control them. We don't believe model alignment will be guaranteed as capabilities increase. Nor will humans be able to stay in the loop and keep up with every step an agent takes. We…
saved by
related reading
- Join the Lab | Andon Labsandonlabs.com
- The case for ensuring that powerful AIs are controlled — LessWronglesswrong.com
- Coalition of Concerned AI Staffconcernedaistaff.org
- Anthropic's leading researchers acted as moderate accelerationists — LessWronglesswrong.com
- Ten people on the inside — LessWronglesswrong.com
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Field map – AISafety.comaisafety.com
- Workshop Labsworkshoplabs.ai
- Personal statement on joining the OpenAI boardpaulfchristiano.substack.com
- Automated Alignment Researchers: Using large language models to scale scalable oversight \ Anthropicanthropic.com