✳flâneur — a map of the web's best reading
AI Evaluations — ML Alignment & Theory Scholars
matsprogram.org · saved by 1 readers
Many stories of AI accident and misuse involve potentially dangerous capabilities, such as sophisticated deception and situational awareness, that have not yet been demonstrated in AI. Can we evaluate such capabilities in existing AI systems to form a foundation for policy and further technical work?
Explore this link on the map →