flâneur — a map of the web's best reading

AI Evaluations — ML Alignment & Theory Scholars

matsprogram.org · saved by 1 readers

Many stories of AI accident and misuse involve potentially dangerous capabilities, such as sophisticated deception and situational awareness, that have not yet been demonstrated in AI. Can we evaluate such capabilities in existing AI systems to form a foundation for policy and further technical work?

Explore this link on the map →

saved by