flâneur

AI Evaluations — ML Alignment & Theory Scholars

matsprogram.org · saved by 1 readers

Many stories of AI accident and misuse involve potentially dangerous capabilities, such as sophisticated deception and situational awareness, that have not yet been demonstrated in AI. Can we evaluate such capabilities in existing AI systems to form a foundation for policy and further technical work?

saved by