flâneur — a map of the web's best reading

Safety cases at AISI | AISI Work

aisi.gov.uk · 1,560 words · saved by 1 readers

As a complement to our empirical evaluations of frontier AI models, AISI is planning a series of collaborations and research projects sketching safety cases for more advanced models than exist today, focusing on risks from loss of control and autonomy. By a safety case, we mean a structured argument that an AI system is safe within a particular training or deployment context.

Note to readers: we changed our name to the AI Security Institute on 14 February 2025. Read more here. Safety cases: structured arguments for frontier AI safety The benefits and risks posed by frontier AI models are changing over time as performance increases, requiring corresponding changes to how we mitigate those risks and the evidence required to trust those mitigations. Existing models lack the capabilities required to pose certain severe risks such as irreversible loss of human control, so arguments based on capabilities evaluations suffice. If future models are much more capable then di

Explore this link on the map →

related reading