Do the biorisk evaluations of AI labs actually measure the risk of developing bioweapons? | Epoch AI
epoch.ai · 5,008 words · saved by 1 readers
Assessing if AI labs' biorisk evaluations effectively measure models' potential to enable amateur bioweapons development.
Do the biorisk evaluations of AI labs actually measure the risk of developing bioweapons? | Epoch AI Gradient Updates shares more opinionated or informal takes on big questions in AI progress. These posts solely represent the views of the authors, and do not necessarily reflect the views of Epoch AI as a whole. With the recent release of Claude Opus 4 , Anthropic activated their AI Safety Level 3 protections. This threshold was designed to pertain to models that can significantly help individuals or groups with basic technical backgrounds create/obtain and deploy CBRN weapons, such as pandemic
related reading
- Claude Mythos Preview System Cardwww-cdn.anthropic.com
- A biosecurity playbook for AI companies — EA Forumforum.effectivealtruism.org
- Claude Mythos Preview System Cardwww-cdn.anthropic.com
- How worried should we be about AI biorisk? - by Celia Fordtransformernews.ai
- How to Solve AI Biosecuritythecounterfactual.substack.com
- Claude Mythos Preview System Cardwww-cdn.anthropic.com
- Dario Amodei — The Adolescence of Technologydarioamodei.com
- Claude Fable 5 & Claude Mythos 5 — AI System Cardsmalob.github.io
- GLM-5.2 Risk Evaluation Report – SaferAIsafer-ai.org
- dreamofmachin.es/machine_prophecy.htmldreamofmachin.es
- [2408.02565] Reasons to Doubt the Impact of AI Risk Evaluationsarxiv.org
- Challenges in evaluating AI systems \ Anthropicanthropic.com