flâneur — a map of the web's best reading

Petri: An open-source auditing tool to accelerate AI safety research \ Anthropic

anthropic.com · 1,371 words · saved by 1 readers

A new automated auditing tool for AI safety research

Alignment Petri: An open-source auditing tool to accelerate AI safety research Oct 6, 2025 Read the technical report Petri (Parallel Exploration Tool for Risky Interactions) is our new open-source tool that enables researchers to explore hypotheses about model behavior with ease. Petri deploys an automated agent to test a target AI system through diverse multi-turn conversations involving simulated users and tools; Petri then scores and summarizes the target’s behavior. This automation handles a significant part of the work that one needs to do to build a broad understanding of a new model, an

Explore this link on the map →

related reading