flâneur — a map of the web's best reading

Open Sourcing Monitorability Evaluations

alignment.openai.com · 2,221 words · saved by 1 readers

We open-source datasets and code from our Monitoring Monitorability paper, and share a new filtering strategy for noise-dominated intervention evaluation instances.

Open Sourcing Monitorability Evaluations ← Back to OpenAI Alignment Blog Open Sourcing Monitorability Evaluations Apr 23, 2026 · Melody Y. Guan, Miles Wang, Micah Carroll, Zehao Dou, Annie Y. Wei, Marcus Williams, Benjamin Arnav, Joost Huizinga, Ian Kivlichan, Mia Glaese, Jakub Pachocki, Bowen Baker TL;DR We are releasing a subset of datasets and reference code from our chain-of-thought monitorability work . The release includes most datasets from our monitorability evaluation suite, code for computing the monitorability metric g-mean 2 , and a new cross-fit filtering strategy that makes

Explore this link on the map →

saved by

related reading