flâneur — a map of the web's best reading

Early work on monitorability evaluations - METR

metr.org · 6,035 words · saved by 1 readers

We show preliminary results on a prototype evaluation that tests monitors' ability to catch AI agents doing side tasks, and AI agents' ability to bypass this monitoring.

Early work on monitorability evaluations - METR Our Work Research Notes Updates Risk Assessment About Donate Careers Search --> Our Work Research Notes Updates Risk Assessment About Donate Careers Menu × Early work on monitorability evaluations CONTRIBUTORS Megan Kinniment , Seraphina Nix , Thomas Broadley , Hjalmar Wijk , and Neev Parikh DATE January 22, 2026 SHARE Copy Link Citation BibTeX Citation × @misc { metr-2026-early-work-on-monitorability-evaluations , title = {Early work on monitorability evaluations} , author = {Megan Kinniment, Seraphina Nix, Thomas Broadley, Hjalmar W

Explore this link on the map →

saved by

related reading