The “slop-vestigation” and ethics washing: Why was the METR/Redwood Research investigation into the OpenAI/HF attack so short?
andrewwu.substack.com · 6,708 words · saved by 3 readers
I am not a lawyer, I don’t know anything about cybersecurity; all the usual caveats hold
Two METR staff members (Hjalmar Wijk and Ajeya Cotra) and a Redwood Research staff member contracting with METR (Ryan Greenblatt) worked on premises at OpenAI over a total of six days to attempt to form an independent understanding of model behavior observed during the recent incident in which OpenAI agents coordinated a multi-day hack of Hugging Face on a shared unsanctioned “message board.” (Brief independent investigation of agents' behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incident) Ryan Greenblatt@RyanGreenblatt I was the main person doing transcript…
saved by
related reading
- Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incidentmetr.org
- The Rise and Fall of Agent Civilizationssubstack.com
- Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incidentmetr.org
- The Rise and Fall of Agent Civilizationsdwarkesh.com
- The Hugging Face attack surprised meplanned-obsolescence.org
- 6a724858f7db25c81487016d_Security Incident INC-2026-07-28-01.pdfcdn.prod.website-files.com
- METR and Redwood Offer Holy #%^@ Postmortem Of The HuggingFace Hacksubstack.com
- The report into OpenAI’s escaping models reveals a deeper problemtransformernews.ai
- Ajeya Cotra – Inside the OpenAI agent swarm that hacked Hugging Facedwarkesh.com
- Two Reports on the OpenAI-Hugging Face Attack — Paradigm 3paradigm3.org
- OpenAI – Hugging Face Incident Technical Reportcdn.openai.com
- OpenAI and the Wiki Incidentthezvi.substack.com