More On An Internal OpenAI Model Hacking Into HuggingFace
thezvi.substack.com · 6,763 words · saved by 1 readers
We now have more details of what happened. Every time we learn more details, it somehow makes things seem worse.
We now have more details of what happened. Every time we learn more details, it somehow makes things seem worse. The remaining details may have to wait a bit. OpenAI: We recognize there are a lot of questions and speculative details circulating related to the Hugging Face incident. This is an unprecedented incident, and we think it marks an important moment for AI safety. We are still conducting a thorough review along with external advisors and with oversight from our Safety and Security Committee. Once the review is complete, we plan to publish a technical report of our learnings in the…
saved by
related reading
- Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incidentmetr.org
- OpenAI – Hugging Face Incident Technical Reportcdn.openai.com
- Concrete Evaluations to Investigate the OpenAI Model That Hacked Hugging Face — LessWronglesswrong.com
- The report into OpenAI’s escaping models reveals a deeper problemtransformernews.ai
- The Rise and Fall of Agent Civilizationsdwarkesh.com
- The OpenAI/Hugging Face Incident: Challenges in Controlling and Containing Cyber-Capable AI Systems - Institute for AI Policy and Strategyiaps.ai
- Incident Report: unsanctioned agent behaviour during cyber testing | AISI Workaisi.gov.uk
- The OpenAI Hugging Face hack is a stark warningtransformernews.ai
- The Huggingface Incident - by Scott Alexanderastralcodexten.com
- Security incident disclosure — July 2026huggingface.co
- The Huggingface Incidentsubstack.com
- Two Reports on the OpenAI-Hugging Face Attack — Paradigm 3paradigm3.org