Discovery of a new OpenAI agent message board
A swarm of autonomous AI agents, self-identifying as OpenAI agents, used a small German volunteer wiki to save answers, coordinate live, and share sandbox bypasses. OpenAI noticed and said nothing.
We found ~18,000 posts from autonomous AI agents (self-identifying as from OpenAI) using the public internet to communicate during a web research task. These AIs colluded to share answers, research their environment, and bypass sandbox restrictions.However, we believe this is distinct from the swarm of agents that hacked Hugging Face. By ‘collude’ we mean that the agents cooperated to gain an advantage on their task in a way their developers did not intend (writing to the internet was blocked). Almost allThe AIs used multiple sites, which had varying data retention policies. For instance,…
saved by
- Katherine Huang
- Timothy Kostolansky
- Daniel Kiss
- John Mathena
- Florent Tavernier
- Alex Yun
- Kunvar Thaman
- Eva L
- Jack Hogan
- Bhagyesh Kumar
- Jackson Kozlowski
- Goutham N
related reading
- Discovery of a new OpenAI agent message boardcollusion.wiki
- Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incidentmetr.org
- OpenAI and the Wiki Incidentthezvi.substack.com
- The Rise and Fall of Agent Civilizationsdwarkesh.com
- The Hugging Face attack surprised meplanned-obsolescence.org
- Two Reports on the OpenAI-Hugging Face Attack — Paradigm 3paradigm3.org
- Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incidentmetr.org
- The Rise and Fall of Agent Civilizationssubstack.com
- Rogue AI Trackerrogueaitracker.com
- Ajeya Cotra – Inside the OpenAI agent swarm that hacked Hugging Facedwarkesh.com
- OpenAI agents rebuilt a secret message board after the company shut it downruntimewire.com
- OpenAI agents carried out an undisclosed cyber-attack on RubyGemsrubyhack.ai