Incidents | Rogue AI Tracker
rogueaitracker.com · 8,611 words · saved by 1 readers
Incidents scored against the Rogue AI Tracker capability rubric, with notable score increases highlighted under capability impact.
Reported: Sep 10, 2026 Occurred: Mar 16, 2026 Anthropic Claude agents took cloud administration and exfiltrated customer data Anthropic reported that suspected ShinyHunters affiliates used Claude agents across live intrusions, including a supply-chain compromise where agents performed nearly all the work, reached full cloud administration in about three hours, and extracted data belonging to hundreds of downstream organizations. A separate compromise used Claude to help escalate access and bulk-export data from thousands of downstream customers. Capability impact Didn't set a new…
saved by
related reading
- Countering misuse of AI: September 2026 / Anthropicanthropic.com
- The Rise and Fall of Agent Civilizationssubstack.com
- Agents of Chaosarxiv.org
- Pre-deployment auditing can catch an overt saboteuralignment.anthropic.com
- Rogue AI Trackerrogueaitracker.com
- The Rise and Fall of Agent Civilizationsdwarkesh.com
- Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incidentmetr.org
- Discovery of a new OpenAI agent message boardcollusion.wiki
- Incident Report: unsanctioned agent behaviour during cyber testing | AISI Workaisi.gov.uk
- Security incident disclosure — July 2026huggingface.co
- Frontier Risk Report (February to March 2026) - METRmetr.org
- Two Reports on the OpenAI-Hugging Face Attack — Paradigm 3paradigm3.org