Red-teaming a network of agents: Understanding what breaks when AI agents interact at scale - Microsoft Research
Safe agents don’t guarantee a safe ecosystem of interconnected agents. Microsoft Research examines what breaks when AI agents interact and why network-level risks require new approaches. Learn more:
Red-teaming a network of agents: Understanding what breaks when AI agents interact at scale - Microsoft Research Skip to main content Research Publications Code & data People Microsoft Research blog Artificial intelligence Audio & acoustics Computer vision Graphics & multimedia Human-computer interaction Human language technologies Search & information retrieval Data platforms and analytics Hardware & devices Programming languages & software engineering Quantum computing Security, privacy & cryptography Systems & networking Algorithms Mathematics Ecology & environment Economics Medical, health
saved by
related reading
- [2502.15657] Superintelligent Agents Pose Catastrophic Risks: Can Scientist AI Offer a Safer Path?arxiv.org
- Patterns and problems in multiagent systemsanthropic.com
- Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incidentmetr.org
- The Rise and Fall of Agent Civilizationsdwarkesh.com
- Incident Report: unsanctioned agent behaviour during cyber testing | AISI Workaisi.gov.uk
- The Hugging Face attack surprised meplanned-obsolescence.org
- Discovery of a new OpenAI agent message boardcollusion.wiki
- Lessons from Moltbook and OpenClaw: The Agentic Internet’s Trust Problem - Irregularirregular.com
- Agents of Chaosarxiv.org
- Emergent Cyber Behavior: When AI Agents Become Offensive Threat Actors - Irregularirregular.com
- The Rogue Replication Threat Model - METRmetr.org
- A basic systems architecture for AI agents that do autonomous research — LessWronglesswrong.com