✳flâneur — a map of the web's best reading
Why multi-agent safety is important — LessWrong
lesswrong.com · 3,455 words · saved by 1 readers
Alternative Title: Mo’ Agents, Mo’ Problems …
x Why multi-agent safety is important — LessWrong Multipolar Scenarios AI Frontpage 10 Why multi-agent safety is important by Akbir Khan 14th Jun 2022 12 min read 2 10 Alternative Title: Mo’ Agents, Mo’ Problems This was a semi- adversarial collaboration with Robert Kirk; views are mostly Akbir’s, clarified with Rob’s critiques. Target Audience: Machine Learning researchers who are interested in Safety and researchers who are focusing on Single Agent Safety problems. Context: Recently I’ve been discussing how to meaningfully de-risk AGI technologies. My argument is that even in the situation w
Explore this link on the map →related reading
- Multi-agent safety — AI Alignment Forumalignmentforum.org
- [2502.14143] Multi-Agent Risks from Advanced AIarxiv.org
- AGI Ruin: A List of Lethalities — LessWronglesswrong.com
- [2605.22748] Superhuman Safe and Agile Racing through Multi-Agent Reinforcement Learningarxiv.org
- [2602.12316] GT-HarmBench: Benchmarking AI Safety Risks Through the Lens of Game Theoryarxiv.org
- ROGUE:arxiv.org
- Red-teaming a network of agents: Understanding what breaks when AI agents interact at scale - Microsoft Researchmicrosoft.com
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- [AN #70]: Agents that help humans who are still learning about their own preferences — LessWronglesswrong.com
- How we built our multi-agent research system \ Anthropicanthropic.com
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Solipsistic Superintelligence is Unlikely to be Cooperativearxiv.org