✳flâneur — a map of the web's best reading
Can We Red Team Our Way to AI Accountability?
techpolicy.press · 1,171 words · saved by 1 readers
Red-teaming can only support AI accountability if laws, regulations, and enforcement to ensure protection from harm are in place.
Can We Red Team Our Way to AI Accountability? | TechPolicy.Press Can We Red Team Our Way to AI Accountability? Ranjit Singh, Borhane Blili-Hamelin, Jacob Metcalf / Aug 18, 2023 Ranjit Singh is a qualitative researcher with the Algorithmic Impact Methods Lab at Data & Society; Borhane Blili-Hamelin is the taxonomy lead at the AI Vulnerability Database, a community partner organization for the Generative Red Team (GRT) challenge at DEF CON ; and Jacob Metcalf is program director, AI on the Ground, at Data & Society. Shutterstock Republish Share Last week’s much publicized Generative Red Team (GR
Explore this link on the map →related reading
- Democratizing Generative AI Red Teams | Andreessen Horowitza16z.com
- Mediumai-alignment.com
- Defining LLM Red Teaming | NVIDIA Technical Blogdeveloper.nvidia.com
- [2004.07213] Toward Trustworthy AI Development: Mechanisms for Supporting Verifiable Claimsarxiv.org
- AI in 2025: gestalt — LessWronglesswrong.com
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- [2312.06942] AI Control: Improving Safety Despite Intentional Subversionarxiv.org
- 2312.06942arxiv.org
- AI Control: Improving Safety Despite Intentional Subversion — LessWronglesswrong.com
- Challenges in evaluating AI systems \ Anthropicanthropic.com
- AI safety - Wikipediaen.wikipedia.org
- [2402.04249] HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusalarxiv.org