Can We Red Team Our Way to AI Accountability?
techpolicy.press · 1,171 words · saved by 1 readers
Red-teaming can only support AI accountability if laws, regulations, and enforcement to ensure protection from harm are in place.
Can We Red Team Our Way to AI Accountability? | TechPolicy.Press Can We Red Team Our Way to AI Accountability? Ranjit Singh, Borhane Blili-Hamelin, Jacob Metcalf / Aug 18, 2023 Ranjit Singh is a qualitative researcher with the Algorithmic Impact Methods Lab at Data & Society; Borhane Blili-Hamelin is the taxonomy lead at the AI Vulnerability Database, a community partner organization for the Generative Red Team (GRT) challenge at DEF CON ; and Jacob Metcalf is program director, AI on the Ground, at Data & Society. Shutterstock Republish Share Last week’s much publicized Generative Red Team (GR
related reading
- Recommendations-for-Using-Red-Teaming-for-AI-Accountability-PolicyBrief.pdfdatasociety.net
- Democratizing Generative AI Red Teams | Andreessen Horowitza16z.com
- GitHub - requie/AI-Red-Teaming-Guide: A comprehensive guide to adversarial testing and security evaluation of AI systems, helping organizations identify vulnerabilities before attackers exploit them.github.com
- Mediumai-alignment.com
- Defining LLM Red Teaming | NVIDIA Technical Blogdeveloper.nvidia.com
- [2004.07213] Toward Trustworthy AI Development: Mechanisms for Supporting Verifiable Claimsarxiv.org
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- [2312.06942] AI Control: Improving Safety Despite Intentional Subversionarxiv.org
- A safe harbor for AI evaluation and red teamingnormaltech.ai
- AI Control: Improving Safety Despite Intentional Subversion — LessWronglesswrong.com
- AI safety is not a model propertysubstack.com
- 2312.06942arxiv.org