AI4GOOD @ NeurIPS 2026 | Trustworthy AI for Good
Sponsors: We're open for sponsorship with PR benefits! Please email zjingchen@cs.toronto.edu directly! Agentic AI systems increasingly shape how billions of people engage with public institutions, civic discourse, and society at large. While much work has focused on making models safer in avoiding harmful output, it is equally important for these improvements to translate into social good at scale. The AI4GOOD workshop brings together the AI safety, AI for social good, and AI policy/governance communities to connect what models can do as individual systems with what they do when deployed across populations. We aim to bridge technical advances in trustworthy AI with real-world societal impact, including protecting democratic institutions and civic discourse. We welcome submissions in a wide range of topics (if you're not sure about your paper, we encourage you to just submit!). We invite work across the following areas. Trustworthy AI models: evaluation, auditing, and red-teaming of mod
AI4GOOD @ NeurIPS 2026 | Trustworthy AI for Good Workshop Trustworthy AI for Good @ NeurIPS 2026 Date: TBD | Location: Paris, France I'd Like to Be a Reviewer: Register Here! View ICML 2026 Workshop Sponsors: We're open for sponsorship with PR benefits! Please email zjingchen@cs.toronto.edu directly! Workshop Overview Agentic AI systems increasingly shape how billions of people engage with public institutions, civic discourse, and society at large. While much work has focused on making models safer in avoiding harmful output, it is equally important for these improvements to translate into soc
Explore this link on the map →related reading
- Planning for AGI and beyond | OpenAIopenai.com
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- AI Control: Improving Safety Despite Intentional Subversion — LessWronglesswrong.com
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- AI safety - Wikipediaen.wikipedia.org
- AI Safety | Arkosevictoriabrook.github.io
- Ten AI safety projects I'd like people to work on — LessWronglesswrong.com
- AI Control: Improving Safety Despite Intentional Subversion — AI Alignment Forumalignmentforum.org
- Spring 2026 Projects - SPARsparai.org
- AI Alignment Cannot Be Top-Down | AI Frontiersai-frontiers.org
- [2004.07213] Toward Trustworthy AI Development: Mechanisms for Supporting Verifiable Claimsarxiv.org