AI4GOOD @ NeurIPS 2026 | Trustworthy AI for Good
Sponsors: We're open for sponsorship with PR benefits! Please email zjingchen@cs.toronto.edu directly! Agentic AI systems increasingly shape how billions of people engage with public institutions, civic discourse, and society at large. While much work has focused on making models safer in avoiding harmful output, it is equally important for these improvements to translate into social good at scale. The AI4GOOD workshop brings together the AI safety, AI for social good, and AI policy/governance communities to connect what models can do as individual systems with what they do when deployed across populations. We aim to bridge technical advances in trustworthy AI with real-world societal impact, including protecting democratic institutions and civic discourse. We welcome submissions in a wide range of topics (if you're not sure about your paper, we encourage you to just submit!). We invite work across the following areas. Trustworthy AI models: evaluation, auditing, and red-teaming of mod
AI4GOOD @ NeurIPS 2026 | Trustworthy AI for Good Workshop Trustworthy AI for Good @ NeurIPS 2026 Date: TBD | Location: Paris, France I'd Like to Be a Reviewer: Register Here! View ICML 2026 Workshop Sponsors: We're open for sponsorship with PR benefits! Please email zjingchen@cs.toronto.edu directly! Workshop Overview Agentic AI systems increasingly shape how billions of people engage with public institutions, civic discourse, and society at large. While much work has focused on making models safer in avoiding harmful output, it is equally important for these improvements to translate into soc
Explore this link on the map →related reading
- Planning for AGI and beyond | OpenAIopenai.com
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- AI Control: Improving Safety Despite Intentional Subversion — LessWronglesswrong.com
- Patterns and problems in multiagent systemsanthropic.com
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Toward A Public Science of Model Behavior | Transluce AItransluce.org
- AI safety - Wikipediaen.wikipedia.org
- AI Control: Improving Safety Despite Intentional Subversion — AI Alignment Forumalignmentforum.org
- Ten AI safety projects I'd like people to work onthirdthing.ai
- AI Safety | Arkosevictoriabrook.github.io
- Field map – AISafety.comaisafety.com