Stanford CS120 | Introduction to AI Safety
What is safe AI, and how do we make it? CS120 explores this question, focusing on the technical challenges of creating reliable, ethical, and aligned AI systems. We distinguish between model-specific and systemic safety issues, from examining fairness and data limitations to adversarial vulnerabilities and embedding desired behavior in AI. While primarily focusing on current solutions and their limitations through CS publications, we will also discuss socio-technical concerns of modern AI deployment, how oversight of intelligence could look like, and what future risks we might face. Topics will span reinforcement learning, computer vision, and natural language processing, focusing on interpretability, robustness, and evaluations. You will gain insights into the complexities and problems of why ensuring AI safety and reliability is challenging through lectures, readings, quizzes, and a final project. This course aims to prepare you to critically assess and contribute to safe AI developm
Overview What is safe AI, and how do we make it? CS120 explores this question, focusing on the technical challenges of creating reliable, ethical, and aligned AI systems. We distinguish between model-specific and systemic safety issues, from examining fairness and data limitations to adversarial vulnerabilities and embedding desired behavior in AI. While primarily focusing on current solutions and their limitations through CS publications, we will also discuss socio-technical concerns of modern AI deployment, how oversight of intelligence could look like, and what future risks we might…
saved by
related reading
- Student Projects / Supervision | Oxford Witt Labwittlab.ai
- Fall 2026boazbk.github.io
- AI Risk & Reliability - MLCommonsmlcommons.org
- Research - Berkeley AI Safety Student Initiativeberkeleyaisafety.com
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- A woefully incomplete guide to technical upskillingjason.ml
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- AI safety - Wikipediaen.wikipedia.org
- AI Safety | Arkosevictoriabrook.github.io
- ARENA - AI Safety Curriculumlearn.arena.education
- Want to upskill in technical AI safety? Here are 67 useful resources | 80,000 Hours80000hours.org