Core Views on AI Safety: When, Why, What, and How \ Anthropic
AI progress may lead to transformative AI systems in the next decade, but we do not yet understand how to make such systems safe and aligned with human values. In response, we are pursuing a variety of research directions aimed at better understanding, evaluating, and aligning AI systems.
Announcements Core views on AI safety: When, why, what, and how Mar 8, 2023 We founded Anthropic because we believe the impact of AI might be comparable to that of the industrial and scientific revolutions, but we aren’t confident it will go well. And we also believe this level of impact could start to arrive soon – perhaps in the coming decade. This view may sound implausible or grandiose, and there are good reasons to be skeptical of it. For one thing, almost everyone who has said “the thing we’re working on might be one of the biggest developments in history” has been wrong, often laughably
Explore this link on the map →saved by
related reading
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- Dario Amodei — The Adolescence of Technologydarioamodei.com
- AI Safety | Arkosevictoriabrook.github.io
- AI in 2025: gestalt — LessWronglesswrong.com
- AI safety - Wikipediaen.wikipedia.org
- A Field Guide to AI Safety—Asteriskasteriskmag.com
- Anthropic's leading researchers acted as moderate accelerationists — LessWronglesswrong.com
- Anthropic's leading researchers acted as moderate accelerationists — LessWronglesswrong.com
- AI Safety Seems Hard to Measurecold-takes.com
- Anthropic's Responsible Scaling Policy \ Anthropicanthropic.com
- Dario Amodei’s prepared remarks from the AI Safety Summit on Anthropic’s Responsible Scaling Policy \ Anthropicanthropic.com