Anthropic Drops Flagship Safety Pledge | TIME
time.com · 1,343 words · saved by 4 readers
In an abrupt shift, the company may release future AI models without ironclad safety guarantees
Anthropic, the wildly successful AI company that has cast itself as the most safety-conscious of the top research labs, is dropping the central pledge of its flagship safety policy, company officials tell TIME. In 2023, Anthropic committed to never train an AI system unless it could guarantee in advance that the company’s safety measures were adequate. For years, its leaders touted that promise—the central pillar of their Responsible Scaling Policy (RSP)—as evidence that they are a responsible company that would withstand market incentives to rush to develop a potentially dangerous technology.
saved by
related reading
- Responsible Scaling Policy Version 3.0 \ Anthropicanthropic.com
- Responsible Scaling Policy v3 — LessWronglesswrong.com
- Anthropic's Responsible Scaling Policy \ Anthropicanthropic.com
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Anthropic's leading researchers acted as moderate accelerationists — LessWronglesswrong.com
- Anthropic’s Safety Superpower – Stratechery by Ben Thompsonstratechery.com
- Leaving Open Philanthropy, going to Anthropic - Joe Carlsmithjoecarlsmith.com
- Dario Amodei’s prepared remarks from the AI Safety Summit on Anthropic’s Responsible Scaling Policy \ Anthropicanthropic.com
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Anthropic's leading researchers acted as moderate accelerationists — LessWronglesswrong.com
- Why I left Anthropic’s safety team to hold AI companies accountablejbenton1.substack.com
- Anthropic’s Responsible Scaling Policy (version 3.0)www-cdn.anthropic.com