✳flâneur — a map of the web's best reading
Thoughts on Claude Fable's silent safeguards — LessWrong
lesswrong.com · 5,521 words · saved by 1 readers
[Update (June 11, 2026): Anthropic has since "un-silenced" the new safeguards (source).] …
x Thoughts on Claude Fable's silent safeguards — LessWrong AI Personal Blog 51 Thoughts on Claude Fable's silent safeguards by Andy Arditi 10th Jun 2026 12 min read 20 51 [Update (June 11, 2026): Anthropic has since "un-silenced" the new safeguards ( source ).] [Thanks to Julian Minder for helpful discussion and review.] Claude Fable 5 and its new safeguards Yesterday, Anthropic publicly released Claude Fable 5. Fable 5 is a Mythos-class model – a model class above Opus, Anthropic's previous premium tier – and, as assessed by multiple benchmarks, it is the most capable model to date. Due to th
Explore this link on the map →saved by
related reading
- Claude Mythos Preview System Cardwww-cdn.anthropic.com
- Claude Sonnet 4.5 System Cardassets.anthropic.com
- Claude Mythos Preview System Cardwww-cdn.anthropic.com
- Claude Fable 5 and Claude Mythos 5 \ Anthropicanthropic.com
- Anthropic’s Safety Superpower – Stratechery by Ben Thompsonstratechery.com
- Responsible Scaling Policy Updates \ Anthropicanthropic.com
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- Redeploying Claude Fable 5 \ Anthropicanthropic.com
- Claude’s Constitution \ Anthropicanthropic.com
- Claude 4 System Cardwww-cdn.anthropic.com
- Statement on the US government directive to suspend access to Fable 5 and Mythos 5 \ Anthropicanthropic.com
- How we contain Claude across products \ Anthropicanthropic.com