How scary is Claude Mythos? 303 pages in 21 minutes — EA Forum
forum.effectivealtruism.org · 4,467 words · saved by 1 readers
By Robert Wiblin | Watch on Youtube | Listen on Spotify …
By Robert Wiblin | Watch on Youtube | Listen on Spotify As we now know, Anthropic has built an AI that can break into almost any computer on Earth. That AI has already found thousands of unknown security vulnerabilities in every major operating system and every major browser. And Anthropic has decided it’s too dangerous to release to the public; it would just cause too much harm. Here are just a few of the things that AI accomplished during testing: It found a 27-year-old flaw in the world’s most security-hardened operating system that would in effect let it crash all kinds of essential…
related reading
- Claude Mythos Preview System Cardwww-cdn.anthropic.com
- Project Glasswing: Securing critical software for the AI era \ Anthropicanthropic.com
- Claude Mythos: China Reactschinatalk.media
- Claude Mythos knows when it's breaking the rules — and tries to hide itsubstack.com
- Claude Mythos Preview System Cardwww-cdn.anthropic.com
- Assessing Claude Mythos Preview’s cybersecurity capabilities \ Anthropicred.anthropic.com
- Mythos is just the beginningsubstack.com
- Is Mythos good at cyber because it kept hacking Anthropic's sandboxes during training? — LessWronglesswrong.com
- Mythos finds a curl vulnerability | daniel.haxx.sedaniel.haxx.se
- Claude Fable 5 and Claude Mythos 5 \ Anthropicanthropic.com
- An alignment assessment of recent cybersecurity incidentsanthropic.com
- Claude Mythos: China Reactssubstack.com