we have a year to fix security everywhere
jyn.dev · 2,738 words · saved by 1 readers
consumer-grade hardware can run an LLM that hacks the planet. we can stop it, but we don't have much time.
GLM 5.3-flash released last week, and that means Project Glasswing and Daybreak are running out of time. Cheap models capable of dangerous hacking are now available to anyone, without the normal safeguards for refusing malicious actions. We need to fix vulnerabilities across the industry so that we aren't caught unawares. And for one of the first times in computing history, we have the ability to! We can use frontier LLMs that move faster than a human to find and fix these issues in the time we have left. The hard remaining part is deploying the fixes. This probably sounds like nonsense…
saved by
related reading
- Project Glasswing: Securing critical software for the AI era \ Anthropicanthropic.com
- Security incident disclosure — July 2026huggingface.co
- AI in 2025: gestalt — LessWronglesswrong.com
- What I learned this week - Can distillation be stopped, Mythos and the cybersecurity equilibrium, Pipeline RLdwarkesh.com
- The End-State Fallacy: Where Is AI Security Headed?endstatefallacy.com
- 2025: The year in LLMssimonwillison.net
- Nicholas Carlininicholas.carlini.com
- Measuring LLMs' impact on N-day exploits \ Anthropicred.anthropic.com
- Greg Brockman on Svbtleblog.gregbrockman.com
- Vulnerability Research Is Cooked - Quarrelsomesockpuppet.org
- My self-sovereign / local / private / secure LLM setup, April 2026vitalik.eth.limo
- How fast is AI improving? - AI Digesttheaidigest.org