How much novel security-critical infrastructure do you need during the singularity? — AI Alignment Forum
I think a lot about the possibility of huge numbers of AI agents doing AI R&D inside an AI company (as depicted in AI 2027). I think particularly about what will happen if those AIs are scheming: coherently and carefully trying to grab power and take over the AI company, as a prelude to taking over the world. And even more particularly, I think about how we might try to mitigate the insider risk posed by these AIs, taking inspiration from traditional computer security, traditional insider threat prevention techniques, and first-principles thinking about the security opportunities posed by the differences between AIs and humans. So to flesh out this situation, I’m imagining a situation something like AI 2027 forecasts for March 2027: I want to know what we need to do to ensure that even if those AIs are conspiring against us, they aren’t able to cause catastrophic security failures.[1] One important question here is how strongly incentivized we are to use the AIs to write security-criti
x How much novel security-critical infrastructure do you need during the singularity? — AI Alignment Forum AI Takeoff Computer Security & Cryptography AI Frontpage 31 How much novel security-critical infrastructure do you need during the singularity? by Buck 4th Jul 2025 6 min read 7 31 I think a lot about the possibility of huge numbers of AI agents doing AI R&D inside an AI company (as depicted in AI 2027). I think particularly about what will happen if those AIs are scheming: coherently and carefully trying to grab power and take over the AI company, as a prelude to taking over the world. A
Explore this link on the map →related reading
- AI 2027ai-2027.com
- Dario Amodei — The Adolescence of Technologydarioamodei.com
- Project Glasswing: Securing critical software for the AI era \ Anthropicanthropic.com
- Nova DasSarma on why information security may be critical to the safe development of AI systems | 80,000 Hours80000hours.org
- AI 2027ai-2027.com
- Quotes from Leopold Aschenbrenner's Situational Awareness Paperthezvi.substack.com
- Fields that I reference when thinking about AI takeover prevention — LessWronglesswrong.com
- A basic systems architecture for AI agents that do autonomous research — LessWronglesswrong.com
- Inside Stargateopenaiglobalaffairs.substack.com
- My picture of the present in AI — LessWronglesswrong.com
- Planning for Extreme AI Risks — AI Alignment Forumalignmentforum.org
- Cybersecurity Looks Like Proof of Work Nowdbreunig.com