My AGI safety research—2024 review, ’25 plans — AI Alignment Forum
“Our greatest fear should not be of failure, but of succeeding at something that doesn't really matter.” –attributed to DL Moody (copied almost word-for-word from Neuroscience of human social instincts: a sketch) My primary neuroscience research goal for the past couple years has been to solve a certain problem, a problem which has had me stumped since the very beginning of when I became interested in neuroscience at all (as a lens into Artificial General Intelligence safety) back in 2019. What is this grand problem? As described in Intro to Brain-Like-AGI Safety, I believe the following: It was a banner year! Basically, for years, I’ve had a vague idea about how human social instincts might work, involving what I call “transient empathetic simulations”. But I didn’t know how to pin it down in more detail than that. One subproblem was: I didn’t have even one example of a specific social instinct based on this putative mechanism—i.e., a hypothesis where a specific innate reaction would
x My AGI safety research—2024 review, ’25 plans — AI Alignment Forum Research Agendas AI Frontpage 45 My AGI safety research—2024 review, ’25 plans by Steven Byrnes 31st Dec 2024 10 min read 5 45 Previous: My AGI safety research—2022 review, ’23 plans . (I guess I skipped it last year.) “Our greatest fear should not be of failure, but of succeeding at something that doesn't really matter.” – attributed to DL Moody Tl;dr Section 1 goes through my main research project, “reverse-engineering human social instincts”: what does that even mean, what’s the path-to-impact, what progress did I make in
Explore this link on the map →related reading
- AGI safetysjbyrnes.com
- My AGI safety research—2025 review, ’26 plans — AI Alignment Forumalignmentforum.org
- AI 2027ai-2027.com
- My take on Jacob Cannell’s take on AGI safety — LessWronglesswrong.com
- LessWronglesswrong.com
- AGI Ruin: A List of Lethalities — LessWronglesswrong.com
- Alignment remains a hard, unsolved problem — LessWronglesswrong.com
- [Intro to brain-like-AGI safety] 1. What's the problem & Why work on it now? — AI Alignment Forumalignmentforum.org
- AI 2027ai-2027.com
- Commentary on AGI Safety from First Principles — AI Alignment Forumalignmentforum.org
- Planning for AGI and beyond | OpenAIopenai.com
- Neurotechnology is Critical for AI Alignmentmilan.cvitkovic.net