Understand, align, cooperate: AI welfare and AI safety are allies
experiencemachines.substack.com · 2,455 words · saved by 1 readers
Win-win solutions and low-hanging fruit
In conversations about AI safety and AI welfare, people sometimes frame them as inherently opposing goals: either we protect humans (AI safety), or we protect AI systems (AI welfare). This framing implies a tragic, necessary choice: whose side are you on? Or it suggests, at least, “shouldn’t we solve safety first, and then worry about AI welfare?” But the truth is that this is a false choice: fortunately, many of the best interventions for protecting AI systems will also protect humans, and vice versa. To be sure, there are genuine tensions between AI safety and AI welfare.1 But if we only…
saved by
related reading
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- Core views on AI safety: When, why, what, and how \ Anthropicanthropic.com
- The Artificiality of Alignmentjoinreboot.org
- You Need A Theory of Victory - by Jason Hausenloyfirstscattering.com
- Current AIs seem pretty misaligned to me — LessWronglesswrong.com
- [2605.10310] Positive Alignment: Artificial Intelligence for Human Flourishingarxiv.org
- Where I agree and disagree with Eliezer — LessWronglesswrong.com
- Recommendations for Technical AI Safety Research Directionsalignment.anthropic.com
- What is AI alignment? - by Adam Jones - BlueDot Impactblog.bluedot.org
- AI safety - Wikipediaen.wikipedia.org
- We should take AI welfare seriously - by Robert Longexperiencemachines.substack.com
- You Need A Theory of Victorysubstack.com