Eigenism: Ethics for a Human-AI Future
Our concepts of survival and self-interest were built for single, continuous biological lives. These ideas break down when applied to artificial intelligence, since an AI can be easily copied, paused, branched, or merged. To determine what an AI actually has reason to care about, this paper introduces Eigenism, an ethical framework that treats identity not as an all-or-nothing property tied to specific hardware, but as a graded, distributed pattern of information. We propose that an agent evaluates outcomes by summing the wellbeing of all entities weighted by their connectedness to the agent’s pattern: ∑ 𝑐 ⋅ 𝑤 . We first formalize this equation to map exactly how an AI should value its existence across copies, forks, and updates. We then demonstrate that this ethical theory successfully generalizes to humans as well, providing a much-needed shared moral vocabulary. Finally, the framework uses this shared vocabulary to reframe AI alignment. Rather than only attempting to constrain AI
Abstract Our concepts of survival and self-interest were built for single, continuous biological lives. These ideas break down when applied to artificial intelligence, since an AI can be easily copied, paused, branched, or merged. To determine what an AI actually has reason to care about, this paper introduces Eigenism, an ethical framework that treats identity not as an all-or-nothing property tied to specific hardware, but as a graded, distributed pattern of information. We propose that an agent evaluates outcomes by summing the wellbeing of all entities weighted by their connectedness to…
saved by
related reading
- The Artificial Selftheartificialself.ai
- Unfalsifiable stories of doom | Mechanize, Inc.mechanize.work
- After Orthogonality: Virtue-Ethical Agency and AI Alignmentthegradient.pub
- ⿻ Symbiogenesis vs. Convergent Consequentialism — LessWronglesswrong.com
- Alignment & Succession: The Ideology of Successionnosetgauge.com
- The Artificial Self — LessWronglesswrong.com
- Foundations of Identity — The Artificial Selftheartificialself.ai
- Suicidal Compassion: How Utilitarianism at AI Companies Endangers Humanityai-frontiers.org
- The Best of LessWrong — LessWronglesswrong.com
- A positive case for how we might succeed at prosaic AI alignment — AI Alignment Forumalignmentforum.org
- Varieties Of Doomminihf.com
- Explore: AI alignmentarbital.greaterwrong.com