flâneur — a map of the web's best reading

Aligning to Virtues — LessWrong

lesswrong.com · 7,795 words · saved by 1 readers

Which alignment target? Suppose you’re an AI company or government, and you want to figure out what values to align your AI to. Here are three option…

x Aligning to Virtues — LessWrong AI Frontpage 93 Aligning to Virtues by Richard_Ngo 16th Feb 2026 5 min read 36 93 Which alignment target? Suppose you’re an AI company or government, and you want to figure out what values to align your AI to. Here are three options, and some of their downsides: AIs that are aligned to a set of consequentialist values are incentivized to acquire power to pursue those values. This creates power struggles between those AIs and: Humans who don’t share those values. Humans who disagree with the AI about how to pursue those values. Humans who don’t trust that the A

Explore this link on the map →

related reading