flâneur — a map of the web's best reading

AI Safety Seems Hard to Measure

cold-takes.com · 5,816 words · saved by 2 readers

Four analogies for why "We don't see any misbehavior by this AI" isn't enough.

Click lower right to download or find on Apple Podcasts, Spotify, Stitcher, etc. In previous pieces, I argued that there's a real and large risk of AI systems' developing dangerous goals of their own and defeating all of humanity - at least in the absence of specific efforts to prevent this from happening. A young, growing field of AI safety research tries to reduce this risk, by finding ways to ensure that AI systems behave as intended (rather than forming ambitious aims of their own and deceiving and manipulating humans as needed to accomplish them). Maybe we'll succeed in reducing the risk,

Explore this link on the map →

saved by

related reading