flâneur — a map of the web's best reading

Aren’t developers regularly making their AIs nice and safe and obedient? | If Anyone Builds It, Everyone Dies | If Anyone Builds It, Everyone Dies

ifanyonebuildsit.com · 2,721 words · saved by 1 readers

Resources and Q&A for the book If Anyone Builds It, Everyone Dies.

Aren’t developers regularly making their AIs nice and safe and obedient? AIs steer in alien directions that only mostly coincide with helpfulness. Modern AIs are pretty helpful (or at least not harmful) to most users, most of the time. But as we noted above , a critical question is how to distinguish an AI that deeply wants to be helpful and do the right thing, from an AI with weirder and more complex drives that happen to line up with helpfulness under typical conditions, but which would prefer other conditions and outcomes even more. * Both sorts of AIs would act helpful in the typical case.

Explore this link on the map →

saved by

related reading