flâneur — a map of the web's best reading

Imitative Generalisation (AKA 'Learning the Prior') — AI Alignment Forum

alignmentforum.org · 6,372 words · saved by 1 readers

Tl;dr We want to be able to supervise models with superhuman knowledge of the world and how to manipulate it. For this we need an overseer to be able…

x Imitative Generalisation (AKA 'Learning the Prior') — AI Alignment Forum Debate (AI safety technique) Distillation & Pedagogy Iterated Amplification OpenAI Outer Alignment AI Frontpage 60 Imitative Generalisation (AKA 'Learning the Prior') by Beth Barnes 10th Jan 2021 13 min read 15 60 Tl;dr We want to be able to supervise models with superhuman knowledge of the world and how to manipulate it. For this we need an overseer to be able to learn or access all the knowledge our models have, in order to be able to understand the consequences of suggestions or decisions from the model. If the overs

Explore this link on the map →

related reading