flâneur — a map of the web's best reading

Iterated Distillation and Amplification | by Ajeya Cotra | AI Alignment

ai-alignment.com · 1,945 words · saved by 1 readers

Guest post summarizing my approach to aligned RL.

Machine Learning Iterated Distillation and Amplification Ajeya Cotra 8 min read · Mar 4, 2018 -- 6 Listen Share This is a guest post summarizing Paul Christiano’s proposed scheme for training machine learning systems that can be robustly aligned to complex and fuzzy values, which I call Iterated Distillation and Amplification (IDA) here. IDA is notably similar to AlphaGoZero and expert iteration . The hope is that if we use IDA to train each learned component of an AI then the overall AI will remain aligned with the user’s interests while achieving state of the art performance at runtime — pro

Explore this link on the map →

related reading