flâneur — a map of the web's best reading

Supervising strong learners by amplifying weak experts — AI Alignment Forum

alignmentforum.org · 302 words · saved by 1 readers

Tomorrow's AI Alignment Forum sequences post will be 'AI safety without goal-directed behavior' by Rohin Shah, in the sequence on Value Learning. The next post in this sequence on Iterated Amplification will be 'AlphaGo Zero and capability amplification', by Paul Christiano, on Tuesday 8th January.

x Supervising strong learners by amplifying weak experts — AI Alignment Forum Iterated Amplification Iterated Amplification Frontpage 9 Supervising strong learners by amplifying weak experts by paulfchristiano 6th Jan 2019 1 min read 1 9 This is a linkpost for https://arxiv.org/pdf/1810.08575.pdf Abstract Many real world learning tasks involve complex or hard-to-specify objectives, and using an easier-to-specify proxy can lead to poor performance or misaligned behavior. One solution is to have humans provide a training signal by demonstrating or judging performance, but this approach fails if

Explore this link on the map →

related reading