flâneur — a map of the web's best reading

Optimal play in human-judged Debate usually won't answer your question - AI Alignment Forum

alignmentforum.org · 5,808 words · saved by 1 readers

Note: I fully expect some readers to find the core of this post almost trivially obvious. If you’re such a reader, please read as “I think [obvious thing] is important”, rather than “I’ve discovered…

x Optimal play in human-judged Debate usually won't answer your question — AI Alignment Forum Debate (AI safety technique) AI Rationality Frontpage 20 Optimal play in human-judged Debate usually won't answer your question by Joe Collman 27th Jan 2021 15 min read 12 20 Epistemic status: highly confident (99%+) this is an issue for optimal play with human consequentialist judges. Thoughts on practical implications are more speculative, and involve much hand-waving (70% sure I’m not overlooking a trivial fix, and that this can’t be safely ignored). Note: I fully expect some readers to find the co

Explore this link on the map →

related reading