flâneur — a map of the web's best reading

How do LLMs generalize when we do training that is intuitively compatible with two off-distribution behaviors? — LessWrong

lesswrong.com · 7,148 words · saved by 1 readers

Authors: Dylan Xu, Alek Westover, Vivek Hebbar, Sebastian Prasanna, Nathan Sheffield, Buck Shlegeris, Julian Stastny …

x How do LLMs generalize when we do training that is intuitively compatible with two off-distribution behaviors? — LessWrong AI Frontpage 62 How do LLMs generalize when we do training that is intuitively compatible with two off-distribution behaviors? by Dylan Xu , Alek Westover , Vivek Hebbar , SebastianP , frisby , Julian Stastny 20th Apr 2026 24 min read 5 62 Authors: Dylan Xu, Alek Westover, Vivek Hebbar, Sebastian Prasanna, Nathan Sheffield, Buck Shlegeris, Julian Stastny Thanks to Eric Gan and Aghyad Deeb for feedback on a draft of this post. EDIT: After further consideration and @nostal

Explore this link on the map →

saved by

related reading