flâneur

user_interactions.pdf

self-distillation.github.io · 8,336 words · saved by 1 readers

N/A

Aligning Language Models from User Interactions Aligning Language Models from User Interactions Thomas Kleine Buening1 Jonas Hübotter1 Barna Pásztor1 Idan Shenfeld2 Giorgia Ramponi3 Andreas Krause1 1 ETH Zurich 2 MIT 3 University of Zurich Abstract Multi-turn user interactions are among the most abundant data produced by language models, yet we lack effective methods to learn from them. While typically discarded, these interactions often contain useful information: follow-up user messages may…

saved by

related reading