✳flâneur — a map of the web's best reading
Ch. 21 - Imitation Learning
underactuated.mit.edu · 6,582 words · saved by 1 readers
© Russ Tedrake, 2024 Last modified 2024-12-6. How to cite these notes, use annotations, and give feedback.
Ch. 21 - Imitation Learning Note: These are working notes used for a course being taught at MIT . They will be updated throughout the Spring 2024 semester. Lecture videos are available on YouTube . Previous Chapter Table of contents Next Chapter Imitation Learning Imitation learning, also known as "learning from demonstrations" (LfD), is the problem of learning a policy from a collection of demonstrations. For state-based feedback, these demonstrations take the form of a set of state-action sequences, $\left[ \bx[\cdot], \bu[\cdot]\right]$. For the richer class of output feedback , this takes
Explore this link on the map →saved by
related reading
- State of Robot Learning, December 2025vedder.io
- Reinforcement Learning via Implicit Imitation Guidancearxiv.org
- On-Policy Distillation - Thinking Machines Labthinkingmachines.ai
- Learning to Imitate | SAIL Blogai.stanford.edu
- e5b5c402bb7bd5e60bede6961d6fe39e-Paper-Conference.pdfproceedings.iclr.cc
- The Pitfalls of Imitation Learning when Actions are Continuousarxiv.org
- Causal Video Models Are Data-Efficient Robot Policy Learners | Rhoda AIrhoda.ai
- Imitation Bootstrapped Reinforcement Learningarxiv.org
- pistar06.pdfpi.website
- RLDG: Robotic Generalist Policy Distillation via Reinforcement Learningarxiv.org
- [2206.11795] Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videosarxiv.org
- Open X-Embodiment: Robotic Learning Datasets and RT-X Modelsarxiv.org