Ch. 21 - Imitation Learning
underactuated.mit.edu · 6,582 words · saved by 2 readers
© Russ Tedrake, 2024 Last modified 2024-12-6. How to cite these notes, use annotations, and give feedback.
Ch. 21 - Imitation Learning Note: These are working notes used for a course being taught at MIT . They will be updated throughout the Spring 2024 semester. Lecture videos are available on YouTube . Previous Chapter Table of contents Next Chapter Imitation Learning Imitation learning, also known as "learning from demonstrations" (LfD), is the problem of learning a policy from a collection of demonstrations. For state-based feedback, these demonstrations take the form of a set of state-action sequences, $\left[ \bx[\cdot], \bu[\cdot]\right]$. For the richer class of output feedback , this takes
saved by
related reading
- State of Robot Learning, December 2025vedder.io
- How Claude Performs on Robotics Tasks \ Anthropicanthropic.com
- [2606.09758] Difference-Aware Retrieval Policies for Imitation Learningarxiv.org
- Reinforcement Learning via Implicit Imitation Guidancearxiv.org
- 2312.10812arxiv.org
- Learning to Imitate | SAIL Blogai.stanford.edu
- e5b5c402bb7bd5e60bede6961d6fe39e-Paper-Conference.pdfproceedings.iclr.cc
- Causal Video Models Are Data-Efficient Robot Policy Learners | Rhoda AIrhoda.ai
- Supervised Policy Learning for Real Robotssupervised-robot-learning.github.io
- [2206.11795] Video PreTraining (VPT): Learning to Act by Watching Unlabeled Online Videosarxiv.org
- Behavioral cloning mysteryseohong.me
- Generative Adversarial Imitation Learningarxiv.org