1011.0686
arxiv.org · 6,833 words · saved by 1 readers
N/A
A Reduction of Imitation Learning and Structured Prediction to No-Regret Online Learning Stéphane Ross Geoffrey J. Gordon J. Andrew Bagnell Robotics Institute Machine Learning Department Robotics Institute Carnegie Mellon University Carnegie Mellon University…
saved by
related reading
- rltheorybook_ABJKS.pdfrltheorybook.github.io
- [2602.11399] Can We Really Learn One Representation to Optimize All Rewards?arxiv.org
- Learning to Imitate | SAIL Blogai.stanford.edu
- Ch. 21 - Imitation Learningunderactuated.mit.edu
- State of Robot Learning, December 2025vedder.io
- Reinforcement Learning via Implicit Imitation Guidancearxiv.org
- Generative Adversarial Imitation Learning.pdfcs.stanford.edu
- Generative Adversarial Imitation Learningarxiv.org
- Should I Use Offline RL or Imitation Learning? – The Berkeley Artificial Intelligence Research Blogbair.berkeley.edu
- The Pitfalls of Imitation Learning when Actions are Continuousarxiv.org
- [2406.04219] Multi-Agent Imitation Learning: Value is Easy, Regret is Hardarxiv.org
- [2606.09758] Difference-Aware Retrieval Policies for Imitation Learningarxiv.org