[2406.04219] Multi-Agent Imitation Learning: Value is Easy, Regret is Hard
arXivLabs is a framework that allows collaborators to develop and share new arXiv features directly on our website. Both individuals and organizations that work with arXivLabs have embraced and accepted our values of openness, community, excellence, and user data privacy. arXiv is committed to these values and only works with partners that adhere to them. Have an idea for a project that will add value for arXiv's community? Learn more about arXivLabs. arXiv Operational Status Get status notifications via email or slack
[2406.04219] Multi-Agent Imitation Learning: Value is Easy, Regret is Hard --> Computer Science > Machine Learning arXiv:2406.04219 (cs) [Submitted on 6 Jun 2024 ( v1 ), last revised 26 Jun 2024 (this version, v2)] Title: Multi-Agent Imitation Learning: Value is Easy, Regret is Hard Authors: Jingwu Tang , Gokul Swamy , Fei Fang , Zhiwei Steven Wu View a PDF of the paper titled Multi-Agent Imitation Learning: Value is Easy, Regret is Hard, by Jingwu Tang and 3 other authors View PDF HTML (experimental) Abstract: We study a multi-agent imitation learning (MAIL) problem where we take the perspect
Explore this link on the map →saved by
related reading
- Learning to Imitate | SAIL Blogai.stanford.edu
- A (Long) Peek into Reinforcement Learning | Lil'Loglilianweng.github.io
- Reinforcement Learning via Implicit Imitation Guidancearxiv.org
- The Pitfalls of Imitation Learning when Actions are Continuousarxiv.org
- [AN #70]: Agents that help humans who are still learning about their own preferences — LessWronglesswrong.com
- Learning Beyond Gradientstrinkle23897.github.io
- Reinforcement Learning in Newcomblike Problemsproceedings.neurips.cc
- pistar06.pdfpi.website
- How DeepMind's Generally Capable Agents Were Trained — LessWronglesswrong.com
- [2305.18784] Collaborative Multi-Agent Heterogeneous Multi-Armed Banditsarxiv.org
- Should I Use Offline RL or Imitation Learning? – The Berkeley Artificial Intelligence Research Blogbair.berkeley.edu
- Ch. 21 - Imitation Learningunderactuated.mit.edu