Building reliable sim driving agents by scaling self-play
This is experimental HTML to improve accessibility. We invite you to report rendering errors. Use Alt+Y to toggle on accessible reporting links and Alt+Shift+Y to toggle off. Learn more about this project and help improve conversions. HTML conversions sometimes display errors due to content that did not convert correctly from the source. This paper uses the following packages that are not yet supported by the HTML conversion tool. Feedback on these issues are not necessary; they are known and are being worked on. Authors: achieve the best HTML results from your LaTeX submissions by following these best practices. Simulation agents are essential for designing and testing systems that interact with humans, such as autonomous vehicles (AVs). These agents serve various purposes, from benchmarking AV performance to stress-testing system limits, but all applications share one key requirement: reliability. To enable systematic experimentation, a simulation agent must behave as intended. It s
Building reliable sim driving agents by scaling self-play Daphne Cornelisse Aarav Pandya Kevin Joseph Joseph Suárez Eugene Vinitsky Abstract Simulation agents are essential for designing and testing systems that interact with humans, such as autonomous vehicles (AVs). These agents serve various purposes, from benchmarking AV performance to stress-testing system limits, but all applications share one key requirement: reliability. To enable systematic experimentation, a simulation agent must behave as intended. It should minimize actions that may lead to undesired outcomes, such as collisions, w
Explore this link on the map →related reading
- Human-compatible driving partners through data-regularized self-play reinforcement learningarxiv.org
- [2304.03442] Generative Agents: Interactive Simulacra of Human Behaviorarxiv.org
- Towards self-driving codebases · Cursorcursor.com
- [2605.22748] Superhuman Safe and Agile Racing through Multi-Agent Reinforcement Learningarxiv.org
- [2312.15122] Scaling Is All You Need: Autonomous Driving with JAX-Accelerated Reinforcement Learningarxiv.org
- The Era of Experience Paper.pdfstorage.googleapis.com
- State of Robot Learning, December 2025vedder.io
- [2411.00114] Project Sid: Many-agent simulations toward AI civilizationarxiv.org
- [2304.03442] Generative Agents: Interactive Simulacra of Human Behaviorarxiv.org
- [AN #70]: Agents that help humans who are still learning about their own preferences — LessWronglesswrong.com
- pdfopenreview.net
- Learning Beyond Gradientstrinkle23897.github.io