Robust Regression for Safe Exploration in Control
We study the problem of safe learning and exploration in sequential control problems. The goal is to safely collect data samples from operating in an environment, in order to learn to achieve a challenging control goal (e.g., an agile maneuver close to a boundary). A central challenge in this setting is how to quantify uncertainty in order to choose provably-safe actions that allow us to collect informative data and reduce uncertainty, thereby achieving both improved controller safety and optimality. To address this challenge, we present a deep robust regression model that is trained to directly predict the uncertainty bounds for safe exploration. We derive generalization bounds for learning, and connect them with safety and stability bounds in control. We demonstrate empirically that our robust regression approach can outperform conventional Gaussian process (GP) based safe exploration in settings where it is difficult to specify a good GP prior. A key challenge in data-driven design
Abstract We study the problem of safe learning and exploration in sequential control problems. The goal is to safely collect data samples from operating in an environment, in order to learn to achieve a challenging control goal (e.g., an agile maneuver close to a boundary). A central challenge in this setting is how to quantify uncertainty in order to choose provably-safe actions that allow us to collect informative data and reduce uncertainty, thereby achieving both improved controller safety and optimality. To address this challenge, we present a deep robust regression model that is…
saved by
related reading
- A Free Lunch from the Noise:Provable and Practical Exploration for Representation Learningarxiv.org
- Ch. 21 - Imitation Learningunderactuated.mit.edu
- State of Robot Learning, December 2025vedder.io
- [1811.01848] Plan Online, Learn Offline: Efficient Learning and Exploration via Model-Based Controlarxiv.org
- Exploration for the Efficient Deployment of Reinforcement Learning Agentsopenreview.net
- Reinforcement Learning and Control as Probabilistic Inference: Tutorial and Reviewarxiv.org
- 2208.05129arxiv.org
- 1011.0686arxiv.org
- ProbabilisticRobotics.pdfdocs.ufpr.br
- [2507.09061] Action Chunking and Exploratory Data Collection Yield Exponential Improvements in Behavior Cloning for Continuous Controlarxiv.org
- 2307.08875arxiv.org
- [1805.12114] Deep Reinforcement Learning in a Handful of Trials using Probabilistic Dynamics Modelsar5iv.labs.arxiv.org