arXiv · 2504.00626
Probabilistically safe and efficient model-based reinforcement learning
Abstract
This paper proposes tackling safety-critical stochastic Reinforcement Learning (RL) tasks with a sample-based, model-based approach. At the core of the method lies a Model Predictive Control (MPC) scheme that acts as function approximation, providing a model-based predictive control policy. To ensure safety, a probabilistic Control Barrier Function (CBF) is integrated into the MPC controller. To approximate the effects of stochasticies in the optimal control formulation and to fulfil the probabilistic CBF condition, a sample-based approach with guarantees is employed. Furthermore, to counterbalance the additional computational burden due to sampling, a learnable terminal cost formulation is included in the MPC objective. An RL algorithm is deployed to learn both the terminal cost and the CBF constraint. Results from a numerical experiment on a constrained LTI problem corroborate the effectiveness of the proposed methodology in reducing computation time while preserving control performance and safety.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Filippo Airaldi, Bart De Schutter, Azita Dabiri. 2025-04-01. Probabilistically safe and efficient model-based reinforcement learning. https://arxiv.org/abs/2504.00626
Cite the original work for its findings. Save a collection to share your selection of sources.