Safely Learning Controlled Stochastic Dynamics
Luc Brogat-Motte, Alessandro Rudi, Riccardo Bonalli
Abstract
We address the problem of safely learning controlled stochastic dynamics from discrete-time trajectory observations, ensuring system trajectories remain within predefined safe regions during both training and deployment. Safety-critical constraints of this kind are crucial in applications such as autonomous robotics, finance, and biomedicine. We introduce a method that ensures safe exploration and efficient estimation of system dynamics by iteratively expanding an initial known safe control set using kernel-based confidence bounds. After training, the learned model enables predictions of the system's dynamics and permits safety verification of any given control. Our approach requires only mild smoothness assumptions and access to an initial safe control set, enabling broad applicability to complex real-world systems. We provide theoretical guarantees for safety and derive adaptive learning rates that improve with increasing Sobolev regularity of the true dynamics. Experimental evaluations demonstrate the practical effectiveness of our method in terms of safety, estimation accuracy, and computational efficiency.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext baad88e8-537d-4441-bf28-eacec134f66eBuilds on5
- Safe Reinforcement Learning in Constrained Markov Decision ProcessesAkifumi Wachi, Yanan SuiICML 2020 · 190 citations
- Safe Reinforcement Learning Using Advantage-Based InterventionNolan Wagener, Byron Boots, Ching-An ChengICML 2021 · 66 citations
- DiffPhyCon: A Generative Approach to Control Complex Physical SystemsLong Wei, Peiyan Hu, Ruiqi Feng, Haodong Feng et al.NeurIPS 2024 · 27 citations
- Information-Theoretic Safe Exploration with Gaussian ProcessesAlessandro G. Bottero, Carlos E. Luis, Julia Vinogradska, Felix Berkenkamp et al.NeurIPS 2022 · 18 citations
- Safe Time-Varying Optimization based on Gaussian Processes with Spatio-Temporal KernelJialin Li, Marta Zagórowska, Giulia De Pasquale, Alisa Rupenyan et al.NeurIPS 2024 · 11 citations
Related papers
- Learning Safe Control via On-the-Fly Bandit ExplorationAlexandre Capone, Ryan Kazuo Cosner, Aaron D. Ames, Sandra HircheICML 2025
- Safety Guarantees for Neural Network Dynamic Systems via Stochastic Barrier FunctionsRayan Mazouz, Karan Muvvala, Akash Ratheesh, Luca Laurenti et al.NeurIPS 2022 · 44 citations
- Almost Surely Stable Deep DynamicsNathan P. Lawrence, Philip D. Loewen, Michael G. Forbes, Johan U. Backström et al.NeurIPS 2020 · 28 citations
- ActSafe: Active Exploration with Safety Constraints for Reinforcement LearningYarden As, Bhavya Sukhija, Lenart Treven, Carmelo Sferrazza et al.ICLR 2025
- Gaussian Process Uniform Error Bounds with Unknown Hyperparameters for Safety-Critical ApplicationsAlexandre Capone, Armin Lederer, Sandra HircheICML 2022 · 26 citations
