Enforcing robust control guarantees within neural network policies
Priya L. Donti, Melrose Roderick, Mahyar Fazlyab, J. Zico Kolter
摘要
When designing controllers for safety-critical systems, practitioners often face a challenging tradeoff between robustness and performance. While robust control methods provide rigorous guarantees on system stability under certain worst-case disturbances, they often yield simple controllers that perform poorly in the average (non-worst) case. In contrast, nonlinear control methods trained using deep learning have achieved state-of-the-art performance on many control tasks, but often lack robustness guarantees. In this paper, we propose a technique that combines the strengths of these two approaches: constructing a generic nonlinear control policy class, parameterized by neural networks, that nonetheless enforces the same provable robustness criteria as robust control. Specifically, our approach entails integrating custom convex-optimization-based projection layers into a neural network-based policy. We demonstrate the power of this approach on several domains, improving in average-case performance over existing robust control methods and in worst-case stability over (non-robust) deep RL methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Safe Pontryagin Differentiable ProgrammingWanxin Jin, Shaoshuai Mou, George J. PappasNeurIPS 2021 · 被引用 65 次
- CROP: Certifying Robust Policies for Reinforcement Learning through Functional SmoothingFan Wu, Linyi Li, Zijian Huang, Yevgeniy Vorobeychik 等ICLR 2022 · 被引用 64 次
- Neural Lyapunov Control for Discrete-Time SystemsJunlin Wu, Andrew Clark, Yiannis Kantaros, Yevgeniy VorobeychikNeurIPS 2023 · 被引用 61 次
- Learning Barrier Certificates: Towards Safe Reinforcement Learning with Zero Training-time ViolationsYuping Luo, Tengyu MaNeurIPS 2021 · 被引用 58 次
- Recurrent Neural Network Controllers Synthesis with Stability Guarantees for Partially Observed SystemsFangda Gu, He Yin, Laurent El Ghaoui, Murat Arcak 等AAAI 2022 · 被引用 33 次
它引用的顶会 Paper1
相关 Paper
- Enforcing Hard Linear Constraints in Deep Learning Models with Decision RulesGonzalo E. Constante, Hao Chen, Can LiNeurIPS 2025 · 被引用 13 次
- Safe Controller Synthesis for Nonlinear Systems via Reinforcement Learning and PAC ApproximationXia Zeng, Banglong Liu, Zhenbing Zeng, Zhiming Liu 等DAC 2024 · 被引用 1 次
- Safe DNN-type Controller Synthesis for Nonlinear Systems via Meta Reinforcement LearningHanrui Zhao, Xia Zeng, Niuniu Qi, Zhengfeng Yang 等DAC 2023 · 被引用 4 次
- The Power of Learned Locally Linear Models for Nonlinear Policy OptimizationDaniel Pfrommer, Max Simchowitz, Tyler Westenbroek, Nikolai Matni 等ICML 2023 · 被引用 4 次
- An Iterative Scheme of Safe Reinforcement Learning for Nonlinear Systems via Barrier Certificate GenerationZhengfeng Yang, Yidan Zhang, Wang Lin, Xia Zeng 等CAV 2021 · 被引用 15 次
