Physics-Regulated Deep Reinforcement Learning: Invariant Embeddings
Hongpeng Cao, Yanbing Mao, Lui Sha, Marco Caccamo
Abstract
This paper proposes the Phy-DRL: a physics-regulated deep reinforcement learning (DRL) framework for safety-critical autonomous systems. The Phy-DRL has three distinguished invariant-embedding designs: i) residual action policy (i.e., integrating data-driven-DRL action policy and physics-model-based action policy), ii) automatically constructed safety-embedded reward, and iii) physics-model-guided neural network (NN) editing, including link editing and activation editing. Theoretically, the Phy-DRL exhibits 1) a mathematically provable safety guarantee and 2) strict compliance of critic and actor networks with physics knowledge about the action-value function and action policy. Finally, we evaluate the Phy-DRL on a cart-pole system and a quadruped robot. The experiments validate our theoretical results and demonstrate that Phy-DRL features guaranteed safety compared to purely data-driven DRL and solely model-based design while offering remarkably fewer learning parameters and fast training towards safety guarantee.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext b3f0a38f-6682-4634-9aa0-384a7efda963Cited by top-tier papers1
Ask how each one uses itBuilds on5
- Incorporating Symmetry into Deep Dynamics Models for Improved GeneralizationRui Wang, Robin Walters, Rose YuICLR 2021 · 201 citations
- Learning Vision-Guided Quadrupedal Locomotion End-to-End with Cross-Modal TransformersRuihan Yang, Minghao Zhang, Nicklas Hansen, Huazhe Xu et al.ICLR 2022 · 146 citations
- Learning Compositional Koopman Operators for Model-Based ControlYunzhu Li, Hao He, Jiajun Wu, Dina Katabi et al.ICLR 2020 · 135 citations
- Towards Physics-informed Deep Learning for Turbulent Flow PredictionRui Wang, Karthik Kashinath, Mustafa Mustafa, Adrian Albert et al.KDD 2020 · 39 citations
- Isometric Transformation Invariant and Equivariant Graph Convolutional NetworksMasanobu Horie, Naoki Morita, Toshiaki Hishinuma, Yu Ihara et al.ICLR 2021 · 25 citations
Related papers
- Safe DNN-type Controller Synthesis for Nonlinear Systems via Meta Reinforcement LearningHanrui Zhao, Xia Zeng, Niuniu Qi, Zhengfeng Yang et al.DAC 2023 · 4 citations
- Reinforcement Learning with Adaptive Regularization for Safe Control of Critical SystemsHaozhe Tian, Homayoun Hamedmoghadam, Robert Shorten, Pietro FerraroNeurIPS 2024 · 11 citations
- Trainify: A CEGAR-Driven Training and Verification Framework for Safe Deep Reinforcement LearningPeng Jin, Jiaxu Tian, Dapeng Zhi, Xuejun Wen et al.CAV 2022 · 28 citations
- Addressing Action Oscillations through Learning Policy InertiaChen Chen, Hongyao Tang, Jianye Hao, Wulong Liu et al.AAAI 2021 · 27 citations
- ConCerNet: A Contrastive Learning Based Framework for Automated Conservation Law Discovery and Trustworthy Dynamical System PredictionWang Zhang, Tsui-Wei Weng, Subhro Das, Alexandre Megretski et al.ICML 2023 · 4 citations
