Physics-Regulated Deep Reinforcement Learning: Invariant Embeddings
Hongpeng Cao, Yanbing Mao, Lui Sha, Marco Caccamo
摘要
This paper proposes the Phy-DRL: a physics-regulated deep reinforcement learning (DRL) framework for safety-critical autonomous systems. The Phy-DRL has three distinguished invariant-embedding designs: i) residual action policy (i.e., integrating data-driven-DRL action policy and physics-model-based action policy), ii) automatically constructed safety-embedded reward, and iii) physics-model-guided neural network (NN) editing, including link editing and activation editing. Theoretically, the Phy-DRL exhibits 1) a mathematically provable safety guarantee and 2) strict compliance of critic and actor networks with physics knowledge about the action-value function and action policy. Finally, we evaluate the Phy-DRL on a cart-pole system and a quadruped robot. The experiments validate our theoretical results and demonstrate that Phy-DRL features guaranteed safety compared to purely data-driven DRL and solely model-based design while offering remarkably fewer learning parameters and fast training towards safety guarantee.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper5
- Incorporating Symmetry into Deep Dynamics Models for Improved GeneralizationRui Wang, Robin Walters, Rose YuICLR 2021 · 被引用 201 次
- Learning Vision-Guided Quadrupedal Locomotion End-to-End with Cross-Modal TransformersRuihan Yang, Minghao Zhang, Nicklas Hansen, Huazhe Xu 等ICLR 2022 · 被引用 146 次
- Learning Compositional Koopman Operators for Model-Based ControlYunzhu Li, Hao He, Jiajun Wu, Dina Katabi 等ICLR 2020 · 被引用 135 次
- Towards Physics-informed Deep Learning for Turbulent Flow PredictionRui Wang, Karthik Kashinath, Mustafa Mustafa, Adrian Albert 等KDD 2020 · 被引用 39 次
- Isometric Transformation Invariant and Equivariant Graph Convolutional NetworksMasanobu Horie, Naoki Morita, Toshiaki Hishinuma, Yu Ihara 等ICLR 2021 · 被引用 25 次
相关 Paper
- Safe DNN-type Controller Synthesis for Nonlinear Systems via Meta Reinforcement LearningHanrui Zhao, Xia Zeng, Niuniu Qi, Zhengfeng Yang 等DAC 2023 · 被引用 4 次
- Reinforcement Learning with Adaptive Regularization for Safe Control of Critical SystemsHaozhe Tian, Homayoun Hamedmoghadam, Robert Shorten, Pietro FerraroNeurIPS 2024 · 被引用 11 次
- Trainify: A CEGAR-Driven Training and Verification Framework for Safe Deep Reinforcement LearningPeng Jin, Jiaxu Tian, Dapeng Zhi, Xuejun Wen 等CAV 2022 · 被引用 28 次
- Addressing Action Oscillations through Learning Policy InertiaChen Chen, Hongyao Tang, Jianye Hao, Wulong Liu 等AAAI 2021 · 被引用 27 次
- ConCerNet: A Contrastive Learning Based Framework for Automated Conservation Law Discovery and Trustworthy Dynamical System PredictionWang Zhang, Tsui-Wei Weng, Subhro Das, Alexandre Megretski 等ICML 2023 · 被引用 4 次
