LipsNet++: Unifying Filter and Controller into a Policy Network
Xujie Song, Liangfa Chen, Tong Liu, Wenxuan Wang, Yinuo Wang, Shentao Qin, Yinsong Ma, Jingliang Duan, Shengbo Eben Li
摘要
Deep reinforcement learning (RL) is effective for decision-making and control tasks like autonomous driving and embodied AI. However, RL policies often suffer from the action fluctuation problem in real-world applications, resulting in severe actuator wear, safety risk, and performance degradation. This paper identifies the two fundamental causes of action fluctuation: observation noise and policy non-smoothness. We propose LipsNet++, a novel policy network with Fourier filter layer and Lipschitz controller layer to separately address both causes. The filter layer incorporates a trainable filter matrix that automatically extracts important frequencies while suppressing noise frequencies in the observations. The controller layer introduces a Jacobian regularization technique to achieve a low Lipschitz constant, ensuring smooth fitting of a policy function. These two layers function analogously to the filter and controller in classical control theory, suggesting that filtering and control capabilities can be seamlessly integrated into a single policy network. Both simulated and real-world experiments demonstrate that LipsNet++ achieves the state-ofthe-art noise robustness and action smoothness. The code and videos are publicly available at https://xjsong99.github.io/LipsNet v2 .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Empowering Multi-Robot Cooperation via Sequential World ModelsZijie Zhao, Honglei Guo, Shengqian Chen, Kaixuan Xu 等ICLR 2026 · 被引用 16 次
- Enhancing Control Policy Smoothness by Aligning Actions with Predictions from Preceding StatesKyoleen Kwak, Hyoseok HwangAAAI 2026 · 被引用 1 次
- Stabilizing the Q-Gradient Field for Policy Smoothness in Actor-Critic MethodsJeong Woon Lee, Kyoleen Kwak, Daeho Kim, Hyoseok HwangICML 2026 · 被引用 1 次
它引用的顶会 Paper11
- Global Filter Networks for Image ClassificationYongming Rao, Wenliang Zhao, Zheng Zhu, Jiwen Lu 等NeurIPS 2021 · 被引用 798 次
- Diffusion Actor-Critic with Entropy RegulatorYinuo Wang, Likun Wang, Yuxuan Jiang, Wenjun Zou 等NeurIPS 2024 · 被引用 105 次
- Deep Reinforcement Learning with Robust and Smooth PolicyQianli Shen, Yan Li, Haoming Jiang, Zhaoran Wang 等ICML 2020 · 被引用 95 次
- Gradient Normalization for Generative Adversarial NetworksYi-Lun Wu, Hong-Han Shuai, Zhi Rui Tam, Hong-Yu ChiuICCV 2021 · 被引用 78 次
- TAAC: Temporally Abstract Actor-Critic for Continuous ControlHaonan Yu, Wei Xu, Haichao ZhangNeurIPS 2021 · 被引用 30 次
相关 Paper
- LipsNet: A Smooth and Robust Neural Network with Adaptive Lipschitz Constant for High Accuracy Optimal ControlXujie Song, Jingliang Duan, Wenxuan Wang, Shengbo Eben Li 等ICML 2023 · 被引用 19 次
- ODE-based Smoothing Neural Network for Reinforcement Learning TasksYinuo Wang, Wenxuan Wang, Xujie Song, Tong Liu 等ICLR 2025
- Improve Robustness of Reinforcement Learning against Observation Perturbations via l∞ Lipschitz Policy NetworksBuqing Nie, Jingtian Ji, Yangqing Fu, Yue GaoAAAI 2024 · 被引用 10 次
- Addressing Action Oscillations through Learning Policy InertiaChen Chen, Hongyao Tang, Jianye Hao, Wulong Liu 等AAAI 2021 · 被引用 27 次
- Enhancing Robustness in Deep Reinforcement Learning: A Lyapunov Exponent ApproachRory Young, Nicolas PugeaultNeurIPS 2024 · 被引用 5 次
