Wasserstein Gradient Flows for Optimizing Gaussian Mixture Policies
Hanna Ziesche, Leonel Rozo
摘要
Robots often rely on a repertoire of previously-learned motion policies for performing tasks of diverse complexities. When facing unseen task conditions or when new task requirements arise, robots must adapt their motion policies accordingly. In this context, policy optimization is the de facto paradigm to adapt robot policies as a function of task-specific objectives. Most commonly-used motion policies carry particular structures that are often overlooked in policy optimization algorithms. We instead propose to leverage the structure of probabilistic policies by casting the policy optimization as an optimal transport problem. Specifically, we focus on robot motion policies that build on Gaussian mixture models (GMMs) and formulate the policy optimization as a Wassertein gradient flow over the GMMs space. This naturally allows us to constrain the policy updates via the L 2 -Wasserstein distance between GMMs to enhance the stability of the policy optimization process. Furthermore, we leverage the geometry of the Bures-Wasserstein manifold to optimize the Gaussian distributions of the GMM policy via Riemannian optimization. We evaluate our approach on common robotic settings: Reaching motions, collision-avoidance behaviors, and multi-goal tasks. Our results show that our method outperforms common policy optimization baselines in terms of task success rate and low-variance solutions. Preprint. Under review.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- SAC Flow: Sample-Efficient Reinforcement Learning of Flow-Based Policies via Velocity-Reparameterized Sequential ModelingYixian Zhang, Shu'ang Yu, Tonghe Zhang, Mo Guang 等ICLR 2026 · 被引用 33 次
- Slicing Wasserstein over Wasserstein via Functional Optimal TransportMoritz Piening, Robert BeinertICLR 2026 · 被引用 5 次
- Learning Anisotropic Value Geometry with Finsler Reinforcement LearningJumman Hossain, Nirmalya RoyICML 2026
它引用的顶会 Paper7
- Maximum Entropy RL (Provably) Solves Some Robust RL ProblemsBenjamin Eysenbach, Sergey LevineICLR 2022 · 被引用 244 次
- Large-Scale Wasserstein Gradient FlowsPetr Mokrov, Alexander Korotin, Lingxiao Li, Aude Genevay 等NeurIPS 2021 · 被引用 112 次
- Neural Dynamic Policies for End-to-End Sensorimotor LearningShikhar Bahl, Mustafa Mukadam, Abhinav Gupta, Deepak PathakNeurIPS 2020 · 被引用 97 次
- On Riemannian Optimization over Positive Definite Matrices with the Bures-Wasserstein GeometryAndi Han, Bamdev Mishra, Pratik Kumar Jawanpuria, Junbin GaoNeurIPS 2021 · 被引用 55 次
- Learning to Score Behaviors for Guided Policy OptimizationAldo Pacchiano, Jack Parker-Holder, Yunhao Tang, Krzysztof Choromanski 等ICML 2020 · 被引用 42 次
相关 Paper
- Flowing Datasets with Wasserstein over Wasserstein Gradient FlowsClément Bonet, Christophe Vauthier, Anna KorbaICML 2025
- Trust Region Policy Optimization with Optimal Transport Discrepancies: Duality and Algorithm for Continuous ActionsAntonio Terpin, Nicolas Lanzetti, Batuhan Yardim, Florian Dörfler 等NeurIPS 2022 · 被引用 14 次
- Continuous-time Riemannian SGD and SVRG Flows on Wasserstein Probabilistic SpaceMingyang Yi, Bohan WangNeurIPS 2025
- Accelerating Motion Planning via Optimal TransportAn T. Le, Georgia Chalvatzaki, Armin Biess, Jan PetersNeurIPS 2023 · 被引用 30 次
- Riemannian Convex Potential MapsSamuel Cohen, Brandon Amos, Yaron LipmanICML 2021 · 被引用 24 次
