Physics-informed Temporal Difference Metric Learning for Robot Motion Planning
Ruiqi Ni, Zherong Pan, Ahmed H. Qureshi
Abstract
The motion planning problem involves finding a collision-free path from a robot's starting to its target configuration. Recently, self-supervised learning methods have emerged to tackle motion planning problems without requiring expensive expert demonstrations. They solve the Eikonal equation for training neural networks and lead to efficient solutions. However, these methods struggle in complex environments because they fail to maintain key properties of the Eikonal equation, such as optimal value functions and geodesic distances. To overcome these limitations, we propose a novel self-supervised temporal difference metric learning approach that solves the Eikonal equation more accurately and enhances performance in solving complex and unseen planning tasks. Our method enforces Bellman's principle of optimality over finite regions, using temporal difference learning to avoid spurious local minima while incorporating metric learning to preserve the Eikonal equation's essential geodesic properties. We demonstrate that our approach significantly outperforms existing self-supervised learning methods in handling complex environments and generalizing to unseen environments, with robot configurations ranging from 2 to 12 degrees of freedom (DOF). The implementation code repository is available at https://github.com/ruiqini/ ntrl-demo .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3ece48e5-ba73-45b1-bf1a-5edcb16b307bCited by top-tier papers1
Ask how each one uses itBuilds on13
- PointNeXt: Revisiting PointNet++ with Improved Training and Scaling StrategiesGuocheng Qian, Yuchen Li, Houwen Peng, Jinjie Mai et al.NeurIPS 2022 · 1,270 citations
- DD-PPO: Learning Near-Perfect PointGoal Navigators from 2.5 Billion FramesErik Wijmans, Abhishek Kadian, Ari Morcos, Stefan Lee et al.ICLR 2020 · 608 citations
- Path Planning using Neural A* SearchRyo Yonetani, Tatsunori Taniai, Mohammadamin Barekatain, Mai Nishimura et al.ICML 2021 · 134 citations
- Optimal Goal-Reaching Reinforcement Learning via Quasimetric LearningTongzhou Wang, Antonio Torralba, Phillip Isola, Amy ZhangICML 2023 · 88 citations
- Learning Invariant Representations for Reinforcement Learning without ReconstructionAmy Zhang, Rowan Thomas McAllister, Roberto Calandra, Yarin Gal et al.ICLR 2021 · 77 citations
Related papers
- Model-Based Visual Planning with Self-Supervised Functional DistancesStephen Tian, Suraj Nair, Frederik Ebert, Sudeep Dasari et al.ICLR 2021 · 69 citations
- Self-Supervised Simultaneous Multi-Step Prediction of Road Dynamics and Cost MapElmira Amirloo Abolfathi, Mohsen Rohani, Ershad Banijamali, Jun Luo et al.CVPR 2021
- Self-Supervised Reinforcement Learning that Transfers using Random FeaturesBoyuan Chen, Chuning Zhu, Pulkit Agrawal, Kaiqing Zhang et al.NeurIPS 2023 · 16 citations
- Dynamical Distance Learning for Semi-Supervised and Unsupervised Skill DiscoveryKristian Hartikainen, Xinyang Geng, Tuomas Haarnoja, Sergey LevineICLR 2020 · 94 citations
- Goal Reaching with Eikonal-Constrained Hierarchical Quasimetric Reinforcement LearningVittorio Giammarino, Ahmed Hussain QureshiICLR 2026 · 7 citations
