Learn TAROT with MENTOR: A Meta-Learned Self-supervised Approach for Trajectory Prediction
Mozhgan Pourkeshavarz, Changhe Chen, Amir Rasouli
Abstract
Predicting diverse yet admissible trajectories that adhere to the map constraints is challenging. Graph-based scene encoders have been proven effective for preserving local structures of maps by defining lane-level connections. However, such encoders do not capture more complex patterns emerging from long-range heterogeneous connections between nonadjacent interacting lanes. To this end, we shed new light on learning common driving patterns by introducing meTA ROad paTh (TAROT) to formulate combinations of various relations between lanes on the road topology. Intuitively, this can be viewed as finding feasible routes. Furthermore, we propose MEta-road NeTwORk (MENTOR) that helps trajectory prediction by providing it with TAROT as navigation tips. More specifically, 1) we define TAROT prediction as a novel self-supervised proxy task to identify the complex heterogeneous structure of the map. 2) For typical driving actions, we establish several TAROTs that result in multiple Heterogeneous Structure Learning (HSL) tasks. These tasks are used in MENTOR, which performs meta-learning by simultaneously predicting trajectories along with proxy tasks, identifying an optimal combination of them, and automatically balancing them to improve the primary task. We show that our model achieves state-of-the-art performance on the Argoverse dataset, especially on diversity and admissibility metrics, achieving up to 20% improvements in challenging scenarios. We further investigate the contribution of proposed modules in ablation studies.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 64416da0-5614-4ac8-9ca5-1321c5007e85Cited by top-tier papers7
- CaDeT: A Causal Disentanglement Approach for Robust Trajectory Prediction in Autonomous DrivingMozhgan Pourkeshavarz, Junrui Zhang, Amir RasouliCVPR 2024 · 15 citations
- T4P: Test-Time Training of Trajectory Prediction via Masked Autoencoder and Actor-Specific Token MemoryDaehee Park, Jaeseok Jeong, Sung-Hoon Yoon, Jaewoo Jeong et al.CVPR 2024 · 14 citations
- Adversarial Backdoor Attack by Naturalistic Data Poisoning on Trajectory Prediction in Autonomous DrivingMozhgan Pourkeshavarz, Mohammad Sabokrou, Amir RasouliCVPR 2024 · 14 citations
- Den-TP: A Density-Balanced Data Curation and Evaluation Framework for Trajectory PredictionRuining Yang, Yi Xu, Yun Fu, Lili SuCVPR 2026
- Can Language Beat Numerical Regression? Language-Based Multimodal Trajectory PredictionInhwan Bae, Junoh Lee, Hae-Gon JeonCVPR 2024
Builds on10
- AgentFormer: Agent-Aware Transformers for Socio-Temporal Multi-Agent ForecastingYe Yuan, Xinshuo Weng, Yanglan Ou, Kris KitaniICCV 2021 · 658 citations
- DenseTNT: End-to-end Trajectory Prediction from Dense Goal SetsJunru Gu, Chen Sun, Hang ZhaoICCV 2021 · 563 citations
- HiVT: Hierarchical Vector Transformer for Multi-Agent Motion PredictionZikang Zhou, Luyao Ye, Jianping Wang, Kui Wu et al.CVPR 2022 · 379 citations
- Scene Transformer: A unified architecture for predicting future trajectories of multiple agentsJiquan Ngiam, Vijay Vasudevan, Benjamin Caine, Zhengdong Zhang et al.ICLR 2022 · 194 citations
- Diverse Trajectory Forecasting with Determinantal Point ProcessesYe Yuan, Kris M. KitaniICLR 2020 · 149 citations
Related papers
- LTP: Lane-based Trajectory Prediction for Autonomous DrivingJingke Wang, Tengju Ye, Ziqing Gu, Junbo ChenCVPR 2022 · 75 citations
- LaPred: Lane-Aware Prediction of Multi-Modal Future Trajectories of Dynamic AgentsByeoungdo Kim, SeongHyeon Park, Seokhwan Lee, Elbek Khoshimjonov et al.CVPR 2021
- Forecast-MAE: Self-supervised Pre-training for Motion Forecasting with Masked AutoencodersJie Cheng, Xiaodong Mei, Ming LiuICCV 2023 · 123 citations
- SEPT: Towards Efficient Scene Representation Learning for Motion PredictionZhiqian Lan, Yuxuan Jiang, Yao Mu, Chen Chen et al.ICLR 2024 · 56 citations
- VectorNet: Encoding HD Maps and Agent Dynamics From Vectorized RepresentationJiyang Gao, Chen Sun, Hang Zhao, Yi Shen et al.CVPR 2020
