Unified Multi-Agent Trajectory Modeling with Masked Trajectory Diffusion
Songru Yang, Zhenwei Shi, Zhengxia Zou
摘要
Understanding movements in multi-agent scenarios is a fundamental problem in intelligent systems. Previous research assumes complete and synchronized observations. However, real-world partial observation caused by occlusions leads to inevitable model failure, which demands a unified framework for coexisting trajectory prediction, imputation, and recovery. Unlike previous attempts that handled observed and unobserved behaviors in a coupled manner, we explore a decoupled denoising diffusion modeling paradigm with a unidirectional information valve to separate the interference from uncertain behaviors. Building on this, we proposed a Unified Masked Trajectory Diffusion model (UniMTD) for arbitrary levels of missing observations. We designed a unidirectional attention as a valve unit to control the direction of information flow between the observed and masked areas, gradually refining the missing observations toward a real-world distribution. We construct it into a unidirectional MoE structure to handle varying proportions of missing observations. A Cached Diffusion model is further designed to improve generation quality while reducing computation and time overhead. Our method has achieved a great leap across human motions and vehicle traffic. UniMTD efficiently achieves 74% improvement in minADE 20 and reaches SOTA with advantages of 91%, 66%, 69%, and 58% across 4 fidelity metrics on out-of-boundary, velocity, and trajectory length. (a) UniMTD UniTraj GC-VRNN SSSD INAM Naomi MAT Transformer LSTM (b)
… Mixed Encoder Masked Observed Trajectory with arbitrary observation loss Noised Latent Space Vanilla Decoder Low-quality results Previous Methods UniMTD(Ours) Accurate Latent Space Observed Behaviors Preservation Masked Behaviors Projection Efficient Cached Diffusion Model Unidirectional Encoder High-quality results Reuse Calculation (c) Real-world Distribution
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper22
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Video Diffusion ModelsJonathan Ho, Tim Salimans, Alexey A. Gritsenko, William Chan 等NeurIPS 2022 · 被引用 2,948 次
- Autoregressive Denoising Diffusion Models for Multivariate Probabilistic Time Series ForecastingKashif Rasul, Calvin Seward, Ingmar Schuster, Roland VollgrafICML 2021 · 被引用 500 次
- HiVT: Hierarchical Vector Transformer for Multi-Agent Motion PredictionZikang Zhou, Luyao Ye, Jianping Wang, Kui Wu 等CVPR 2022 · 被引用 379 次
相关 Paper
- A Universal Model for Human Mobility PredictionQingyue Long, Yuan Yuan, Yong LiKDD 2025 · 被引用 8 次
- Unified Uncertainty-Aware Diffusion for Multi-Agent Trajectory ModelingGuillem Capellera, Antonio Rubio, Luis Ferraz, Antonio AgudoCVPR 2025
- Sports-Traj: A Unified Trajectory Generation Model for Multi-Agent Movement in SportsYi Xu, Yun FuICLR 2025
- Uncovering the Missing Pattern: Unified Framework Towards Trajectory Imputation and PredictionYi Xu, Armin Bazarjani, Hyung-Gun Chi, Chiho Choi 等CVPR 2023
- UniMotion: A Unified Motion Framework for Simulation, Prediction and PlanningNan Song, Junzhe Jiang, Jingyu Li, Xiatian Zhu 等NeurIPS 2025 · 被引用 2 次
