PRANK: motion Prediction based on RANKing
Yuriy Biktairov, Maxim Stebelev, Irina Rudenko, Oleh Shliazhko, Boris Yangel
Abstract
Predicting the motion of agents such as pedestrians or human-driven vehicles is one of the most critical problems in the autonomous driving domain. The overall safety of driving and the comfort of a passenger directly depend on its successful solution. The motion prediction problem also remains one of the most challenging problems in autonomous driving engineering, mainly due to high variance of the possible agent's future behavior given a situation. The two phenomena responsible for the said variance are the multimodality caused by the uncertainty of the agent's intent (e.g., turn right or move forward) and uncertainty in the realization of a given intent (e.g., which lane to turn into). To be useful within a real-time autonomous driving pipeline, a motion prediction system must provide efficient ways to describe and quantify this uncertainty, such as computing posterior modes and their probabilities or estimating density at the point corresponding to a given trajectory. It also should not put substantial density on physically impossible trajectories, as they can confuse the system processing the predictions. In this paper, we introduce the PRANK method, which satisfies these requirements. PRANK takes rasterized bird-eye images of agent's surroundings as an input and extracts features of the scene with a convolutional neural network. It then produces the conditional distribution of agent's trajectories plausible in the given scene. The key contribution of PRANK is a way to represent that distribution using nearest-neighbor methods in latent trajectory space, which allows for efficient inference in real time. We evaluate PRANK on the in-house and Argoverse datasets, where it shows competitive results.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5e23cd08-ee72-4433-adc6-2aa45adf8ab0Cited by top-tier papers6
- Motion Transformer with Global Intention Localization and Local Movement RefinementShaoshuai Shi, Li Jiang, Dengxin Dai, Bernt SchieleNeurIPS 2022 · 515 citations
- Latent Variable Sequential Set Transformers for Joint Multi-Agent Motion PredictionRoger Girgis, Florian Golemo, Felipe Codevilla, Martin Weiss et al.ICLR 2022 · 200 citations
- ADAPT: Efficient Multi-Agent Trajectory Prediction with AdaptationGörkay Aydemir, Adil Kaan Akan, Fatma GüneyICCV 2023 · 85 citations
- Vehicle trajectory prediction works, but not everywhereMohammadhossein Bahari, Saeed Saadatnejad, Ahmad Rahimi, Mohammad Shahverdikondori et al.CVPR 2022 · 65 citations
- Reasoning Multi-Agent Behavioral Topology for Interactive Autonomous DrivingHaochen Liu, Li Chen, Yu Qiao, Chen Lv et al.NeurIPS 2024 · 52 citations
Builds on3
- PRECOG: PREdiction Conditioned on Goals in Visual Multi-Agent SettingsNicholas Rhinehart, Rowan McAllister, Kris Kitani, Sergey LevineICCV 2019 · 407 citations
- VectorNet: Encoding HD Maps and Agent Dynamics From Vectorized RepresentationJiyang Gao, Chen Sun, Hang Zhao, Yi Shen et al.CVPR 2020
- CoverNet: Multimodal Behavior Prediction Using Trajectory SetsTung Phan-Minh, Elena Corina Grigore, Freddy A. Boulton, Oscar Beijbom et al.CVPR 2020
Related papers
- ProphNet: Efficient Agent-Centric Motion Forecasting with Anchor-Informed ProposalsXishun Wang, Tong Su, Fang Da, Xiaodong YangCVPR 2023
- Gaussian-Mixture Latent Flow for Stochastic 3D Human Motion PredictionYue Ma, Frederick W. B. Li, Xiaohui LiangCVPR 2026
- FJMP: Factorized Joint Multi-Agent Motion Prediction over Learned Directed Acyclic Interaction GraphsLuke Rowe, Martin Ethier, Eli-Henry Dykhne, Krzysztof CzarneckiCVPR 2023
- Multimodal Motion Prediction With Stacked TransformersYicheng Liu, Jinghuai Zhang, Liangji Fang, Qinhong Jiang et al.CVPR 2021
- Real-Time Motion Prediction via Heterogeneous Polyline Transformer with Relative Pose EncodingZhejun Zhang, Alexander Liniger, Christos Sakaridis, Fisher Yu et al.NeurIPS 2023 · 79 citations
