Safe Local Motion Planning With Self-Supervised Freespace Forecasting
Peiyun Hu, Aaron Huang, John M. Dolan, David Held, Deva Ramanan
Abstract
Figure 1 : What are good 3D representations that support planning in dynamic environments? We visualize a typical urban motion planning scenario from a bird's-eye view, where an autonomous vehicle (AV) awaits an unprotected left turn. We highlight a candidate plan with a blue arrow, whose endpoint represents where the AV will be in 1s. An object-centric representation (left), as adopted by standard perception stacks, focuses on object properties (their shape, orientation, position, etc.) both at the current time step and the future. Alternatively, a freespace-centric representation directly captures the freespace of the surrounding scene and can be readily obtained by raycasting measurements from a depth (e.g., LiDAR) sensor. Forecasting a future version (in 1s) of either representation could help the AV identify a potential collision associated with the candidate plan, however at wildly different annotation costs. Forecasting future object trajectories requires a massive amount of object and track labels to train perceptual modules. Instead, we explore future freespace, whose forecasting can be naturally self-supervised by simply letting time move forward and raycasting future sensor measurements. We propose approaches to planning with forecasted freespace and learning to plan with future freespace.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 694c2776-f92f-4e4f-bbba-d511da3437a1Cited by top-tier papers29
- VAD: Vectorized Scene Representation for Efficient Autonomous DrivingBo Jiang, Shaoyu Chen, Qing Xu, Bencheng Liao et al.ICCV 2023 · 602 citations
- Behavior Generation with Latent ActionsSeungjae Lee, Yibin Wang, Haritheja Etukuru, H. Jin Kim et al.ICML 2024 · 154 citations
- OpenDriveVLA: Towards End-to-end Autonomous Driving with Large Vision Language Action ModelXingcheng Zhou, Xuyuan Han, Feng Yang, Yunpu Ma et al.AAAI 2026 · 119 citations
- DME-Driver: Integrating Human Decision Logic and 3D Scene Perception in Autonomous DrivingWencheng Han, Dongqian Guo, Cheng-Zhong Xu, Jianbing ShenAAAI 2025 · 64 citations
- VLP: Vision Language Planning for Autonomous DrivingChenbin Pan, Burhaneddin Yaman, Tommaso Nesti, Abhirup Mallik et al.CVPR 2024 · 53 citations
Builds on3
- Exploring the Limitations of Behavior Cloning for Autonomous DrivingFelipe Codevilla, Eder Santana, Antonio M. López, Adrien GaidonICCV 2019 · 666 citations
- Deep Imitative Models for Flexible Inference, Planning, and ControlNicholas Rhinehart, Rowan McAllister, Sergey LevineICLR 2020 · 159 citations
- Predicting Semantic Map Representations From Images Using Pyramid Occupancy NetworksThomas Roddick, Roberto CipollaCVPR 2020
Related papers
- What You See is What You Get: Exploiting Visibility for 3D Object DetectionPeiyun Hu, Jason Ziglar, David Held, Deva RamananCVPR 2020
- Point Cloud Forecasting as a Proxy for 4D Occupancy ForecastingTarasha Khurana, Peiyun Hu, David Held, Deva RamananCVPR 2023
- TREND: Unsupervised 3D Representation Learning via Temporal Forecasting for LiDAR PerceptionRunjian Chen, Hyoungseob Park, Bo Zhang, Wenqi Shao et al.NeurIPS 2025 · 4 citations
- Self-Supervised Simultaneous Multi-Step Prediction of Road Dynamics and Cost MapElmira Amirloo Abolfathi, Mohsen Rohani, Ershad Banijamali, Jun Luo et al.CVPR 2021
- DriveX: Omni Scene Modeling for Learning Generalizable World Knowledge in Autonomous DrivingChen Shi, Shaoshuai Shi, Kehua Sheng, Bo Zhang et al.ICCV 2025
