ZeroFlow: Scalable Scene Flow via Distillation
Kyle Vedder, Neehar Peri, Nathaniel Chodosh, Ishan Khatri, Eric Eaton, Dinesh Jayaraman, Yang Liu, Deva Ramanan, James Hays
Abstract
Scene flow estimation is the task of describing the 3D motion field between temporally successive point clouds. State-of-the-art methods use strong priors and test-time optimization techniques, but require on the order of tens of seconds to process full-size point clouds, making them unusable as computer vision primitives for real-time applications such as open world object detection. Feedforward methods are considerably faster, running on the order of tens to hundreds of milliseconds for full-size point clouds, but require expensive human supervision. To address both limitations, we propose Scene Flow via Distillation, a simple, scalable distillation framework that uses a label-free optimization method to produce pseudo-labels to supervise a feedforward model. Our instantiation of this framework, ZeroFlow, achieves state-of-the-art performance on the Argoverse 2 Self-Supervised Scene Flow Challenge while using zero human labels by simply training on large-scale, diverse unlabeled data. At test-time, ZeroFlow is over 1000x faster than label-free state-of-the-art optimization-based methods on full-size point clouds (34 FPS vs 0.028 FPS) and over 1000x cheaper to train on unlabeled data compared to the cost of human annotation ($394 vs $750,000). To facilitate further research, we release our code, trained model weights, and high quality pseudo-labels for the Argoverse 2 and Waymo Open datasets at https://vedder.io/zeroflow.html
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3d0a12f8-e63b-4743-b9b4-fd41fe41df9fCited by top-tier papers4
- 4DSegStreamer: Streaming 4D Panoptic Segmentation via Dual ThreadsLing Liu, Jun Tian, Li YiICCV 2025
- Floxels: Fast Unsupervised Voxel Based Scene Flow EstimationDavid T. Hoffmann, Syed Haseeb Raza, Hanqiu Jiang, Denis Tananaev et al.CVPR 2025
- RaLiFlow: Scene Flow Estimation with 4D Radar and LiDAR Point CloudsJingyun Fu, Zhiyu Xiang, Na ZhaoAAAI 2026
- Neural Eulerian Scene Flow FieldsKyle Vedder, Neehar Peri, Ishan Khatri, Siyi Li et al.ICLR 2025
Builds on15
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao et al.ICCV 2023 · 13,211 citations
- PointOdyssey: A Large-Scale Synthetic Dataset for Long-Term Point TrackingYang Zheng, Adam W. Harley, Bokui Shen, Gordon Wetzstein et al.ICCV 2023 · 255 citations
- LIV: Language-Image Representations and Rewards for Robotic ControlYecheng Jason Ma, Vikash Kumar, Amy Zhang, Osbert Bastani et al.ICML 2023 · 212 citations
- Neural Scene Flow PriorXueqian Li, Jhony Kaesemodel Pontes, Simon LuceyNeurIPS 2021 · 136 citations
Related papers
- TeFlow: Enabling Multi-frame Supervision for Self-Supervised Feed-forward Scene Flow EstimationQingwen Zhang, Chenhan Jiang, Xiaomeng Zhu, Yunqi Miao et al.CVPR 2026 · 5 citations
- ICP-Flow: LiDAR Scene Flow Estimation with ICPYancong Lin, Holger CaesarCVPR 2024
- SCOOP: Self-Supervised Correspondence and Optimization-Based Scene FlowItai Lang, Dror Aiger, Forrester Cole, Shai Avidan et al.CVPR 2023
- DeltaFlow: An Efficient Multi-frame Scene Flow Estimation MethodQingwen Zhang, Xiaomeng Zhu, Yushan Zhang, Yixi Cai et al.NeurIPS 2025 · 8 citations
- RigidFlow: Self-Supervised Scene Flow Learning on Point Clouds by Local Rigidity PriorRuibo Li, Chi Zhang, Guosheng Lin, Zhe Wang et al.CVPR 2022 · 47 citations
