Zero-Shot 4D Lidar Panoptic Segmentation
Yushan Zhang, Aljosa Osep, Laura Leal-Taixé, Tim Meinhardt
Abstract
Prior methods (left) for zero-shot Lidar panoptic segmentation process individual (3D) point clouds in isolation. In contrast, our data-driven approach (right) operates directly on sequences of point clouds, jointly performing object segmentation, tracking, and zero-shot recognition based on text prompts specified at test time. Our method localizes and tracks any object and provides a temporally coherent semantic interpretation of dynamic scenes. We can correctly segment canonical objects, such as car, and objects beyond the vocabularies of standard Lidar datasets, such as advertising stand. Best seen in color, zoomed.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a09e019e-42a4-4d0c-9f66-c23454818999Builds on40
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao et al.ICCV 2023 · 13,211 citations
- Fourier Features Let Networks Learn High Frequency Functions in Low Dimensional DomainsMatthew Tancik, Pratul P. Srinivasan, Ben Mildenhall, Sara Fridovich-Keil et al.NeurIPS 2020 · 4,036 citations
- KPConv: Flexible and Deformable Convolution for Point CloudsHugues Thomas, Charles R. Qi, Jean-Emmanuel Deschaud, Beatriz Marcotegui et al.ICCV 2019 · 3,193 citations
- SemanticKITTI: A Dataset for Semantic Scene Understanding of LiDAR SequencesJens Behley, Martin Garbade, Andres Milioto, Jan Quenzel et al.ICCV 2019 · 2,345 citations
Related papers
- Towards Learning to Complete Anything in LidarAyça Takmaz, Cristiano Saltori, Neehar Peri, Tim Meinhardt et al.ICML 2025
- 3D Spatial Recognition Without Spatially Labeled 3DZhongzheng Ren, Ishan Misra, Alexander G. Schwing, Rohit GirdharCVPR 2021
- OpenESS: Event-Based Semantic Scene Understanding with Open VocabulariesLingdong Kong, Youquan Liu, Lai Xing Ng, Benoit R. Cottereau et al.CVPR 2024
- LiDAR-Based Panoptic Segmentation via Dynamic Shifting NetworkFangzhou Hong, Hui Zhou, Xinge Zhu, Hongsheng Li et al.CVPR 2021
- 4D Panoptic LiDAR SegmentationMehmet Aygun, Aljosa Osep, Mark Weber, Maxim Maximov et al.CVPR 2021
