SCTN: Sparse Convolution-Transformer Network for Scene Flow Estimation
Bing Li, Cheng Zheng, Silvio Giancola, Bernard Ghanem
Abstract
We propose a novel scene flow estimation approach to capture and infer 3D motions from point clouds. Estimating 3D motions for point clouds is challenging, since a point cloud is unordered and its density is significantly non-uniform. Such unstructured data poses difficulties in matching corresponding points between point clouds, leading to inaccurate flow estimation. We propose a novel architecture named Sparse Convolution-Transformer Network (SCTN) that equips the sparse convolution with the transformer. Specifically, by leveraging the sparse convolution, SCTN transfers irregular point cloud into locally consistent flow features for estimating continuous and consistent motions within an object/local object part. We further propose to explicitly learn point relations using a point transformer module, different from exiting methods. We show that the learned relation-based contextual information is rich and helpful for matching corresponding points, benefiting scene flow estimation. In addition, a novel loss function is proposed to adaptively encourage flow consistency according to feature similarity. Extensive experiments demonstrate that our proposed approach achieves a new state of the art in scene flow estimation. Our approach achieves an error of 0.038 and 0.037 (EPE3D) on FlyingThings3D and KITTI Scene Flow respectively, which significantly outperforms previous methods by large margins.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 00f6a5f5-9e73-409b-834b-02813d169817Cited by top-tier papers15
- Focal Modulation NetworksJianwei Yang, Chunyuan Li, Xiyang Dai, Jianfeng GaoNeurIPS 2022 · 494 citations
- Focal Attention for Long-Range Interactions in Vision TransformersJianwei Yang, Chunyuan Li, Pengchuan Zhang, Xiyang Dai et al.NeurIPS 2021 · 228 citations
- RPEFlow: Multimodal Fusion of RGB-PointCloud-Event for Joint Optical Flow and Scene Flow EstimationZhexiong Wan, Yuxin Mao, Jing Zhang, Yuchao DaiICCV 2023 · 35 citations
- GMSF: Global Matching Scene FlowYushan Zhang, Johan Edstedt, Bastian Wandt, Per-Erik Forssén et al.NeurIPS 2023 · 27 citations
- EgoLoc: Revisiting 3D Object Localization from Egocentric Videos with Visual QueriesJinjie Mai, Abdullah Hamdi, Silvio Giancola, Chen Zhao et al.ICCV 2023 · 26 citations
Builds on16
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li et al.ICLR 2021 · 7,353 citations
- KPConv: Flexible and Deformable Convolution for Point CloudsHugues Thomas, Charles R. Qi, Jean-Emmanuel Deschaud, Beatriz Marcotegui et al.ICCV 2019 · 3,193 citations
- ShellNet: Efficient Point Cloud Convolutional Neural Networks Using Concentric Shells StatisticsZhiyuan Zhang, Binh-Son Hua, Sai-Kit YeungICCV 2019 · 400 citations
- SENSE: A Shared Encoder Network for Scene-Flow EstimationHuaizu Jiang, Deqing Sun, Varun Jampani, Zhaoyang Lv et al.ICCV 2019 · 86 citations
- Point TransformerHengshuang Zhao, Li Jiang, Jiaya Jia, Philip H. S. Torr et al.ICCV 2021 · 23 citations
Related papers
- HCRF-Flow: Scene Flow From Point Clouds With Continuous High-Order CRFs and Position-Aware Flow EmbeddingRuibo Li, Guosheng Lin, Tong He, Fayao Liu et al.CVPR 2021
- PointConvFormer: Revenge of the Point-based ConvolutionWenxuan Wu, Fuxin Li, Qi ShanCVPR 2023
- RPPformer-Flow: Relative Position Guided Point Transformer for Scene Flow EstimationHanlin Li, Guanting Dong, Yueyi Zhang, Xiaoyan Sun et al.ACM MM 2022 · 7 citations
- Self-Supervised Robust Scene Flow Estimation via the Alignment of Probability Density FunctionsPan He, Patrick Emami, Sanjay Ranka, Anand RangarajanAAAI 2022 · 13 citations
- PV-RAFT: Point-Voxel Correlation Fields for Scene Flow Estimation of Point CloudsYi Wei, Ziyi Wang, Yongming Rao, Jiwen Lu et al.CVPR 2021
