SCTN: Sparse Convolution-Transformer Network for Scene Flow Estimation
Bing Li, Cheng Zheng, Silvio Giancola, Bernard Ghanem
摘要
We propose a novel scene flow estimation approach to capture and infer 3D motions from point clouds. Estimating 3D motions for point clouds is challenging, since a point cloud is unordered and its density is significantly non-uniform. Such unstructured data poses difficulties in matching corresponding points between point clouds, leading to inaccurate flow estimation. We propose a novel architecture named Sparse Convolution-Transformer Network (SCTN) that equips the sparse convolution with the transformer. Specifically, by leveraging the sparse convolution, SCTN transfers irregular point cloud into locally consistent flow features for estimating continuous and consistent motions within an object/local object part. We further propose to explicitly learn point relations using a point transformer module, different from exiting methods. We show that the learned relation-based contextual information is rich and helpful for matching corresponding points, benefiting scene flow estimation. In addition, a novel loss function is proposed to adaptively encourage flow consistency according to feature similarity. Extensive experiments demonstrate that our proposed approach achieves a new state of the art in scene flow estimation. Our approach achieves an error of 0.038 and 0.037 (EPE3D) on FlyingThings3D and KITTI Scene Flow respectively, which significantly outperforms previous methods by large margins.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- Focal Modulation NetworksJianwei Yang, Chunyuan Li, Xiyang Dai, Jianfeng GaoNeurIPS 2022 · 被引用 494 次
- Focal Attention for Long-Range Interactions in Vision TransformersJianwei Yang, Chunyuan Li, Pengchuan Zhang, Xiyang Dai 等NeurIPS 2021 · 被引用 228 次
- RPEFlow: Multimodal Fusion of RGB-PointCloud-Event for Joint Optical Flow and Scene Flow EstimationZhexiong Wan, Yuxin Mao, Jing Zhang, Yuchao DaiICCV 2023 · 被引用 35 次
- GMSF: Global Matching Scene FlowYushan Zhang, Johan Edstedt, Bastian Wandt, Per-Erik Forssén 等NeurIPS 2023 · 被引用 27 次
- EgoLoc: Revisiting 3D Object Localization from Egocentric Videos with Visual QueriesJinjie Mai, Abdullah Hamdi, Silvio Giancola, Chen Zhao 等ICCV 2023 · 被引用 26 次
它引用的顶会 Paper16
- Deformable DETR: Deformable Transformers for End-to-End Object DetectionXizhou Zhu, Weijie Su, Lewei Lu, Bin Li 等ICLR 2021 · 被引用 7,353 次
- KPConv: Flexible and Deformable Convolution for Point CloudsHugues Thomas, Charles R. Qi, Jean-Emmanuel Deschaud, Beatriz Marcotegui 等ICCV 2019 · 被引用 3,193 次
- ShellNet: Efficient Point Cloud Convolutional Neural Networks Using Concentric Shells StatisticsZhiyuan Zhang, Binh-Son Hua, Sai-Kit YeungICCV 2019 · 被引用 400 次
- SENSE: A Shared Encoder Network for Scene-Flow EstimationHuaizu Jiang, Deqing Sun, Varun Jampani, Zhaoyang Lv 等ICCV 2019 · 被引用 86 次
- Point TransformerHengshuang Zhao, Li Jiang, Jiaya Jia, Philip H. S. Torr 等ICCV 2021 · 被引用 23 次
相关 Paper
- HCRF-Flow: Scene Flow From Point Clouds With Continuous High-Order CRFs and Position-Aware Flow EmbeddingRuibo Li, Guosheng Lin, Tong He, Fayao Liu 等CVPR 2021
- PointConvFormer: Revenge of the Point-based ConvolutionWenxuan Wu, Fuxin Li, Qi ShanCVPR 2023
- RPPformer-Flow: Relative Position Guided Point Transformer for Scene Flow EstimationHanlin Li, Guanting Dong, Yueyi Zhang, Xiaoyan Sun 等ACM MM 2022 · 被引用 7 次
- Self-Supervised Robust Scene Flow Estimation via the Alignment of Probability Density FunctionsPan He, Patrick Emami, Sanjay Ranka, Anand RangarajanAAAI 2022 · 被引用 13 次
- PV-RAFT: Point-Voxel Correlation Fields for Scene Flow Estimation of Point CloudsYi Wei, Ziyi Wang, Yongming Rao, Jiwen Lu 等CVPR 2021
