Occ2Net: Robust Image Matching Based on 3D Occupancy Estimation for Occluded Regions
Miao Fan, Mingrui Chen, Chen Hu, Shuchang Zhou
摘要
Image matching is a fundamental and critical task in various visual applications, such as Simultaneous Localization and Mapping (SLAM) and image retrieval, which require accurate pose estimation. However, most existing methods ignore the occlusion relations between objects caused by camera motion and scene structure. In this paper, we propose Occ 2 Net, a novel image matching method that models occlusion relations using 3D occupancy and infers matching points in occluded regions. Thanks to the inductive bias encoded in the Occupancy Estimation (OE) module, it greatly simplifies bootstrapping of a multi-view consistent 3D representation that can then integrate information from multiple views. Together with an Occlusion-Aware (OA) module, it incorporates attention layers and rotation alignment to enable matching between occluded and visible points. We evaluate our method on both real-world and simulated datasets and demonstrate its superior performance over state-of-the-art methods on several metrics, especially in occlusion scenarios.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- ETO: Efficient Transformer-based Local Feature Matching by Organizing Multiple Homography HypothesesJunjie Ni, Guofeng Zhang, Guanglin Li, Yijin Li 等NeurIPS 2024 · 被引用 14 次
- Alligat0R: Pre-Training through Covisibility Segmentation for Relative Camera Pose RegressionThibaut Loiseau, Guillaume Bourmaud, Vincent LepetitNeurIPS 2025 · 被引用 11 次
- PRISM: PRogressive dependency maxImization for Scale-invariant image MatchingXudong Cai, Yongcai Wang, Lun Luo, Minhang Wang 等ACM MM 2024 · 被引用 3 次
- RUBIK: A Structured Benchmark for Image Matching across Geometric ChallengesThibaut Loiseau, Guillaume BourmaudCVPR 2025
它引用的顶会 Paper16
- DROID-SLAM: Deep Visual SLAM for Monocular, Stereo, and RGB-D CamerasZachary Teed, Jia DengNeurIPS 2021 · 被引用 1,248 次
- DISK: Learning local features with policy gradientMichal J. Tyszkiewicz, Pascal Fua, Eduard TrullsNeurIPS 2020 · 被引用 652 次
- Rethinking and Improving Relative Position Encoding for Vision TransformerKan Wu, Houwen Peng, Minghao Chen, Jianlong Fu 等ICCV 2021 · 被引用 427 次
- MAT: Mask-Aware Transformer for Large Hole Image InpaintingWenbo Li, Zhe Lin, Kun Zhou, Lu Qi 等CVPR 2022 · 被引用 382 次
- Learning Two-View Correspondences and Geometry Using Order-Aware NetworkJiahui Zhang, Dawei Sun, Zixin Luo, Anbang Yao 等ICCV 2019 · 被引用 362 次
相关 Paper
- Deep Two-View Structure-From-Motion RevisitedJianyuan Wang, Yiran Zhong, Yuchao Dai, Stan Birchfield 等CVPR 2021
- DVMNet: Computing Relative Pose for Unseen Objects Beyond HypothesesChen Zhao, Tong Zhang, Zheng Dang, Mathieu SalzmannCVPR 2024 · 被引用 5 次
- OAMaskFlow: Occlusion-Aware Motion Mask for Scene FlowXiongfeng Peng, Zhihua Liu, Weiming Li, Yamin Mao 等AAAI 2025
- OCR-Pose: Occlusion-aware Contrastive Representation for Unsupervised 3D Human Pose EstimationJunjie Wang, Zhenbo Yu, Zhengyan Tong, Hang Wang 等ACM MM 2022 · 被引用 12 次
- DCNet: Dense Correspondence Neural Network for 6DoF Object Pose Estimation in Occluded ScenesZhi Chen, Wei Yang, Zhenbo Xu, Xike Xie 等ACM MM 2020 · 被引用 3 次
