DiffPoseNet: Direct Differentiable Camera Pose Estimation
Chethan M. Parameshwara, Gokul Hari, Cornelia Fermüller, Nitin J. Sanket, Yiannis Aloimonos
摘要
Current deep neural network approaches for camera pose estimation rely on scene structure for 3D motion estimation, but this decreases the robustness and thereby makes cross-dataset generalization difficult. In contrast, classical approaches to structure from motion estimate 3D motion utilizing optical flow and then compute depth. Their accuracy, however, depends strongly on the quality of the optical flow. To avoid this issue, direct methods have been proposed, which separate 3D motion from depth estimation, but compute 3D motion using only image gradients in the form of normal flow. In this paper, we introduce a network NFlowNet, for normal flow estimation which is used to enforce robust and direct constraints. In particular, normal flow is used to estimate relative camera pose based on the cheirality (depth positivity) constraint. We achieve this by formulating the optimization problem as a differentiable cheirality layer, which allows for end-to-end learning of camera pose. We perform extensive qualitative and quantitative evaluation of the proposed DiffPoseNet's sensitivity to noise and its generalization across datasets. We compare our approach to existing state-of-the-art methods on KITTI, TartanAir, and TUM-RGBD datasets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Scal3R: Scalable Test-Time Training for Large-Scale 3D ReconstructionTao Xie, Peishan Yang, Yudong Jin, Yingfeng Cai 等CVPR 2026 · 被引用 26 次
- Learning Normal Flow Directly from EventsDehao Yuan, Levi Burner, Jiayi Wu, Minghui Liu 等ICCV 2025 · 被引用 2 次
- Robust Frame-to-Frame Camera Rotation Estimation in Crowded ScenesFabien Delattre, David Dirnfeld, Phat Nguyen, Stephen Scarano 等ICCV 2023 · 被引用 2 次
- Flow-Guided Online Stereo Rectification for Wide Baseline StereoAnush Kumar, Fahim Mannan, Omid Hosseini Jafari, Shile Li 等CVPR 2024
- AlignDiff: Learning Physically-Grounded Camera Alignment via DiffusionLiuyue Xie, Jiancong Guo, Ozan Cakmakci, Andre Araujo 等ICCV 2025
它引用的顶会 Paper3
- Attacking Optical FlowAnurag Ranjan, Joel Janai, Andreas Geiger, Michael J. BlackICCV 2019 · 被引用 93 次
- MaskFlownet: Asymmetric Feature Matching With Learnable Occlusion MaskShengyu Zhao, Yilun Sheng, Yue Dong, Eric I-Chao Chang 等CVPR 2020
- Towards Better Generalization: Joint Depth-Pose Learning Without PoseNetWang Zhao, Shaohui Liu, Yezhi Shu, Yong-Jin LiuCVPR 2020
相关 Paper
- Deep Two-View Structure-From-Motion RevisitedJianyuan Wang, Yiran Zhong, Yuchao Dai, Stan Birchfield 等CVPR 2021
- FlowCam: Training Generalizable 3D Radiance Fields without Camera Poses via Pixel-Aligned Scene FlowCameron Smith, Yilun Du, Ayush Tewari, Vincent SitzmannNeurIPS 2023 · 被引用 43 次
- Consensus Learning with Deep Sets for Essential Matrix EstimationDror Moran, Yuval Margalit, Guy Trostianetsky, Fadi Khatib 等NeurIPS 2024 · 被引用 4 次
- PanoPose: Self-supervised Relative Pose Estimation for Panoramic ImagesDiantao Tu, Hainan Cui, Xianwei Zheng, Shuhan ShenCVPR 2024 · 被引用 4 次
- Polarimetric Relative Pose EstimationZhaopeng Cui, Viktor Larsson, Marc PollefeysICCV 2019 · 被引用 24 次
