Camera Pose Matters: Improving Depth Prediction by Mitigating Pose Distribution Bias
Yunhan Zhao, Shu Kong, Charless C. Fowlkes
摘要
Monocular depth predictors are typically trained on large-scale training sets which are naturally biased w.r.t the distribution of camera poses. As a result, trained predictors fail to make reliable depth predictions for testing examples captured under uncommon camera poses. To address this issue, we propose two novel techniques that exploit the camera pose during training and prediction. First, we introduce a simple perspective-aware data augmentation that synthesizes new training examples with more diverse views by perturbing the existing ones in a geometrically consistent manner. Second, we propose a conditional model that exploits the per-image camera pose as prior knowledge by encoding it as a part of the input. We show that jointly applying the two methods improves depth prediction on images captured under uncommon and even never-before-seen camera poses. We show that our methods improve performance when applied to a range of different predictor architectures. Lastly, we show that explicitly encoding the camera pose distribution improves the generalization performance of a synthetically trained depth predictor when evaluated on real images.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper8
- Long- Tailed Recognition via Weight BalancingShaden Alshammari, Yu-Xiong Wang, Deva Ramanan, Shu KongCVPR 2022 · 被引用 133 次
- Unified Domain Generalization and Adaptation for Multi-View 3D Object DetectionGyusam Chang, Jiwon Lee, Donghyun Kim, Jinkyu Kim 等NeurIPS 2024 · 被引用 19 次
- Geometry-Guided Domain Generalization for Monocular 3D Object DetectionFan Yang, Hui Chen, Yuwei He, Sicheng Zhao 等AAAI 2024 · 被引用 12 次
- Instance Tracking in 3D Scenes from Egocentric VideosYunhan Zhao, Haoyu Ma, Shu Kong, Charless C. FowlkesCVPR 2024 · 被引用 5 次
- Egocentric Scene Understanding via Multimodal Spatial RectifierTien Do, Khiem Vuong, Hyun Soo ParkCVPR 2022 · 被引用 5 次
它引用的顶会 Paper7
- Enforcing Geometric Constraints of Virtual Normal for Depth PredictionWei Yin, Yifan Liu, Chunhua Shen, Youliang YanICCV 2019 · 被引用 487 次
- Pseudo-LiDAR++: Accurate Depth for 3D Object Detection in Autonomous DrivingYurong You, Yan Wang, Wei-Lun Chao, Divyansh Garg 等ICLR 2020 · 被引用 439 次
- How Do Neural Networks See Depth in Single Images?Tom van Dijk, Guido de CroonICCV 2019 · 被引用 210 次
- 3D Scene Reconstruction With Multi-Layer Depth and Epipolar TransformersDaeyun Shin, Zhile Ren, Erik B. Sudderth, Charless C. FowlkesICCV 2019 · 被引用 67 次
- UprightNet: Geometry-Aware Camera Orientation Estimation From Single ImagesWenqi Xian, Zhengqi Li, Noah Snavely, Matthew Fisher 等ICCV 2019 · 被引用 52 次
相关 Paper
- Exploring Geometric Consistency for Monocular 3D Object DetectionQing Lian, Botao Ye, Ruijia Xu, Weilong Yao 等CVPR 2022 · 被引用 34 次
- Cameras as Relative Positional EncodingRuilong Li, Brent Yi, Junchen Liu, Hang Gao 等NeurIPS 2025 · 被引用 113 次
- Vivid4D: Improving 4D Reconstruction from Monocular Video by Video InpaintingJiaxin Huang, Sheng Miao, Bangbang Yang, Yuewen Ma 等ICCV 2025 · 被引用 2 次
- UniDepth: Universal Monocular Metric Depth EstimationLuigi Piccinelli, Yung-Hsu Yang, Christos Sakaridis, Mattia Segù 等CVPR 2024 · 被引用 122 次
- Cascaded Deep Monocular 3D Human Pose Estimation With Evolutionary Training DataShichao Li, Lei Ke, Kevin Pratama, Yu-Wing Tai 等CVPR 2020
