Input-level Inductive Biases for 3D Reconstruction
Wang Yifan, Carl Doersch, Relja Arandjelovic, João Carreira, Andrew Zisserman
摘要
Much of the recent progress in 3D vision has been driven by the development of specialized architectures that incorporate geometrical inductive biases. In this paper we tackle 3D reconstruction using a domain agnostic architecture and study how to inject the same type of inductive biases directly as extra inputs to the model. This approach makes it possible to apply existing general models, such as Perceivers, on this rich domain, without the need for architectural changes, while simultaneously maintaining data efficiency of bespoke models. In particular we study how to encode cameras, projective ray incidence and epipolar geometry as model inputs, and demonstrate competitive multi-view depth estimation performance on multiple benchmarks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- Towards Zero-Shot Scale-Aware Monocular Depth EstimationVitor Guizilini, Igor Vasiljevic, Dian Chen, Rares Ambrus 等ICCV 2023 · 被引用 129 次
- Long-Term Photometric Consistent Novel View Synthesis with Diffusion ModelsJason J. Yu, Fereshteh Forghani, Konstantinos G. Derpanis, Marcus A. BrubakerICCV 2023 · 被引用 71 次
- Dens3R: A Foundation Model for 3D Geometry PredictionXianze Fang, Jingnan Gao, Zhe Wang, Zhuo Chen 等ICLR 2026 · 被引用 45 次
- Attention-based Neural Cellular AutomataMattie Tesfaldet, Derek Nowrouzezahrai, Chris PalNeurIPS 2022 · 被引用 33 次
- IINet: Implicit Intra-inter Information Fusion for Real-Time Stereo MatchingXimeng Li, Chen Zhang, Wanjuan Su, Wenbing TaoAAAI 2024 · 被引用 22 次
它引用的顶会 Paper18
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Perceiver: General Perception with Iterative AttentionAndrew Jaegle, Felix Gimeno, Andy Brock, Oriol Vinyals 等ICML 2021 · 被引用 1,399 次
- Hierarchical Neural Architecture Search for Deep Stereo MatchingXuelian Cheng, Yiran Zhong, Mehrtash Harandi, Yuchao Dai 等NeurIPS 2020 · 被引用 436 次
- CrossTransformers: spatially-aware few-shot transferCarl Doersch, Ankush Gupta, Andrew ZissermanNeurIPS 2020 · 被引用 420 次
- Revisiting Stereo Depth Estimation From a Sequence-to-Sequence Perspective with TransformersZhaoshuo Li, Xingtong Liu, Nathan Drenkow, Andy S. Ding 等ICCV 2021 · 被引用 380 次
相关 Paper
- Equivariant Ray Embeddings for Implicit Multi-View Depth EstimationYinshuang Xu, Dian Chen, Katherine Liu, Sergey Zakharov 等NeurIPS 2024 · 被引用 11 次
- MonoSDF: Exploring Monocular Geometric Cues for Neural Implicit Surface ReconstructionZehao Yu, Songyou Peng, Michael Niemeyer, Torsten Sattler 等NeurIPS 2022 · 被引用 670 次
- UniDepth: Universal Monocular Metric Depth EstimationLuigi Piccinelli, Yung-Hsu Yang, Christos Sakaridis, Mattia Segù 等CVPR 2024 · 被引用 122 次
- Learning Efficient Photometric Feature Transform for Multi-view StereoKaizhang Kang, Cihui Xie, Ruisheng Zhu, Xiaohe Ma 等ICCV 2021 · 被引用 3 次
- PE3R: Perception-Efficient 3D ReconstructionJie Hu, Shizun Wang, Xinchao WangCVPR 2026 · 被引用 9 次
