Multi-View Stereo by Temporal Nonparametric Fusion
Yuxin Hou, Juho Kannala, Arno Solin
摘要
We propose a novel idea for depth estimation from multiview image-pose pairs, where the model has capability to leverage information from previous latent-space encodings of the scene. This model uses pairs of images and poses, which are passed through an encoder-decoder model for disparity estimation. The novelty lies in soft-constraining the bottleneck layer by a nonparametric Gaussian process prior. We propose a pose-kernel structure that encourages similar poses to have resembling latent spaces. The flexibility of the Gaussian process (GP) prior provides adapting memory for fusing information from previous views. We train the encoder-decoder and the GP hyperparameters jointly end-to-end. In addition to a batch method, we derive a lightweight estimation scheme that circumvents standard pitfalls in scaling Gaussian process inference, and demonstrate how our scheme can run in real-time on smart devices.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper29
- NerfingMVS: Guided Optimization of Neural Radiance Fields for Indoor Multi-view StereoYi Wei, Shaohui Liu, Yongming Rao, Wang Zhao 等ICCV 2021 · 被引用 286 次
- TransformerFusion: Monocular RGB Scene Reconstruction using TransformersAljaz Bozic, Pablo R. Palafox, Justus Thies, Angela Dai 等NeurIPS 2021 · 被引用 185 次
- VolumeFusion: Deep Depth Fusion for 3D Scene ReconstructionJaesung Choe, Sunghoon Im, François Rameau, Minjun Kang 等ICCV 2021 · 被引用 83 次
- Time Will Tell: New Outlooks and A Baseline for Temporal Multi-View 3D Object DetectionJinhyung Park, Chenfeng Xu, Shijia Yang, Kurt Keutzer 等ICLR 2023 · 被引用 71 次
- PlanarRecon: Realtime 3D Plane Detection and Reconstruction from Posed Monocular VideosYiming Xie, Matheus Gadelha, Fengting Yang, Xiaowei Zhou 等CVPR 2022 · 被引用 32 次
相关 Paper
- A Global Depth-Range-Free Multi-View Stereo Transformer Network with Pose EmbeddingYitong Dong, Yijin Li, Zhaoyang Huang, Weikang Bian 等NeurIPS 2024 · 被引用 7 次
- PFDepth: Heterogeneous Pinhole-Fisheye Joint Depth Estimation via Distortion-aware Gaussian-Splatted Volumetric FusionZhiwei Zhang, Ruikai Xu, Weijian Zhang, Zhizhong Zhang 等ACM MM 2025 · 被引用 2 次
- Gaussian Process Priors for View-Aware InferenceYuxin Hou, Ari Heljakka, Arno SolinAAAI 2021 · 被引用 1 次
- Mutual Adaptive Reasoning for Monocular 3D Multi-Person Pose EstimationJuze Zhang, Jingya Wang, Ye Shi, Fei Gao 等ACM MM 2022 · 被引用 15 次
- Augmenting Depth Estimation with Geospatial ContextScott Workman, Hunter BlantonICCV 2021 · 被引用 6 次
