Temporally Consistent Online Depth Estimation Using Point-Based Fusion
Numair Khan, Eric Penner, Douglas Lanman, Lei Xiao
摘要
Depth estimation is an important step in many computer vision problems such as 3D reconstruction, novel view synthesis, and computational photography. Most existing work focuses on depth estimation from single frames. When applied to videos, the result lacks temporal consistency, showing flickering and swimming artifacts. In this paper we aim to estimate temporally consistent depth maps of video streams in an online setting. This is a difficult problem as future frames are not available and the method must choose between enforcing consistency and correcting errors from previous estimations. The presence of dynamic objects further complicates the problem. We propose to address these challenges by using a global point cloud that is dynamically updated each frame, along with a learned fusion approach in image space. Our approach encourages consistency while simultaneously allowing updates to handle errors and dynamic objects. Qualitative and quantitative results show that our method achieves state-of-the-art quality for consistent video depth estimation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- DAGE: Dual-Stream Architecture for Efficient and Fine-Grained Geometry EstimationTuan Duc Ngo, Jiahui Huang, Seoung Wug Oh, Kevin Blackburn-Matzen 等CVPR 2026 · 被引用 3 次
- FlashDepth: Real-Time Streaming Video Depth Estimation at 2K ResolutionGene Chou, Wenqi Xian, Guandao Yang, Mohamed Abdelfattah 等ICCV 2025 · 被引用 1 次
- Uni4D: Unifying Visual Foundation Models for 4D Modeling from a Single VideoDavid Yifan Yao, Albert J. Zhai, Shenlong WangCVPR 2025
它引用的顶会 Paper21
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 被引用 2,647 次
- Alias-Free Generative Adversarial NetworksTero Karras, Miika Aittala, Samuli Laine, Erik Härkönen 等NeurIPS 2021 · 被引用 2,126 次
- Hierarchical Neural Architecture Search for Deep Stereo MatchingXuelian Cheng, Yiran Zhong, Mehrtash Harandi, Yuchao Dai 等NeurIPS 2020 · 被引用 436 次
- Consistent video depth estimationXuan Luo, Jia-Bin Huang, Richard Szeliski, Kevin Matzen 等SIGGRAPH 2020 · 被引用 321 次
- DeepV2D: Video to Depth with Differentiable Structure from MotionZachary Teed, Jia DengICLR 2020 · 被引用 314 次
相关 Paper
- Robust Consistent Video Depth EstimationJohannes Kopf, Xuejian Rong, Jia-Bin HuangCVPR 2021
- Consistent depth of moving objects in videoZhoutong Zhang, Forrester Cole, Richard Tucker, William T. Freeman 等SIGGRAPH 2021 · 被引用 26 次
- DepthCrafter: Generating Consistent Long Depth Sequences for Open-world VideosWenbo Hu, Xiangjun Gao, Xiaoyu Li, Sijie Zhao 等CVPR 2025
- Geometry-guided Online 3D Video Synthesis with Multi-View Temporal ConsistencyHyunho Ha, Lei Xiao, Christian Richardt, Thu Nguyen-Phuoc 等CVPR 2025
- Learning to Fuse Monocular and Multi-view Cues for Multi-frame Depth Estimation in Dynamic ScenesRui Li, Dong Gong, Wei Yin, Hao Chen 等CVPR 2023
