Consistent depth of moving objects in video
Zhoutong Zhang, Forrester Cole, Richard Tucker, William T. Freeman, Tali Dekel
摘要
We present a method to estimate depth of a dynamic scene, containing arbitrary moving objects, from an ordinary video captured with a moving camera. We seek a geometrically and temporally consistent solution to this under-constrained problem: the depth predictions of corresponding points across frames should induce plausible, smooth motion in 3D. We formulate this objective in a new test-time training framework where a depth-prediction CNN is trained in tandem with an auxiliary scene-flow prediction MLP over the entire input video. By recursively unrolling the scene-flow prediction MLP over varying time steps, we compute both short-range scene flow to impose local smooth motion priors directly in 3D, and long-range scene flow to impose multi-view consistency constraints with wide baselines. We demonstrate accurate and temporally coherent results on a variety of challenging videos containing diverse moving objects (pets, people, cars), as well as camera motion. Our depth maps give rise to a number of depth-and-motion aware video editing effects such as object and lighting insertion.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper39
- SceneScape: Text-Driven Consistent Scene GenerationRafail Fridman, Amit Abecasis, Yoni Kasten, Tali DekelNeurIPS 2023 · 被引用 196 次
- Source-free Depth for Object Pop-outZongwei Wu, Danda Pani Paudel, Deng-Ping Fan, Jingjing Wang 等ICCV 2023 · 被引用 110 次
- Uncertainty-aware State Space Transformer for Egocentric 3D Hand Trajectory ForecastingWentao Bao, Lele Chen, Libing Zeng, Zhong Li 等ICCV 2023 · 被引用 34 次
- 4D Gaussian Splatting in the Wild with Uncertainty-Aware RegularizationMijeong Kim, Jongwoo Lim, Bohyung HanNeurIPS 2024 · 被引用 33 次
- Pseudo-Generalized Dynamic View Synthesis from a VideoXiaoming Zhao, Alex Colburn, Fangchang Ma, Miguel Ángel Bautista 等ICLR 2024 · 被引用 31 次
它引用的顶会 Paper8
- Nerfies: Deformable Neural Radiance FieldsKeunhong Park, Utkarsh Sinha, Jonathan T. Barron, Sofien Bouaziz 等ICCV 2021 · 被引用 1,442 次
- Consistent video depth estimationXuan Luo, Jia-Bin Huang, Richard Szeliski, Kevin Matzen 等SIGGRAPH 2020 · 被引用 321 次
- Occupancy Flow: 4D Reconstruction by Learning Particle DynamicsMichael Niemeyer, Lars M. Mescheder, Michael Oechsle, Andreas GeigerICCV 2019 · 被引用 314 次
- Self-Supervised Learning With Geometric Constraints in Monocular Video: Connecting Flow, Depth, and CameraYuhua Chen, Cordelia Schmid, Cristian SminchisescuICCV 2019 · 被引用 265 次
- Novel View Synthesis of Dynamic Scenes With Globally Coherent Depths From a Monocular CameraJae Shin Yoon, Kihwan Kim, Orazio Gallo, Hyun Soo Park 等CVPR 2020
相关 Paper
- Self-Supervised Monocular Scene Flow EstimationJunhwa Hur, Stefan RothCVPR 2020
- Neural Scene Flow Fields for Space-Time View Synthesis of Dynamic ScenesZhengqi Li, Simon Niklaus, Noah Snavely, Oliver WangCVPR 2021
- Temporally Consistent Online Depth Estimation Using Point-Based FusionNumair Khan, Eric Penner, Douglas Lanman, Lei XiaoCVPR 2023
- Dynamo-Depth: Fixing Unsupervised Depth Estimation for Dynamical ScenesYihong Sun, Bharath HariharanNeurIPS 2023 · 被引用 58 次
- Depth-Aware Test-Time Training for Zero-Shot Video Object SegmentationWeihuang Liu, Xi Shen, Haolun Li, Xiuli Bi 等CVPR 2024
