Monocular Dynamic View Synthesis: A Reality Check
Hang Gao, Ruilong Li, Shubham Tulsiani, Bryan Russell, Angjoo Kanazawa
摘要
We study the recent progress on dynamic view synthesis (DVS) from monocular video. Though existing approaches have demonstrated impressive results, we show a discrepancy between the practical capture process and the existing experimental protocols, which effectively leaks in multi-view signals during training. We define effective multi-view factors (EMFs) to quantify the amount of multi-view signal present in the input capture sequence based on the relative camera-scene motion. We introduce two new metrics: co-visibility masked image metrics and correspondence accuracy, which overcome the issue in existing protocols. We also propose a new iPhone dataset that includes more diverse real-life deformation sequences. Using our proposed experimental protocol, we show that the state-of-the-art approaches observe a 1-2 dB drop in masked PSNR in the absence of multi-view cues and 4-5 dB drop when modeling complex motion. Code and data can be found at https://hangg7.com/dycheck .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper114
- NeRFPlayer: A Streamable Dynamic Scene Representation with Decomposed Neural Radiance FieldsLiangchen Song, Anpei Chen, Zhong Li, Zhang Chen 等IEEE VR 2023 · 被引用 246 次
- Tracking Everything Everywhere All at OnceQianqian Wang, Yen-Yu Chang, Ruojin Cai, Zhengqi Li 等ICCV 2023 · 被引用 238 次
- Consistent4D: Consistent 360° Dynamic Object Generation from Monocular VideoYanqin Jiang, Li Zhang, Jin Gao, Weiming Hu 等ICLR 2024 · 被引用 120 次
- NerfAcc: Efficient Sampling Accelerates NeRFsRuilong Li, Hang Gao, Matthew Tancik, Angjoo KanazawaICCV 2023 · 被引用 119 次
- Nerfbusters: Removing Ghostly Artifacts from Casually Captured NeRFsFrederik Warburg, Ethan Weber, Matthew Tancik, Aleksander Holynski 等ICCV 2023 · 被引用 92 次
它引用的顶会 Paper24
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 被引用 2,647 次
- Mip-NeRF 360: Unbounded Anti-Aliased Neural Radiance FieldsJonathan T. Barron, Ben Mildenhall, Dor Verbin, Pratul P. Srinivasan 等CVPR 2022 · 被引用 1,603 次
- Non-Rigid Neural Radiance Fields: Reconstruction and Novel View Synthesis of a Dynamic Scene From Monocular VideoEdgar Tretschk, Ayush Tewari, Vladislav Golyanik, Michael Zollhöfer 等ICCV 2021 · 被引用 617 次
- Dynamic View Synthesis from Dynamic Monocular VideoChen Gao, Ayush Saraf, Johannes Kopf, Jia-Bin HuangICCV 2021 · 被引用 522 次
- HumanNeRF: Free-viewpoint Rendering of Moving People from Monocular VideoChung-Yi Weng, Brian Curless, Pratul P. Srinivasan, Jonathan T. Barron 等CVPR 2022 · 被引用 411 次
相关 Paper
- Deep 3D Mask Volume for View Synthesis of Dynamic ScenesKai-En Lin, Lei Xiao, Feng Liu, Guowei Yang 等ICCV 2021 · 被引用 42 次
- Scaling4D: Pushing the Frontier of Video Novel View Synthesis through Large-Scale Monocular VideosHongrui Cai, Junjie Luo, Zhihong Fu, Shengnan Zhu 等CVPR 2026
- Enhanced Stable View SynthesisNishant Jain, Suryansh Kumar, Luc Van GoolCVPR 2023
- Novel View Synthesis with View-Dependent Effects from a Single ImageJuan Luis Gonzalez Bello, Munchurl KimCVPR 2024
- WildRayZer: Self-supervised Large View Synthesis in Dynamic EnvironmentsXuweiyi Chen, Wentao Zhou, Zezhou ChengCVPR 2026 · 被引用 5 次
