Deep 3D Mask Volume for View Synthesis of Dynamic Scenes
Kai-En Lin, Lei Xiao, Feng Liu, Guowei Yang, Ravi Ramamoorthi
摘要
Image view synthesis has seen great success in reconstructing photorealistic visuals, thanks to deep learning and various novel representations. The next key step in immersive virtual experiences is view synthesis of dynamic scenes. However, several challenges exist due to the lack of high-quality training datasets, and the additional time dimension for videos of dynamic scenes. To address this issue, we introduce a multi-view video dataset, captured with a custom 10-camera rig in 120FPS. The dataset contains 96 high-quality scenes showing various visual effects and human interactions in outdoor scenes. We develop a new algorithm, Deep 3D Mask Volume, which enables temporally-stable view extrapolation from binocular videos of dynamic scenes, captured by static cameras. Our algorithm addresses the temporal inconsistency of disocclusions by identifying the error-prone areas with a 3D mask volume, and replaces them with static background observed throughout the video. Our method enables manipulation in 3D space as opposed to simple 2D masks, We demonstrate better temporal stability than frame-by-frame static view synthesis methods, or those that use 2D masks. The resulting view synthesis videos show minimal flickering artifacts and allow for larger translational movements.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper14
- DynPoint: Dynamic Neural Point For View SynthesisKaichen Zhou, Jia-Xing Zhong, Sangyun Shin, Kai Lu 等NeurIPS 2023 · 被引用 46 次
- 3D Moments from Near-Duplicate PhotosQianqian Wang, Zhengqi Li, David Salesin, Noah Snavely 等CVPR 2022 · 被引用 17 次
- Replay: Multi-modal Multi-view Acted Videos for Casual HolographyRoman Shapovalov, Yanir Kleiman, Ignacio Rocco, David Novotný 等ICCV 2023 · 被引用 11 次
- Factorized Motion Fields for Fast Sparse Input Dynamic View SynthesisNagabhushan Somraj, Kapil Choudhary, Sai Harsha Mupparaju, Rajiv SoundararajanSIGGRAPH 2024 · 被引用 6 次
- DiVa-360: The Dynamic Visual Dataset for Immersive Neural FieldsCheng-You Lu, Peisen Zhou, Angela Xing, Chandradeep Pokhariya 等CVPR 2024 · 被引用 4 次
它引用的顶会 Paper6
- Immersive light field video with a layered mesh representationMichael Broxton, John Flynn, Ryan S. Overbeck, Daniel Erickson 等SIGGRAPH 2020 · 被引用 271 次
- IBRNet: Learning Multi-View Image-Based RenderingQianqian Wang, Zhicheng Wang, Kyle Genova, Pratul P. Srinivasan 等CVPR 2021
- Local Implicit Grid Representations for 3D ScenesChiyu Max Jiang, Avneesh Sud, Ameesh Makadia, Jingwei Huang 等CVPR 2020
- Novel View Synthesis of Dynamic Scenes With Globally Coherent Depths From a Monocular CameraJae Shin Yoon, Kihwan Kim, Orazio Gallo, Hyun Soo Park 等CVPR 2020
- 4D Visualization of Dynamic Events From Unconstrained Multi-View VideosAayush Bansal, Minh Vo, Yaser Sheikh, Deva Ramanan 等CVPR 2020
相关 Paper
- DynamicStereo: Consistent Dynamic Depth from Stereo VideosNikita Karaev, Ignacio Rocco, Benjamin Graham, Natalia Neverova 等CVPR 2023
- Vivid4D: Improving 4D Reconstruction from Monocular Video by Video InpaintingJiaxin Huang, Sheng Miao, Bangbang Yang, Yuewen Ma 等ICCV 2025 · 被引用 2 次
- NVFi: Neural Velocity Fields for 3D Physics Learning from Dynamic VideosJinxi Li, Ziyang Song, Bo YangNeurIPS 2023 · 被引用 39 次
- DreamScene4D: Dynamic Multi-Object Scene Generation from Monocular VideosWen-Hsuan Chu, Lei Ke, Katerina FragkiadakiNeurIPS 2024 · 被引用 75 次
- MPI-Flow: Learning Realistic Optical Flow with Multiplane ImagesYingping Liang, Jiaming Liu, Debing Zhang, Ying FuICCV 2023 · 被引用 12 次
