Stereo Vision Conversion from Planar Videos Based on Temporal Multiplane Images
Shanding Diao, Yuan Chen, Yang Zhao, Wei Jia, Zhao Zhang, Ronggang Wang
Abstract
With the rapid development of 3D movie and light-field displays, there is a growing demand for stereo videos. However, generating high-quality stereo videos from planar videos remains a challenging task. Traditional depth-image-based rendering techniques struggle to effectively handle the problem of occlusion exposure, which occurs when the occluded contents become visible in other views. Recently, the single-view multiplane images (MPI) representation has shown promising performance for planar video stereoscopy. However, the MPI still lacks real details that are occluded in the current frame, resulting in blurry artifacts in occlusion exposure regions. In fact, planar videos can leverage complementary information from adjacent frames to predict a more complete scene representation for the current frame. Therefore, this paper extends the MPI from still frames to the temporal domain, introducing the temporal MPI (TMPI). By extracting complementary information from adjacent frames based on optical flow guidance, obscured regions in the current frame can be effectively repaired. Additionally, a new module called masked optical flow warping (MOFW) is introduced to improve the propagation of pixels along optical flow trajectories. Experimental results demonstrate that the proposed method can generate high-quality stereoscopic or light-field videos from a single view and reproduce better occluded details than other state-of-the-art (SOTA) methods. https://github.com/Dio3ding/TMPI
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 40041053-8233-4cce-8043-6735f35ad7e4Cited by top-tier papers1
Ask how each one uses itBuilds on15
- Vision Transformers for Dense PredictionRené Ranftl, Alexey Bochkovskiy, Vladlen KoltunICCV 2021 · 2,647 citations
- Digging Into Self-Supervised Monocular Depth EstimationClément Godard, Oisin Mac Aodha, Michael Firman, Gabriel J. BrostowICCV 2019 · 2,416 citations
- MUSIQ: Multi-scale Image Quality TransformerJunjie Ke, Qifei Wang, Yilin Wang, Peyman Milanfar et al.ICCV 2021 · 1,325 citations
- MINE: Towards Continuous Depth MPI with NeRF for Novel View SynthesisJiaxin Li, Zijian Feng, Qi She, Henghui Ding et al.ICCV 2021 · 189 citations
- Towards An End-to-End Framework for Flow-Guided Video InpaintingZhen Li, Chengze Lu, Jianhua Qin, Chun-Le Guo et al.CVPR 2022 · 136 citations
Related papers
- MPI-Flow: Learning Realistic Optical Flow with Multiplane ImagesYingping Liang, Jiaming Liu, Debing Zhang, Ying FuICCV 2023 · 12 citations
- Tiled Multiplane Images for Practical 3D PhotographyNumair Khan, Lei Xiao, Douglas LanmanICCV 2023 · 15 citations
- Single-View View Synthesis in the Wild with Learned Adaptive Multiplane ImagesYuxuan Han, Ruicheng Wang, Jiaolong YangSIGGRAPH 2022 · 65 citations
- M2Flow: A Motion Information Fusion Framework for Enhanced Unsupervised Optical Flow Estimation in Autonomous DrivingXunpei Sun, Gang Chen, Zuoxun HouAAAI 2025 · 4 citations
- SVG: 3D Stereoscopic Video Generation via Denoising Frame MatrixPeng Dai, Feitong Tan, Qiangeng Xu, David Futschik et al.ICLR 2025
