MPI-Flow: Learning Realistic Optical Flow with Multiplane Images
Yingping Liang, Jiaming Liu, Debing Zhang, Ying Fu
Abstract
The accuracy of learning-based optical flow estimation models heavily relies on the realism of the training datasets. Current approaches for generating such datasets either employ synthetic data or generate images with limited realism. However, the domain gap of these data with real-world scenes constrains the generalization of the trained model to real-world applications. To address this issue, we investigate generating realistic optical flow datasets from real-world images. Firstly, to generate highly realistic new images, we construct a layered depth representation, known as multiplane images (MPI), from single-view images. This allows us to generate novel view images that are highly realistic. To generate optical flow maps that correspond accurately to the new image, we calculate the optical flows of each plane using the camera matrix and plane depths. We then project these layered optical flows into the output optical flow map with volume rendering. Secondly, to ensure the realism of motion, we present an independent object motion module that can separate the camera and dynamic object motion in MPI. This module addresses the deficiency in MPI-based single-view methods, where optical flow is generated only by camera motion and does not account for any object movement. We additionally devise a depth-aware inpainting module to merge new images with dynamic objects and address unnatural motion occlusions. We show the superior performance of our method through extensive experiments on real-world datasets. Moreover, our approach achieves state-of-the-art performance in both unsupervised and supervised training of learning-based models. The code will be made publicly available at: https://github.com/Sharpiless/MPI-Flow.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 002f603c-5a3b-4af0-be86-66f06c15cff3Cited by top-tier papers4
- StreamFlow: Streamlined Multi-Frame Optical Flow Estimation for Video SequencesShangkun Sun, Jiaming Liu, Huaxia Li, Guoqing Liu et al.NeurIPS 2024 · 19 citations
- Enhancing Unregistered Hyperspectral Image Super-Resolution via Unmixing-based Abundance Fusion LearningYingkai Zhang, Tao Zhang, Jing Nie, Ying FuCVPR 2026 · 6 citations
- Rethinking Unsupervised Cross-modal Flow Estimation: Learning from Decoupled Optimization and Consistency ConstraintRunmin Zhang, Jialiang Wang, Si-Yuan Cao, Zhu Yu et al.ICLR 2026 · 1 citation
- Distilling Monocular Foundation Model for Fine-grained Depth CompletionYingping Liang, Yutao Hu, Wenqi Shao, Ying FuCVPR 2025
Builds on15
- Geometry-Free View Synthesis: Transformers and no 3D PriorsRobin Rombach, Patrick Esser, Björn OmmerICCV 2021 · 115 citations
- PixelSynth: Generating a 3D-Consistent Experience from a Single ImageChris Rockwell, David F. Fouhey, Justin JohnsonICCV 2021 · 98 citations
- Learning Optical Flow with Adaptive Graph ReasoningAo Luo, Fan Yang, Kunming Luo, Xin Li et al.AAAI 2022 · 73 citations
- Single-View View Synthesis in the Wild with Learned Adaptive Multiplane ImagesYuxuan Han, Ruicheng Wang, Jiaolong YangSIGGRAPH 2022 · 65 citations
- Learning Optical Flow with Kernel Patch AttentionAo Luo, Fan Yang, Xin Li, Shuaicheng LiuCVPR 2022 · 63 citations
Related papers
- Single-View View Synthesis With Multiplane ImagesRichard Tucker, Noah SnavelyCVPR 2020
- Learning Optical Flow From Still ImagesFilippo Aleotti, Matteo Poggi, Stefano MattocciaCVPR 2021
- Tiled Multiplane Images for Practical 3D PhotographyNumair Khan, Lei Xiao, Douglas LanmanICCV 2023 · 15 citations
- Stereo Vision Conversion from Planar Videos Based on Temporal Multiplane ImagesShanding Diao, Yuan Chen, Yang Zhao, Wei Jia et al.AAAI 2024 · 1 citation
- Efficient View Synthesis and 3D-based Multi-Frame Denoising with Multiplane Feature RepresentationsThomas Tanay, Ales Leonardis, Matteo MaggioniCVPR 2023
