Flow2Stereo: Effective Self-Supervised Learning of Optical Flow and Stereo Matching
Pengpeng Liu, Irwin King, Michael R. Lyu, Jia Xu
Abstract
In this paper, we propose a unified method to jointly learn optical flow and stereo matching. Our first intuition is stereo matching can be modeled as a special case of optical flow, and we can leverage 3D geometry behind stereoscopic videos to guide the learning of these two forms of correspondences. We then enroll this knowledge into the state-ofthe-art self-supervised learning framework, and train one single network to estimate both flow and stereo. Second, we unveil the bottlenecks in prior self-supervised learning approaches, and propose to create a new set of challenging proxy tasks to boost performance. These two insights yield a single model that achieves the highest accuracy among all existing unsupervised flow and stereo methods on KITTI 2012 and 2015 benchmarks. More remarkably, our self-supervised method even outperforms several state-ofthe-art fully supervised methods, including PWC-Net and FlowNet2 on KITTI 2012.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers20
- R-MSFM: Recurrent Multi-Scale Feature Modulation for Monocular Depth EstimatingZhongkai Zhou, Xinnan Fan, Pengfei Shi, Yuanxue XinICCV 2021 · 150 citations
- Learning Monocular Depth in Dynamic Scenes via Instance-Aware Projection ConsistencySeokju Lee, Sunghoon Im, Stephen Lin, In So KweonAAAI 2021 · 107 citations
- Displacement-Invariant Matching Cost Learning for Accurate Optical Flow EstimationJianyuan Wang, Yiran Zhong, Yuchao Dai, Kaihao Zhang et al.NeurIPS 2020 · 83 citations
- LSVC: A Learning-based Stereo Video Compression FrameworkZhenghao Chen, Guo Lu, Zhihao Hu, Shan Liu et al.CVPR 2022 · 39 citations
- SelfD: Self-Learning Large-Scale Driving Policies From the WebJimuyang Zhang, Ruizhao Zhu, Eshed Ohn-BarCVPR 2022 · 17 citations
Related papers
- Feature-Level Collaboration: Joint Unsupervised Learning of Optical Flow, Stereo Depth and Camera MotionCheng Chi, Qingjie Wang, Tianyu Hao, Peng Guo et al.CVPR 2021
- Imposing Consistency for Optical Flow EstimationJisoo Jeong, Jamie Menjay Lin, Fatih Porikli, Nojun KwakCVPR 2022 · 40 citations
- Revealing the Reciprocal Relations between Self-Supervised Stereo and Monocular Depth EstimationZhi Chen, Xiaoqing Ye, Wei Yang, Zhenbo Xu et al.ICCV 2021 · 34 citations
- Two-in-One Depth: Bridging the Gap Between Monocular and Binocular Self-supervised Depth EstimationZhengming Zhou, Qiulei DongICCV 2023 · 16 citations
- DualNet: Robust Self-Supervised Stereo Matching with Pseudo-Label SupervisionYun Wang, Jiahao Zheng, Chenghao Zhang, Zhanjie Zhang et al.AAAI 2025 · 12 citations
