Rethinking Unsupervised Cross-modal Flow Estimation: Learning from Decoupled Optimization and Consistency Constraint
Runmin Zhang, Jialiang Wang, Si-Yuan Cao, Zhu Yu, Junchen Yu, Guangyi Zhang, Hui-liang Shen
Abstract
This work presents DCFlow, a novel self-supervised cross-modal flow estimation framework that integrates a decoupled optimization strategy and a cross-modal consistency constraint. Unlike previous unsupervised approaches that implicitly learn flow estimation solely from appearance similarity, we introduce a decoupled optimization strategy with task-specific supervision to address modality discrepancy and geometric misalignment distinctly. This is achieved by collaboratively training a modality transfer network and a flow estimation network. To enable reliable motion supervision without ground-truth flow, we propose a geometry-aware data synthesis pipeline combined with an outlier-robust loss. Additionally, we introduce a cross-modal consistency constraint to jointly optimize both networks, significantly improving flow prediction accuracy. For evaluation, we construct a comprehensive cross-modal flow benchmark by repurposing public datasets. Experimental results demonstrate that DCFlow can be integrated with various flow estimation networks and achieves state-of-the-art performance among unsupervised approaches.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 30bf6845-4e3a-4fbe-9378-2c586f1c5293Builds on17
- Learning to Estimate Hidden Motions with Global Motion AggregationShihao Jiang, Dylan Campbell, Yao Lu, Hongdong Li et al.ICCV 2021 · 402 citations
- RFNet: Unsupervised Network for Mutually Reinforcing Multi-modal Image Registration and FusionHan Xu, Jiayi Ma, Jiteng Yuan, Zhuliang Le et al.CVPR 2022 · 161 citations
- UniDepth: Universal Monocular Metric Depth EstimationLuigi Piccinelli, Yung-Hsu Yang, Christos Sakaridis, Mattia Segù et al.CVPR 2024 · 122 citations
- Promoting Single-Modal Optical Flow Network for Diverse Cross-Modal Flow EstimationShili Zhou, Weimin Tan, Bo YanAAAI 2022 · 33 citations
- Towards Robust Image Stitching: An Adaptive Resistance Learning against Compatible AttacksZhiying Jiang, Xingyuan Li, Jinyuan Liu, Xin Fan et al.AAAI 2024 · 16 citations
Related papers
- MAPConNet: Self-supervised 3D Pose Transfer with Mesh and Point Contrastive LearningJiaze Sun, Zhixiang Chen, Tae-Kyun KimICCV 2023 · 2 citations
- Learning by Analogy: Reliable Supervision From Transformations for Unsupervised Optical Flow EstimationLiang Liu, Jiangning Zhang, Ruifei He, Yong Liu et al.CVPR 2020
- Self-Supervised Bird's Eye View Motion Prediction with Cross-Modality SignalsShaoheng Fang, Zuhong Liu, Mingyu Wang, Chenxin Xu et al.AAAI 2024 · 8 citations
- Flow2Stereo: Effective Self-Supervised Learning of Optical Flow and Stereo MatchingPengpeng Liu, Irwin King, Michael R. Lyu, Jia XuCVPR 2020
- Unsupervised Space-Time Network for Temporally-Consistent Segmentation of Multiple MotionsEtienne Meunier, Patrick BouthemyCVPR 2023
