Promoting Single-Modal Optical Flow Network for Diverse Cross-Modal Flow Estimation
Shili Zhou, Weimin Tan, Bo Yan
摘要
In recent years, optical flow methods develop rapidly, achieving unprecedented high performance. Most of the methods only consider single-modal optical flow under the well-known brightness-constancy assumption. However, in many application systems, images of different modalities need to be aligned, which demands to estimate cross-modal flow between the cross-modal image pairs. A lot of cross-modal matching methods are designed for some specific cross-modal scenarios. We argue that the prior knowledge of the advanced optical flow models can be transferred to the cross-modal flow estimation, which may be a simple but unified solution for diverse cross-modal matching tasks. To verify our hypothesis, we design a self-supervised framework to promote the single-modal optical flow networks for diverse corss-modal flow estimation. Moreover, we add a Cross-Modal-Adapter block as a plugin to the state-of-the-art optical flow model RAFT for better performance in cross-modal scenarios. Our proposed Modality Promotion Framework and Cross-Modal Adapter have multiple advantages compared to the existing methods. The experiments demonstrate that our method is effective on multiple datasets of different cross-modal scenarios.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Rethinking Unsupervised Cross-modal Flow Estimation: Learning from Decoupled Optimization and Consistency ConstraintRunmin Zhang, Jialiang Wang, Si-Yuan Cao, Zhu Yu 等ICLR 2026 · 被引用 1 次
- AerialFusion: Co-Motion-Driven Unified Registration and Fusion on Multi-modal Data Streams from Aerial ViewJunhui Qiu, Xiang Xiang, Hongyun Wang, Jiaqi GuiAAAI 2026
- A Hybrid Space Model for Misaligned Multi-modality Image FusionYi Xiao, Jia Wang, Zhu Liu, Di Wang 等AAAI 2026
它引用的顶会 Paper4
- Unsupervised Multi-Modal Image Registration via Geometry Preserving Image-to-Image TranslationMoab Arar, Yiftach Ginger, Dov Danon, Amit H. Bermano 等CVPR 2020
- Cross-Spectral Face Hallucination via Disentangling Independent FactorsBoyan Duan, Chaoyou Fu, Yi Li, Xingguang Song 等CVPR 2020
- Learning by Analogy: Reliable Supervision From Transformations for Unsupervised Optical Flow EstimationLiang Liu, Jiangning Zhang, Ruifei He, Yong Liu 等CVPR 2020
- GLU-Net: Global-Local Universal Network for Dense Flow and CorrespondencesPrune Truong, Martin Danelljan, Radu TimofteCVPR 2020
相关 Paper
- GMFlow: Learning Optical Flow via Global MatchingHaofei Xu, Jing Zhang, Jianfei Cai, Hamid Rezatofighi 等CVPR 2022 · 被引用 353 次
- CRAFT: Cross-Attentional Flow Transformer for Robust Optical FlowXiuchao Sui, Shaohua Li, Xue Geng, Yan Wu 等CVPR 2022 · 被引用 114 次
- SMURF: Self-Teaching Multi-Frame Unsupervised RAFT With Full-Image WarpingAustin Stone, Daniel Maurer, Alper Ayvaci, Anelia Angelova 等CVPR 2021
- CRFT: Consistent-Recurrent Feature Flow Transformer for Cross-Modal Image RegistrationXuecong Liu, Mengzhu Ding, Zixuan Sun, Zhang Li 等CVPR 2026 · 被引用 4 次
- ADFactory: An Effective Framework for Generalizing Optical Flow With NeRFHan Ling, Quansen Sun, Yinghui Sun, Xian Xu 等CVPR 2024
