CRFT: Consistent-Recurrent Feature Flow Transformer for Cross-Modal Image Registration
Xuecong Liu, Mengzhu Ding, Zixuan Sun, Zhang Li, Xichao Teng
摘要
We present Consistent-Recurrent Feature Flow Transformer (CRFT), a unified coarse-to-fine framework based on feature flow learning for robust cross-modal image registration. CRFT learns a modality-independent feature flow representation within a transformer-based architecture that jointly performs feature alignment and flow estimation. The coarse stage establishes global correspondences through multi-scale feature correlation, while the fine stage refines local details via hierarchical feature fusion and adaptive spatial reasoning. To enhance geometric adaptability, an iterative discrepancy-guided attention mechanism with a Spatial Geometric Transform (SGT) recurrently refines the flow field, progressively capturing subtle spatial inconsistencies and enforcing feature-level consistency. This design enables accurate alignment under large affine and scale variations while maintaining structural coherence across modalities. Extensive experiments on diverse cross-modal datasets demonstrate that CRFT consistently outperforms state-of-the-art registration methods in both accuracy and robustness. Beyond registration, CRFT provides a generalizable paradigm for multimodal spatial correspondence, offering broad applicability to remote sensing, autonomous navigation, and medical imaging. Code and datasets are publicly available at https://github.com/NEU-Liuxuecong/CRFT.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper29
- LightGlue: Local Feature Matching at Light SpeedPhilipp Lindenberger, Paul-Edouard Sarlin, Marc PollefeysICCV 2023 · 被引用 936 次
- GMFlow: Learning Optical Flow via Global MatchingHaofei Xu, Jing Zhang, Jianfei Cai, Hamid Rezatofighi 等CVPR 2022 · 被引用 353 次
- SARDet-100K: Towards Open-Source Benchmark and ToolKit for Large-Scale SAR Object DetectionYuxuan Li, Xiang Li, Weijie Li, Qibin Hou 等NeurIPS 2024 · 被引用 145 次
- XFeat: Accelerated Features for Lightweight Image MatchingGuilherme A. Potje, Felipe Cadar, André Araújo, Renato Martins 等CVPR 2024 · 被引用 128 次
- Efficient LoFTR: Semi-Dense Local Feature Matching with Sparse-Like SpeedYifan Wang, Xingyi He, Sida Peng, Dongli Tan 等CVPR 2024 · 被引用 126 次
相关 Paper
- A Consistency-Aware Spot-Guided Transformer for Versatile and Hierarchical Point Cloud RegistrationRenlang Huang, Yufan Tang, Jiming Chen, Liang LiNeurIPS 2024 · 被引用 17 次
- Auto-Regressive Transformation for Image AlignmentKanggeon Lee, Soochahn Lee, Kyoung Mu LeeICCV 2025 · 被引用 1 次
- 2D3D-MATR: 2D-3D Matching Transformer for Detection-free Registration between Images and Point CloudsMinhao Li, Zheng Qin, Zhirui Gao, Renjiao Yi 等ICCV 2023 · 被引用 30 次
- Affine Medical Image Registration with Coarse-to-Fine Vision TransformerTony C. W. Mok, Albert C. S. ChungCVPR 2022 · 被引用 94 次
- Correlation-aware Coarse-to-fine MLPs for Deformable Medical Image RegistrationMingyuan Meng, Dagan Feng, Lei Bi, Jinman KimCVPR 2024 · 被引用 47 次
