Interpretable Matching of Optical-SAR Image via Dynamically Conditioned Diffusion Models
Shuiping Gou, Xin Wang, Xinlin Wang, Yunzhi Chen
摘要
Driven by the complementary information fusion of optical and synthetic aperture radar (SAR) images, the optical-SAR image matching has drawn much attention. However, the significant radiometric differences between them imposes great challenges on accurate matching. Most existing approaches convert SAR and optical images into a shared feature space to perform the matching, but these methods often fail to achieve the robust matching since the feature spaces are unknown and uninterpretable. Motivated by the interpretable latent space of diffusion models, this paper formulates an optical-SAR image translation and matching framework via a dynamically conditioned diffusion model (DCDM) to achieve the interpretable and robust optical-SAR cross-modal image matching. Specifically, in the denoising process, to filter out outlier matching regions, a gated dynamic sparse cross-attention module is proposed to facilitate efficient and effective long-range interactions of multi-grained features between the cross-modal data. In addition, a spatial position consistency constraint is designed to promote the cross-attention features to perceive the spatial corresponding relation in different modalities, improving the matching precision. Experimental results demonstrate that the proposed method outperforms state-of-the-art methods in terms of both the matching accuracy and the interpretability.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- Latent Space Consistency for Sparse-View CT ReconstructionDuoyou Chen, Yunqing Chen, Can Zhang, Zhou Wang 等ACM MM 2025 · 被引用 1 次
- SAR-DisentDM: A Semantic-Disentangled Diffusion Model for Limited-Data SAR Image SynthesisYue Yang, Song Tang, Qijun Zhao, Hailun Zhang 等AAAI 2026 · 被引用 1 次
- Satellite to GroundScape - Large-scale Consistent Ground View Generation from Satellite ViewsNingli Xu, Rongjun QinCVPR 2025
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Any2Any: Unified Arbitrary Modality Translation for Remote SensingHaoyang Chen, Jing Zhang, Di Wang, Hebaixu Wang 等ICML 2026 · 被引用 5 次
