Equivariant Multi-Modality Image Fusion
Zixiang Zhao, Haowen Bai, Jiangshe Zhang, Yulun Zhang, Kai Zhang, Shuang Xu, Dongdong Chen, Radu Timofte, Luc Van Gool
Abstract
Multi-modality image fusion is a technique that combines information from different sensors or modalities, enabling the fused image to retain complementary features from each modality, such as functional highlights and texture details. However, effective training of such fusion models is challenging due to the scarcity of ground truth fusion data. To tackle this issue, we propose the Equivariant Multi-Modality imAge fusion (EMMA) paradigm for end-to-end self-supervised learning. Our approach is rooted in the prior knowledge that natural imaging responses are equivariant to certain transformations. Consequently, we introduce a novel training paradigm that encompasses a fusion module, a pseudo-sensing module, and an equivariant fusion module. These components enable the net training to follow the principles of the natural sensing-imaging process while satisfying the equivariant imaging prior. Extensive experiments confirm that EMMA yields high-quality fusion results for infrared-visible and medical images, concurrently facilitating downstream multi-modal segmentation and detection tasks. The code is available at https: //github.com/Zhaozixiang1228/MMIF-EMMA .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers53
- DDFM: Denoising Diffusion Model for Multi-Modality Image FusionZixiang Zhao, Haowen Bai, Yuanzhi Zhu, Jiangshe Zhang et al.ICCV 2023 · 350 citations
- Text-DiFuse: An Interactive Multi-Modal Image Fusion Framework based on Text-modulated Diffusion ModelHao Zhang, Lei Cao, Jiayi MaNeurIPS 2024 · 69 citations
- Spherical Space Feature Decomposition for Guided Depth Map Super-ResolutionZixiang Zhao, Jiangshe Zhang, Xiang Gu, Chengli Tan et al.ICCV 2023 · 55 citations
- Degradation-Resistant Unfolding Network for Heterogeneous Image FusionChunming He, Kai Li, Guoxia Xu, Yulun Zhang et al.ICCV 2023 · 55 citations
- BSAFusion: A Bidirectional Stepwise Feature Alignment Network for Unaligned Medical Image FusionHuafeng Li, Dayong Su, Qing Cai, Yafei ZhangAAAI 2025 · 43 citations
Builds on15
- Target-aware Dual Adversarial Learning and a Multi-scenario Multi-Modality Benchmark to Fuse Infrared and Visible for Object DetectionJinyuan Liu, Xin Fan, Zhanbo Huang, Guanyao Wu et al.CVPR 2022 · 929 citations
- Rethinking the Image Fusion: A Fast Unified Image Fusion Network based on Proportional Maintenance of Gradient and IntensityHao Zhang, Han Xu, Yang Xiao, Xiaojie Guo et al.AAAI 2020 · 583 citations
- FusionDN: A Unified Densely Connected Network for Image FusionHan Xu, Jiayi Ma, Zhuliang Le, Junjun Jiang et al.AAAI 2020 · 559 citations
- DDFM: Denoising Diffusion Model for Multi-Modality Image FusionZixiang Zhao, Haowen Bai, Yuanzhi Zhu, Jiangshe Zhang et al.ICCV 2023 · 350 citations
- Multi-interactive Feature Learning and a Full-time Multi-modality Benchmark for Image Fusion and SegmentationJinyuan Liu, Zhu Liu, Guanyao Wu, Long Ma et al.ICCV 2023 · 287 citations
Related papers
- SigFusion: Unified Signal-Level Self-Supervised Learning Paradigm for Image FusionZeyu Wang, Jiawei Feng, Jiayu Wang, Pengjie Wang et al.AAAI 2026
- Searching a Hierarchically Aggregated Fusion Architecture for Fast Multi-Modality Image FusionRisheng Liu, Zhu Liu, Jinyuan Liu, Xin FanACM MM 2021 · 64 citations
- Beyond Strict Pairing: Arbitrarily Paired Training for High-Performance Infrared and Visible Image FusionYanglin Deng, Tianyang Xu, Chunyang Cheng, Hui Li et al.CVPR 2026
- Self-supervised Multiplex Consensus Mamba for General Image FusionYingying Wang, Rongjin Zhuang, Hui Zheng, Xuanhua He et al.AAAI 2026 · 2 citations
- E2E-MFD: Towards End-to-End Synchronous Multimodal Fusion DetectionJiaqing Zhang, Mingxiang Cao, Weiying Xie, Jie Lei et al.NeurIPS 2024 · 68 citations
