Learning a Graph Neural Network with Cross Modality Interaction for Image Fusion
Jiawei Li, Jiansheng Chen, Jinyuan Liu, Huimin Ma
Abstract
Infrared and visible image fusion has gradually proved to be a vital fork in the field of multi-modality imaging technologies. In recent developments, researchers not only focus on the quality of fused images but also evaluate their performance in downstream tasks. Nevertheless, the majority of methods seldom put their eyes on mutual learning from different modalities, resulting in fused images lacking significant details and textures. To overcome this issue, we propose an interactive graph neural network (GNN)-based architecture between cross modality for fusion, called IGNet. Specifically, we first apply a multi-scale extractor to achieve shallow features, which are employed as the necessary input to build graph structures. Then, the graph interaction module can construct the extracted intermediate features of the infrared/visible branch into graph structures. Meanwhile, the graph structures of two branches interact for cross-modality and semantic learning, so that fused images can maintain the important feature expressions and enhance the performance of downstream tasks. Besides, the proposed leader nodes can improve information propagation in the same modality. Finally, we merge all graph features to get the fusion result. Extensive experiments on different datasets (i.e. TNO, MFNet, and M3FD) demonstrate that our IGNet can generate visually appealing fused images while scoring averagely 2.59% [email protected] and 7.77% mIoU higher in detection and segmentation than the compared state-of-the-art methods. The source code of the proposed IGNet can be available at https://github.com/lok-18/IGNet.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 85a9506b-bd1c-4f46-9202-81e60608cac2Cited by top-tier papers12
- WaveMamba: Wavelet-Driven Mamba Fusion for RGB-Infrared Object DetectionHaodong Zhu, Wenhao Dong, Linlin Yang, Hong Li et al.ICCV 2025 · 38 citations
- SMR-Net: Semantic-Guided Mutually Reinforcing Network for Cross-Modal Image Fusion and Salient Object DetectionGuobao Xiao, Xinyu Liu, Zebin Lin, Rui MingAAAI 2025 · 11 citations
- Residual Prior-driven Frequency-aware Network for Image FusionZheng Guan, Xue Wang, Wenhua Qian, Peng Liu et al.ACM MM 2025 · 10 citations
- A²RNet: Adversarial Attack Resilient Network for Robust Infrared and Visible Image FusionJiawei Li, Hongwei Yu, Jiansheng Chen, Xinlong Ding et al.AAAI 2025 · 6 citations
- DADet: Safeguarding Image Conditional Diffusion Models Against Adversarial and Backdoor Attacks via Diffusion Anomaly DetectionHongwei Yu, Xinlong Ding, Jiawei Li, Jinlong Wang et al.ICCV 2025 · 4 citations
Builds on7
- Target-aware Dual Adversarial Learning and a Multi-scenario Multi-Modality Benchmark to Fuse Infrared and Visible for Object DetectionJinyuan Liu, Xin Fan, Zhanbo Huang, Guanyao Wu et al.CVPR 2022 · 929 citations
- DetFusion: A Detection-driven Infrared and Visible Image Fusion NetworkYiming Sun, Bing Cao, Pengfei Zhu, Qinghua HuACM MM 2022 · 165 citations
- Towards All Weather and Unobstructed Multi-Spectral Image Stitching: Algorithm and BenchmarkZhiying Jiang, Zengxi Zhang, Xin Fan, Risheng LiuACM MM 2022 · 46 citations
- Cross-Patch Graph Convolutional Network for Image DenoisingYao Li, Xueyang Fu, Zheng-Jun ZhaICCV 2021 · 24 citations
- Graph-BAS3Net: Boundary-Aware Semi-Supervised Segmentation Network with Bilateral Graph ConvolutionHuimin Huang, Lanfen Lin, Yue Zhang, Yingying Xu et al.ICCV 2021 · 18 citations
Related papers
- Multi-modal Gated Mixture of Local-to-Global Experts for Dynamic Image FusionBing Cao, Yiming Sun, Pengfei Zhu, Qinghua HuICCV 2023 · 110 citations
- MetaFusion: Infrared and Visible Image Fusion via Meta-Feature Embedding from Object DetectionWenda Zhao, Shigeng Xie, Fan Zhao, You He et al.CVPR 2023
- FusionRegister: Every Infrared and Visible Image Fusion Deserves RegistrationCongcong Bian, Haolong Ma, Hui Li, Zhongwei Shen et al.CVPR 2026 · 3 citations
- CDDFuse: Correlation-Driven Dual-Branch Feature Decomposition for Multi-Modality Image FusionZixiang Zhao, Haowen Bai, Jiangshe Zhang, Yulun Zhang et al.CVPR 2023
- One Model for ALL: Low-Level Task Interaction Is a Key to Task-Agnostic Image FusionChunyang Cheng, Tianyang Xu, Zhenhua Feng, Xiaojun Wu et al.CVPR 2025
