Every SAM Drop Counts: Embracing Semantic Priors for Multi-Modality Image Fusion and Beyond
Guanyao Wu, Haoyu Liu, Hongming Fu, Yichuan Peng, Jinyuan Liu, Xin Fan, Risheng Liu
摘要
Multi-modality image fusion, particularly infrared and visible, plays a crucial role in integrating diverse modalities to enhance scene understanding. Although early research prioritized visual quality, preserving fine details and adapting to downstream tasks remains challenging. Recent approaches attempt task-specific design but rarely achieve "The Best of Both Worlds" due to inconsistent optimization goals. To address these issues, we propose a novel method that leverages the semantic knowledge from the Segment Anything Model (SAM) to Grow the quality of fusion results and Enable downstream task adaptability, namely SAGE. Specifically, we design a Semantic Persistent Attention (SPA) Module that efficiently maintains source information via the persistent repository while extracting high-level semantic priors from SAM. More importantly, to eliminate the impractical dependence on SAM during inference, we introduce a bi-level optimization-driven distillation mechanism with triplet losses, which allow the student network to effectively extract knowledge. Extensive experiments show that our method achieves a balance between high-quality visual results and downstream task adaptability while maintaining practical deployment efficiency. The code is available at https://github . com/RollingPlain/SAGE_IVIF.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper21
- Customized Fusion: A Closed-Loop Dynamic Network for Adaptive Multi-Task-Aware Infrared-Visible Image FusionZengyi Yang, Yu Liu, Juan Cheng, Zhiqin Zhu 等CVPR 2026 · 被引用 7 次
- Text-Guided Channel Perturbation and Pre-Trained Knowledge Integration for Unified Multi-Modality Image FusionXilai Li, Xiaosong Li, Weijun JiangAAAI 2026 · 被引用 4 次
- Image Stitching in Adverse Condition: A Bidirectional-Consistency Learning Framework and BenchmarkZengxi Zhang, Junchen Ge, Zhiying Jiang, Miao Zhang 等NeurIPS 2025 · 被引用 3 次
- Missing No More: Dictionary-Guided Cross-Modal Image Fusion under Missing InfraredYafei Zhang, Meng Ma, Huafeng Li, Yu LiuCVPR 2026 · 被引用 3 次
- Multi-Modal Image Fusion via Intervention-Stable Feature LearningXue Wang, Zheng Guan, Wenhua Qian, Chengchao Wang 等CVPR 2026 · 被引用 3 次
它引用的顶会 Paper17
- SegFormer: Simple and Efficient Design for Semantic Segmentation with TransformersEnze Xie, Wenhai Wang, Zhiding Yu, Anima Anandkumar 等NeurIPS 2021 · 被引用 9,661 次
- Target-aware Dual Adversarial Learning and a Multi-scenario Multi-Modality Benchmark to Fuse Infrared and Visible for Object DetectionJinyuan Liu, Xin Fan, Zhanbo Huang, Guanyao Wu 等CVPR 2022 · 被引用 929 次
- Toward Fast, Flexible, and Robust Low-Light Image EnhancementLong Ma, Tengyu Ma, Risheng Liu, Xin Fan 等CVPR 2022 · 被引用 928 次
- Segment Everything Everywhere All at OnceXueyan Zou, Jianwei Yang, Hao Zhang, Feng Li 等NeurIPS 2023 · 被引用 889 次
- DDFM: Denoising Diffusion Model for Multi-Modality Image FusionZixiang Zhao, Haowen Bai, Yuanzhi Zhu, Jiangshe Zhang 等ICCV 2023 · 被引用 350 次
相关 Paper
- Distilling Semantic Priors from SAM to Efficient Image Restoration ModelsQuan Zhang, Xiaoyu Liu, Wei Li, Hanting Chen 等CVPR 2024 · 被引用 20 次
- Unleashing the Power of Generic Segmentation Model: A Simple Baseline for Infrared Small Target DetectionMingjin Zhang, Chi Zhang, Qiming Zhang, Yunsong Li 等ACM MM 2024 · 被引用 33 次
- SAM-Guided Semantic Knowledge Fusion for Visible-Infrared Object DetectionTing Li, Songtao Li, Shuaifeng Li, Xiaolin Qin 等ACM MM 2025 · 被引用 2 次
- Improving SAM for Camouflaged Object Detection via Dual Stream AdaptersJiaming Liu, Linghe Kong, Guihai ChenICCV 2025 · 被引用 5 次
- Segment Anything Model Meets Semi-supervised Medical Image Segmentation: A Novel PerspectiveHaifeng Zhao, Haiyang Li, Lei-Lei Ma, Dengdi SunNeurIPS 2025 · 被引用 1 次
