MDFL: Multi-Domain Diffusion-Driven Feature Learning
Daixun Li, Weiying Xie, Jiaqing Zhang, Yunsong Li
Abstract
High-dimensional images, known for their rich semantic information, are widely applied in remote sensing and other fields. The spatial information in these images reflects the object's texture features, while the spectral information reveals the potential spectral representations across different bands. Currently, the understanding of high-dimensional images remains limited to a single-domain perspective with performance degradation. Motivated by the masking texture effect observed in the human visual system, we present a multidomain diffusion-driven feature learning network (MDFL) , a scheme to redefine the effective information domain that the model really focuses on. This method employs diffusionbased posterior sampling to explicitly consider joint information interactions between the high-dimensional manifold structures in the spectral, spatial, and frequency domains, thereby eliminating the influence of masking texture effects in visual models. Additionally, we introduce a feature reuse mechanism to gather deep and raw features of high-dimensional data. We demonstrate that MDFL significantly improves the feature extraction performance of highdimensional data, thereby providing a powerful aid for revealing the intrinsic patterns and structures of such data. The experimental results on three multi-modal remote sensing datasets show that MDFL reaches an average overall accuracy of 98.25%, outperforming various state-of-the-art baseline schemes. The code will be released, contributing to the computer vision community.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 51ae2ea2-47e2-4669-8216-f13c04b59157Cited by top-tier papers4
- Structure-Adaptive Multi-View Graph Clustering for Remote Sensing DataRenxiang Guan, Wenxuan Tu, Siwei Wang, Jiyuan Liu et al.AAAI 2025 · 26 citations
- FusionSAM: Visual Multi-Modal Learning with Segment Anything ModelDaixun Li, Weiying Xie, Mingxiang Cao, Yunke Wang et al.KDD 2025 · 2 citations
- Cross-View Progressive Feature Filtering for Multi-View Graph Clustering in Remote SensingBowen Liu, Xin Peng, Wenxuan Tu, Chengyao Wei et al.AAAI 2026
- Can Generative Geospatial Diffusion Models Excel as Discriminative Geospatial Foundation Models?Yuru Jia, Valerio Marsocci, Ziyang Gong, Xue Yang et al.ICCV 2025
Builds on12
- Improved Denoising Diffusion Probabilistic ModelsAlexander Quinn Nichol, Prafulla DhariwalICML 2021 · 5,234 citations
- Elucidating the Design Space of Diffusion-Based Generative ModelsTero Karras, Miika Aittala, Timo Aila, Samuli LaineNeurIPS 2022 · 3,959 citations
- Multiscale Vision TransformersHaoqi Fan, Bo Xiong, Karttikeya Mangalam, Yanghao Li et al.ICCV 2021 · 1,611 citations
- Label-Efficient Semantic Segmentation with Diffusion ModelsDmitry Baranchuk, Andrey Voynov, Ivan Rubachev, Valentin Khrulkov et al.ICLR 2022 · 700 citations
- MAT: Mask-Aware Transformer for Large Hole Image InpaintingWenbo Li, Zhe Lin, Kun Zhou, Lu Qi et al.CVPR 2022 · 382 citations
Related papers
- Frequency-Aware Vision-Language Multimodality Generalization Network for Remote Sensing Image ClassificationJunjie Zhang, Feng Zhao, Hanqiang Liu, Jun YuAAAI 2026
- Revealing the Invisible: Latent Structure Modeling for Semantically Consistent Cloud RemovalJingwei Xin, Kai Guo, Jie Li, Nannan WangAAAI 2026
- Effective Cloud Removal for Remote Sensing Images by an Improved Mean-Reverting Denoising Model with Elucidated Design SpaceYi Liu, Wengen Li, Jihong Guan, Shuigeng Zhou et al.CVPR 2025
- EMR-Diff: Edge-aware Multimodal Residual Diffusion Model for Hyperspectral Image Super-resolutionTao Zhang, Shengtao Yao, Rong Zeng, Zunjie Zhu et al.CVPR 2026
- Remote Sensing Image Super-Resolution for Imbalanced Textures: A Texture-Aware Diffusion FrameworkEnzhuo Zhang, Sijie Zhao, Dilxat Muhtar, Zhenshi Li et al.CVPR 2026
