DIGEST: Dynamic Graph Refinement with Dual Contrastive Semantic Transfer for Multimodal Recommendation
Xiangyu Sai, Meysam Madadi, Sergio Escalera, Yong Xu
Abstract
Multimodal recommendation benefits from leveraging rich content signals such as images and texts to alleviate interaction sparsity, yet existing graph-based approaches are still hindered by (i) noisy user—item edges that are treated as static during training and (ii) inconsistent representation spaces across interaction-driven and modality-induced graph views. To address these issues, we propose DIGEST, a multi-graph framework that propagates trainable ID embeddings on a denoised user—item graph and a fused modality-induced item—item graph, and interleaves message passing with dynamic graph refinement that iteratively reweights existing edges to suppress noisy connections. To enable reliable semantic transfer across views, DIGEST further introduces a dual contrastive alignment that (i) aligns the collaborative and semantic item views and (ii) constrains the semantic graph representations to projected multimodal features, together with a lightweight dimension decorrelation regularizer and adaptive gated fusion to reduce redundancy and stabilize multi-view learning. Extensive experiments on three Amazon benchmark datasets demonstrate that DIGEST consistently outperforms state-of-the-art multimodal recommenders, achieving up to 8.43% relative improvement on NDCG@20 and 7.66% on Recall@20 over the strongest baselines.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 5348996b-e302-4c6c-8a17-d49a1bade706Related papers
- DHMRec: Collaboration-Guided Multimodal Disentanglement and Hierarchical Fusion for RecommendationXiaohan Zhan, Yuliang Shi, Jihu Wang, Shijun Liu et al.AAAI 2026
- Improving Multi-modal Recommender Systems by Denoising and Aligning Multi-modal Content and User FeedbackGuipeng Xv, Xinyu Li, Ruobing Xie, Chen Lin et al.KDD 2024 · 25 citations
- Multi-Modality is All You Need for Transferable Recommender SystemsYouhua Li, Hanwen Du, Yongxin Ni, Pengpeng Zhao et al.ICDE 2024 · 20 citations
- A Tale of Two Graphs: Freezing and Denoising Graph Structures for Multimodal RecommendationXin Zhou, Zhiqi ShenACM MM 2023 · 234 citations
- Frequency-refined Graph Convolution Network with Cross-modal Wavelet Denoising for RecommendationFeiyu Peng, Chaobo He, Junwei Cheng, Huijuan Hu et al.ACM MM 2025 · 6 citations
