Multi-View Graph Convolutional Network for Multimedia Recommendation
Penghang Yu, Zhiyi Tan, Guanming Lu, Bing-Kun Bao
Abstract
Multimedia recommendation has received much attention in recent years. It models user preferences based on both behavior information and item multimodal information. Though current GCN-based methods achieve notable success, they suffer from two limitations: (1) Modality noise contamination to the item representations. Existing methods often mix modality features and behavior features in a single view (e.g., user-item view) for propagation, the noise in the modality features may be amplified and coupled with behavior features. In the end, it leads to poor feature discriminability; (2) Incomplete user preference modeling caused by equal treatment of modality features. Users often exhibit distinct modality preferences when purchasing different items. Equally fusing each modality feature ignores the relative importance among different modalities, leading to the suboptimal user preference modeling.
To tackle the above issues, we propose a novel Multi-View Graph Convolutional Network (MGCN) for the multimedia recommendation. Specifically, to avoid modality noise contamination, the modality features are first purified with the aid of item behavior information. Then, the purified modality features of items and behavior features are enriched in separate views, including the useritem view and the item-item view. In this way, the distinguishability of features is enhanced. Meanwhile, a behavior-aware fuser is designed to comprehensively model user preferences by adaptively learning the relative importance of different modality features. Furthermore, we equip the fuser with a self-supervised auxiliary task. This task is expected to maximize the mutual information between the fused multimodal features and behavior features, so as to capture complementary and supplementary preference information simultaneously. Extensive experiments on three public datasets
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9e03fe02-a144-47b8-8cf0-33926e8a1708Cited by top-tier papers32
- Modality-Independent Graph Neural Networks with Global Transformers for Multimodal RecommendationJun Hu, Bryan Hooi, Bingsheng He, Yinwei WeiAAAI 2025 · 31 citations
- Improving Multi-modal Recommender Systems by Denoising and Aligning Multi-modal Content and User FeedbackGuipeng Xv, Xinyu Li, Ruobing Xie, Chen Lin et al.KDD 2024 · 25 citations
- Mind Individual Information! Principal Graph Learning for Multimedia RecommendationPenghang Yu, Zhiyi Tan, Guanming Lu, Bing-Kun BaoAAAI 2025 · 25 citations
- Multi-Modal Multi-Behavior Sequential Recommendation with Conditional Diffusion-Based Feature DenoisingXiaoxi Cui, Weihai Lu, Yu Tong, Yiheng Li et al.SIGIR 2025 · 21 citations
- Modality-Balanced Learning for Multimedia RecommendationJinghao Zhang, Guofan Liu, Qiang Liu, Shu Wu et al.ACM MM 2024 · 21 citations
Builds on12
- LightGCN: Simplifying and Powering Graph Convolution Network for RecommendationXiangnan He, Kuan Deng, Xiang Wang, Yan Li et al.SIGIR 2020 · 4,448 citations
- Measuring and Relieving the Over-Smoothing Problem for Graph Neural Networks from the Topological ViewDeli Chen, Yankai Lin, Wei Li, Peng Li et al.AAAI 2020 · 1,353 citations
- Revisiting Graph Based Collaborative Filtering: A Linear Residual Graph Convolutional Network ApproachLei Chen, Le Wu, Richang Hong, Kun Zhang et al.AAAI 2020 · 634 citations
- AM-GCN: Adaptive Multi-channel Graph Convolutional NetworksXiao Wang, Meiqi Zhu, Deyu Bo, Peng Cui et al.KDD 2020 · 464 citations
- Graph-Refined Convolutional Network for Multimedia Recommendation with Implicit FeedbackYinwei Wei, Xiang Wang, Liqiang Nie, Xiangnan He et al.ACM MM 2020 · 374 citations
Related papers
- Learning Hybrid Behavior Patterns for Multimedia RecommendationZongshen Mu, Yueting Zhuang, Jie Tan, Jun Xiao et al.ACM MM 2022 · 47 citations
- Multi-behavior Recommendation with Graph Convolutional NetworksBowen Jin, Chen Gao, Xiangnan He, Depeng Jin et al.SIGIR 2020 · 420 citations
- Breaking Isolation: Multimodal Graph Fusion for Multimedia Recommendation by Edge-wise ModulationFeiyu Chen, Junjie Wang, Yinwei Wei, Hai-Tao Zheng et al.ACM MM 2022 · 33 citations
- LightGT: A Light Graph Transformer for Multimedia RecommendationYinwei Wei, Wenqi Liu, Fan Liu, Xiang Wang et al.SIGIR 2023 · 71 citations
- COHESION: Composite Graph Convolutional Network with Dual-Stage Fusion for Multimodal RecommendationJinfeng Xu, Zheyu Chen, Wei Wang, Xiping Hu et al.SIGIR 2025 · 20 citations
