Adaptive Anti-Bottleneck Multi-Modal Graph Learning Network for Personalized Micro-video Recommendation
Desheng Cai, Shengsheng Qian, Quan Fang, Jun Hu, Changsheng Xu
Abstract
Micro-video recommendation has attracted extensive research attention with the increasing popularity of micro-video sharing platforms. There exists a substantial amount of excellent efforts made to the micro-video recommendation task. Recently, homogeneous (or heterogeneous) GNN-based approaches utilize graph convolutional operators (or meta-path based similarity measures) to learn meaningful representations for users and micro-videos and show promising performance for the micro-video recommendation task. However, these methods may suffer from the following problems: (1) fail to aggregate information from distant or long-range nodes; (2) ignore the varying intensity of users' preferences for different items in micro-video recommendations; (3) neglect the similarities of multi-modal contents of micro-videos for recommendation tasks. In this paper, we propose a novel Adaptive Anti-Bottleneck Multi-Modal Graph Learning Network for personalized micro-video recommendation. Specifically, we design a collaborative representation learning module and a semantic representation learning module to fully exploit user-video interaction information and the similarities of micro-videos, respectively. Furthermore, we utilize an anti-bottleneck module to automatically learn the importance weights of short-range and long-range neighboring nodes to obtain more expressive representations of users and micro-videos. Finally, to consider the varying intensity of users' preferences for different micro-videos, we design and optimize an adaptive recommendation loss to train our model in an end-to-end manner. We evaluate our method on three real-world datasets and the results demonstrate that the proposed model outperforms the baselines.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get b57635ba-d4bd-4c35-a044-b3b4bfe8ffc8Cited by top-tier papers7
- Constructing Holistic Spatio-Temporal Scene Graph for Video Semantic Role LabelingYu Zhao, Hao Fei, Yixin Cao, Bobo Li et al.ACM MM 2023 · 31 citations
- Modality-Independent Graph Neural Networks with Global Transformers for Multimodal RecommendationJun Hu, Bryan Hooi, Bingsheng He, Yinwei WeiAAAI 2025 · 31 citations
- Contrastive Intra- and Inter-Modality Generation for Enhancing Incomplete Multimedia RecommendationZhenghong Lin, Yanchao Tan, Yunfei Zhan, Weiming Liu et al.ACM MM 2023 · 27 citations
- Unveiling the Impact of Multi-modal Content in Multi-modal Recommender SystemsGuipeng Xv, Xinyu Li, Yi Liu, Chen Lin et al.ACM MM 2025 · 3 citations
- Exploiting Fine-Grained Skip Behaviors for Micro-Video RecommendationSanghyuck Lee, Sangkeun Park, Jaesung LeeAAAI 2025 · 2 citations
Related papers
- Multi-modal Bipartite Graph Structure Learning with Information Bottleneck for Micro-video RecommendationYing He, Desheng Cai, Shengsheng Qian, Quan Fang et al.WWW 2026
- Learning Fine-grained User Interests for Micro-video RecommendationYu Shang, Chen Gao, Jiansheng Chen, Depeng Jin et al.SIGIR 2023 · 19 citations
- What Aspect Do You Like: Multi-scale Time-aware User Interest Modeling for Micro-video RecommendationHao Jiang, Wenjie Wang, Yinwei Wei, Zan Gao et al.ACM MM 2020 · 65 citations
- Micro-video Tagging via Jointly Modeling Social Influence and Tag RelationXiao Wang, Tian Gan, Yinwei Wei, Jianlong Wu et al.ACM MM 2022 · 8 citations
- Attentional Graph Convolutional Networks for Knowledge Concept Recommendation in MOOCs in a Heterogeneous ViewJibing Gong, Shen Wang, Jinlong Wang, Wenzheng Feng et al.SIGIR 2020 · 180 citations
