Collaboration of Large Language Models and Small Recommendation Models for Device-Cloud Recommendation
Zheqi Lv, Tianyu Zhan, Wenjie Wang, Xinyu Lin, Shengyu Zhang, Wenqiao Zhang, Jiwei Li, Kun Kuang, Fei Wu
摘要
Large Language Models (LLMs) for Recommendation (LLM4Rec) is a promising research direction that has demonstrated exceptional performance in this field. However, its inability to capture real-time user preferences greatly limits the practical application of LLM4Rec because (i) LLMs are costly to train and infer frequently, and (ii) LLMs struggle to access real-time data (its large number of parameters poses an obstacle to deployment on devices). Fortunately, small recommendation models (SRMs) can effectively supplement these shortcomings of LLM4Rec diagrams by consuming minimal resources for frequent training and inference, and by conveniently accessing real-time data on devices. In light of this, we designed the Device-Cloud LLM-SRM Collaborative Recommendation Framework (LSC4Rec) under a device-cloud collaboration setting. LSC4Rec aims to integrate the advantages of both LLMs and SRMs, as well as the benefits of cloud and edge computing, achieving a complementary synergy. We enhance the practicability of LSC4Rec by designing three strategies: collaborative training, collaborative inference, and intelligent request. During training, LLM generates candidate lists to enhance the ranking ability of SRM in collaborative scenarios and enables SRM to update adaptively to capture real-time user interests. During inference, LLM and SRM are deployed on the cloud and on the device, respectively. LLM generates candidate lists and initial ranking results based on user behavior, and SRM get reranking results based on the candidate list, with final results integrating
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper11
- GraphCLIP: Enhancing Transferability in Graph Foundation Models for Text-Attributed GraphsYun Zhu, Haizhou Shi, Xiaotang Wang, Yongchao Liu 等WWW 2025 · 被引用 54 次
- Disentangled Knowledge Tracing for Alleviating Cognitive BiasYiyun Zhou, Zheqi Lv, Shengyu Zhang, Jingyuan ChenWWW 2025 · 被引用 18 次
- ThinkRec: Thinking-based recommendation via LLMQihang Yu, Kairui Fu, Zheqi Lv, Shengyu Zhang 等WWW 2026 · 被引用 10 次
- Joint Similarity Item Exploration and Overlapped User Guidance for Multi-Modal Cross-Domain RecommendationWeiming Liu, Chaochao Chen, Jiahe Xu, Xinting Liao 等WWW 2025 · 被引用 3 次
- Decoding Correlation-Induced Misalignment in the Stable Diffusion Workflow for Text-to-Image GenerationYunze Tong, Fengda Zhang, Didi Zhu, Jun Xiao 等ICCV 2025 · 被引用 1 次
它引用的顶会 Paper25
- NExT-GPT: Any-to-Any Multimodal LLMShengqiong Wu, Hao Fei, Leigang Qu, Wei Ji 等ICML 2024 · 被引用 786 次
- Intent Contrastive Learning for Sequential RecommendationYongjun Chen, Zhiwei Liu, Jia Li, Julian J. McAuley 等WWW 2022 · 被引用 429 次
- Mining Latent Structures for Multimedia RecommendationJinghao Zhang, Yanqiao Zhu, Qiang Liu, Shu Wu 等ACM MM 2021 · 被引用 350 次
- Is ChatGPT Good at Search? Investigating Large Language Models as Re-Ranking AgentsWeiwei Sun, Lingyong Yan, Xinyu Ma, Shuaiqiang Wang 等EMNLP 2023 · 被引用 182 次
- Denoising Diffusion Recommender ModelJujia Zhao, Wenjie Wang, Yiyan Xu, Teng Sun 等SIGIR 2024 · 被引用 86 次
相关 Paper
- LSRP: A Leader-Subordinate Retrieval Framework for Privacy-Preserving Cloud-Device CollaborationYingyi Zhang, Pengyue Jia, Xianneng Li, Derong Xu 等KDD 2025 · 被引用 2 次
- Lost in Sequence: Do Large Language Models Understand Sequential Recommendation?Sein Kim, Hongseok Kang, Kibum Kim, Jiwan Kim 等KDD 2025 · 被引用 3 次
- Semantic Convergence: Harmonizing Recommender Systems via Two-Stage Alignment and Behavioral Semantic TokenizationGuanghan Li, Xun Zhang, Yufei Zhang, Yifan Yin 等AAAI 2025 · 被引用 18 次
- A Structure-Agnostic Co-Tuning Framework for LLMs and SLMs in Cloud-Edge SystemsYuze Liu, Yunhan Wang, Tiehua Zhang, Zhishu Shen 等WWW 2026 · 被引用 1 次
- DELRec: Distilling Sequential Pattern to Enhance LLMs-Based Sequential RecommendationHaoyi Zhang, Guohao Sun, Jinhu Lu, Guanfeng Liu 等ICDE 2025 · 被引用 1 次
