ClickPrompt: CTR Models are Strong Prompt Generators for Adapting Language Models to CTR Prediction
Jianghao Lin, Bo Chen, Hangyu Wang, Yunjia Xi, Yanru Qu, Xinyi Dai, Kangning Zhang, Ruiming Tang, Yong Yu, Weinan Zhang
Abstract
Click-through rate (CTR) prediction has become increasingly indispensable for various Internet applications. Traditional CTR models convert the multi-field categorical data into ID features via one-hot encoding, and extract the collaborative signals among features. Such a paradigm suffers from the problem of semantic information loss. Another line of research explores the potential of pretrained language models (PLMs) for CTR prediction by converting input data into textual sentences through hard prompt templates. Although semantic signals are preserved, they generally fail to capture the collaborative information (e.g., feature interactions, pure ID features), not to mention the unacceptable inference overhead brought by the huge model size. In this paper, we aim to model both the semantic knowledge and collaborative knowledge for accurate CTR estimation, and meanwhile address the inference inefficiency issue. To benefit from both worlds and close their gaps, we propose a novel model-agnostic framework (i.e., ClickPrompt), where we incorporate CTR models to generate interaction-aware soft prompts for PLMs. We design a prompt-augmented masked language modeling (PA-MLM) pretraining task, where PLM has to recover the masked tokens based on the language context, as well as the soft prompts generated by CTR model. The collaborative and semantic knowledge from ID and textual features would be explicitly aligned and interacted via the prompt interface. Then, we can either tune the CTR model with PLM for superior performance, or solely tune the CTR model without PLM for inference efficiency. Experiments on four real-world datasets validate the effectiveness of ClickPrompt compared with existing baselines.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext dce802e2-8306-42c5-891f-517e54cc3ee7Cited by top-tier papers9
- M-scan: A Multi-Scenario Causal-driven Adaptive Network for RecommendationJiachen Zhu, Yichao Wang, Jianghao Lin, Jiarui Qin et al.WWW 2024 · 6 citations
- Field Matters: A Lightweight LLM-enhanced Method for CTR PredictionYu Cui, Feng Liu, Jiawei Chen, Xingyu Lou et al.WWW 2026 · 5 citations
- Efficiency Unleashed: Inference Acceleration for LLM-based Recommender Systems with Speculative DecodingYunjia Xi, Hangyu Wang, Bo Chen, Jianghao Lin et al.SIGIR 2025 · 5 citations
- PepRec: Progressive Enhancement of Prompting for RecommendationYakun Yu, Shiang Qi, Baochun Li, Di NiuEMNLP 2024 · 2 citations
- RecBase: Generative Foundation Model Pretraining for Zero-Shot RecommendationSashuai Zhou, Weinan Gan, Qijiong Liu, Ke Lei et al.EMNLP 2025 · 1 citation
Builds on16
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- DCN V2: Improved Deep & Cross Network and Practical Lessons for Web-scale Learning to Rank SystemsRuoxi Wang, Rakesh Shivanna, Derek Zhiyuan Cheng, Sagar Jain et al.WWW 2021 · 793 citations
- Towards Universal Sequence Representation Learning for Recommender SystemsYupeng Hou, Shanlei Mu, Wayne Xin Zhao, Yaliang Li et al.KDD 2022 · 245 citations
- ReLLa: Retrieval-enhanced Large Language Models for Lifelong Sequential Behavior Comprehension in RecommendationJianghao Lin, Rong Shan, Chenxu Zhu, Kounianhua Du et al.WWW 2024 · 151 citations
- Text Is All You Need: Learning Language Representations for Sequential RecommendationJiacheng Li, Ming Wang, Jin Li, Jinmiao Fu et al.KDD 2023 · 134 citations
Related papers
- Topic Guided Multi-faceted Semantic Disentanglement for CTR predictionFengxin Li, Zhiqian Yin, Hongyan Liu, Jingcai Guo et al.ACM MM 2025
- MGTA: Multi-scale Graph Tokens Alignment for CTR Prediction via Pre-trained Language ModelsZhongzhen Wu, Yating Ren, Shuochen Li, Huobin TanKDD 2026
- HPT: Hierarchy-aware Prompt Tuning for Hierarchical Text ClassificationZihan Wang, Peiyi Wang, Tianyu Liu, Binghuai Lin et al.EMNLP 2022 · 42 citations
- MAP: A Model-agnostic Pretraining Framework for Click-through Rate PredictionJianghao Lin, Yanru Qu, Wei Guo, Xinyi Dai et al.KDD 2023 · 28 citations
- Prompt Learning for News RecommendationZizhuo Zhang, Bang WangSIGIR 2023 · 76 citations
