Generalizable Multi-Pass Training of Ads Recommendation Models with Foundation Model Guidance
Yunzhe Qi, Qinghai Zhou, Boyang Liu, Can Cui, Hua Zheng, Qiuling Suo, Mingfu Liang, Kenny Lov, Xi Liu, Shali Jiang, Laming Chen, Wen-Yen Chen
摘要
High-capacity Click-Through Rate (CTR) models for ads recommendation often exhibit pronounced one-epoch overfitting: performance peaks after a single training pass (epoch) over the data, while additional epochs degrade generalization as the model memorizes noise in high-variance click labels. To address this challenge, we propose TeMPO, a principled framework for generalizable multi-pass training of recommendation models that leverages supervision from a teacher recommendation Foundation Model (FM). Our key idea is to move beyond maximizing the conditional likelihood of noisy click labels. Instead, we maximize the joint likelihood of observing both click labels and the teacher's rich representational knowledge, combining the task loss with teacher-guided alignment objectives. We propose a two-stage optimization strategy to stabilize multi-pass learning. Our theoretical analysis shows that distillation acts as complexity regularization, yielding tighter generalization bounds than standard empirical risk minimization under noisy labels. Extensive experiments on public benchmarks and Meta's production-scale ads application demonstrate that TeMPO enables additional training passes to improve, rather than degrade, performance.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- Generalize for Future: Slow and Fast Trajectory Learning for CTR PredictionJian Zhu, Congcong Liu, Xue Jiang, Changping Peng 等AAAI 2024 · 被引用 2 次
- A Preference Learning Decoupling Framework for User Cold-Start RecommendationChunyang Wang, Yanmin Zhu, Aixin Sun, Zhaobo Wang 等SIGIR 2023 · 被引用 15 次
- Reconsidering Learning Objectives in Unbiased Recommendation: A Distribution Shift PerspectiveTeng Xiao, Zhengyu Chen, Suhang WangKDD 2023 · 被引用 8 次
- Recurrent Meta-Learning against Generalized Cold-start Problem in CTR PredictionJunyu Chen, Qianqian Xu, Zhiyong Yang, Ke Ma 等ACM MM 2022 · 被引用 2 次
- Meta-Learning with Self-Improving Momentum TargetJihoon Tack, Jongjin Park, Hankook Lee, Jaeho Lee 等NeurIPS 2022 · 被引用 17 次
