Lune

KDD2026顶会

Generalizable Multi-Pass Training of Ads Recommendation Models with Foundation Model Guidance

Yunzhe Qi, Qinghai Zhou, Boyang Liu, Can Cui, Hua Zheng, Qiuling Suo, Mingfu Liang, Kenny Lov, Xi Liu, Shali Jiang, Laming Chen, Wen-Yen Chen

2026年份

摘要

High-capacity Click-Through Rate (CTR) models for ads recommendation often exhibit pronounced one-epoch overfitting: performance peaks after a single training pass (epoch) over the data, while additional epochs degrade generalization as the model memorizes noise in high-variance click labels. To address this challenge, we propose TeMPO, a principled framework for generalizable multi-pass training of recommendation models that leverages supervision from a teacher recommendation Foundation Model (FM). Our key idea is to move beyond maximizing the conditional likelihood of noisy click labels. Instead, we maximize the joint likelihood of observing both click labels and the teacher's rich representational knowledge, combining the task loss with teacher-guided alignment objectives. We propose a two-stage optimization strategy to stabilize multi-pass learning. Our theoretical analysis shows that distillation acts as complexity regularization, yielding tighter generalization bounds than standard empirical risk minimization under noisy labels. Extensive experiments on public benchmarks and Meta's production-scale ads application demonstrate that TeMPO enables additional training passes to improve, rather than degrade, performance.

问问这篇 Paper

问问你的智能体。

Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。

可以从这些问题问起

智能体调用

Lunesearch_papers

在 Lune 里问

免费开始,无需绑卡

相关 Paper

黄昏的海面,两侧是细线勾勒的悬崖