DCN V2: Improved Deep & Cross Network and Practical Lessons for Web-scale Learning to Rank Systems
Ruoxi Wang, Rakesh Shivanna, Derek Zhiyuan Cheng, Sagar Jain, Dong Lin, Lichan Hong, Ed H. Chi
Abstract
Learning effective feature crosses is the key behind building recommender systems. However, the sparse and large feature space requires exhaustive search to identify effective crosses. Deep & Cross Network (DCN) was proposed to automatically and efficiently learn bounded-degree predictive feature interactions. Unfortunately, in models that serve web-scale traffic with billions of training examples, DCN showed limited expressiveness in its cross network at learning more predictive feature interactions. Despite significant research progress made, many deep learning models in production still rely on traditional feed-forward neural networks to learn feature crosses inefficiently. In light of the pros/cons of DCN and existing feature interaction learning approaches, we propose an improved framework DCN-V2 to make DCN more practical in large-scale industrial settings. In a comprehensive experimental study with extensive hyper-parameter search and model tuning, we observed that DCN-V2 approaches outperform all the state-of-the-art algorithms on popular benchmark datasets. The improved DCN-V2 is more expressive yet remains cost efficient at feature interaction learning, especially when coupled with a mixture of low-rank architecture. DCN-V2 is simple, can be easily adopted as building blocks, and has delivered significant offline accuracy and online business metrics gains across many web-scale learning to rank systems at Google. Our code and tutorial are open-sourced as part of TensorFlow Recommenders (TFRS)1.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 5e2eecd2-f185-4570-91bb-dabce224e87fCited by top-tier papers112
- Revisiting Deep Learning Models for Tabular DataYury Gorishniy, Ivan Rubachev, Valentin Khrulkov, Artem BabenkoNeurIPS 2021 · 1,847 citations
- Actions Speak Louder than Words: Trillion-Parameter Sequential Transducers for Generative RecommendationsJiaqi Zhai, Lucy Liao, Xing Liu, Yueming Wang et al.ICML 2024 · 200 citations
- ReLLa: Retrieval-enhanced Large Language Models for Lifelong Sequential Behavior Comprehension in RecommendationJianghao Lin, Rong Shan, Chenxu Zhu, Kounianhua Du et al.WWW 2024 · 151 citations
- FinalMLP: An Enhanced Two-Stream MLP Model for CTR PredictionKelong Mao, Jieming Zhu, Liangcai Su, Guohao Cai et al.AAAI 2023 · 142 citations
- Jury Learning: Integrating Dissenting Voices into Machine Learning ModelsMitchell L. Gordon, Michelle S. Lam, Joon Sung Park, Kayur Patel et al.CHI 2022 · 134 citations
Builds on1
Related papers
- FCN: Fusing Exponential and Linear Cross Network for Click-Through Rate PredictionHonghao Li, Yiwen Zhang, Yi Zhang, Hanwei Li et al.KDD 2026 · 7 citations
- Progressive Feature Interaction Search for Deep Sparse NetworkChen Gao, Yinfeng Li, Quanming Yao, Depeng Jin et al.NeurIPS 2021 · 17 citations
- Detecting Arbitrary Order Beneficial Feature Interactions for Recommender SystemsYixin Su, Yunxiang Zhao, Sarah M. Erfani, Junhao Gan et al.KDD 2022 · 25 citations
- AutoGroup: Automatic Feature Grouping for Modelling Explicit High-Order Feature Interactions in CTR PredictionBin Liu, Niannan Xue, Huifeng Guo, Ruiming Tang et al.SIGIR 2020 · 48 citations
- Towards Hybrid-grained Feature Interaction Selection for Deep Sparse NetworkFuyuan Lyu, Xing Tang, Dugang Liu, Chen Ma et al.NeurIPS 2023 · 4 citations
