NASRec: Weight Sharing Neural Architecture Search for Recommender Systems
Tunhou Zhang, Dehua Cheng, Yuchen He, Zhengxing Chen, Xiaoliang Dai, Liang Xiong, Feng Yan, Hai Li, Yiran Chen, Wei Wen
Abstract
The rise of deep neural networks offers new opportunities in optimizing recommender systems. However, optimizing recommender systems using deep neural networks requires delicate architecture fabrication. We propose NASRec, a paradigm that trains a single supernet and efficiently produces abundant models/sub-architectures by weight sharing. To overcome the data multi-modality and architecture heterogeneity challenges in the recommendation domain, NASRec establishes a large supernet (i.e., search space) to search the full architectures. The supernet incorporates versatile choice of operators and dense connectivity to minimize human efforts for finding priors. The scale and heterogeneity in NASRec impose several challenges, such as training inefficiency, operator-imbalance, and degraded rank correlation. We tackle these challenges by proposing single-operator any-connection sampling, operator-balancing interaction modules, and post-training fine-tuning. Our crafted models, NASRecNet, show promising results on three Click-Through Rates (CTR) prediction benchmarks, indicating that NASRec outperforms both manually designed models and existing NAS methods with state-of-the-art performance. Our work is publicly available here.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext fc99f295-fb7c-41ac-a17d-d85d4cfb2730Cited by top-tier papers2
- Forward Once for All: Structural Parameterized Adaptation for Efficient Cloud-coordinated On-device RecommendationKairui Fu, Zheqi Lv, Shengyu Zhang, Fan Wu et al.KDD 2025 · 3 citations
- Drs.NAS: Ultra-Efficient Neural Architecture Search for Recommendation SystemsRuixuan Wang, Xun JiaoOSDI 2026
Builds on8
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Once-for-All: Train One Network and Specialize it for Efficient DeploymentHan Cai, Chuang Gan, Tianzhe Wang, Zhekai Zhang et al.ICLR 2020 · 1,522 citations
- DCN V2: Improved Deep & Cross Network and Practical Lessons for Web-scale Learning to Rank SystemsRuoxi Wang, Rakesh Shivanna, Derek Zhiyuan Cheng, Sagar Jain et al.WWW 2021 · 793 citations
- HAT: Hardware-Aware Transformers for Efficient Natural Language ProcessingHanrui Wang, Zhanghao Wu, Zhijian Liu, Han Cai et al.ACL 2020 · 215 citations
- Towards Automated Neural Interaction Discovery for Click-Through Rate PredictionQingquan Song, Dehua Cheng, Hanning Zhou, Jiyan Yang et al.KDD 2020 · 63 citations
Related papers
- Distribution Consistent Neural Architecture SearchJunyi Pan, Chong Sun, Yizhou Zhou, Ying Zhang et al.CVPR 2022 · 9 citations
- A General Method For Automatic Discovery of Powerful Interactions In Click-Through Rate PredictionZe Meng, Jinnian Zhang, Yumeng Li, Jiancheng Li et al.SIGIR 2021 · 17 citations
- Towards Automatic Discovering of Deep Hybrid Network Architecture for Sequential RecommendationMingyue Cheng, Zhiding Liu, Qi Liu, Shenyang Ge et al.WWW 2022 · 36 citations
- SUMNAS: Supernet with Unbiased Meta-Features for Neural Architecture SearchHyeonmin Ha, Ji-Hoon Kim, Semin Park, Byung-Gon ChunICLR 2022 · 5 citations
- SiGeo: Sub-One-Shot NAS via Geometry of Loss LandscapeHua Zheng, Kuang-Hung Liu, Igor Fedorov, Xin Zhang et al.KDD 2024 · 1 citation
