On the Generalizability and Predictability of Recommender Systems
Duncan C. McElfresh, Sujay Khandagale, Jonathan Valverde, John Dickerson, Colin White
Abstract
While other areas of machine learning have seen more and more automation, designing a high-performing recommender system still requires a high level of human effort. Furthermore, recent work has shown that modern recommender system algorithms do not always improve over well-tuned baselines. A natural follow-up question is, "how do we choose the right algorithm for a new dataset and performance metric?" In this work, we start by giving the first large-scale study of recommender system approaches by comparing 24 algorithms and 100 sets of hyperparameters across 85 datasets and 315 metrics. We find that the best algorithms and hyperparameters are highly dependent on the dataset and performance metric. However, there is also a strong correlation between the performance of each algorithm and various meta-features of the datasets. Motivated by these findings, we create RecZilla, a meta-learning approach to recommender systems that uses a model to predict the best algorithm and hyperparameters for new, unseen datasets. By using far more meta-training data than prior work, RecZilla is able to substantially reduce the level of human involvement when faced with a new recommender system application. We not only release our code and pretrained RecZilla models, but also all of our raw experimental results, so that practitioners can train a RecZilla model for their desired performance metric: https://github.com/naszilla/reczilla .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers4
- STAIR: Manipulating Collaborative and Multimodal Information for E-Commerce RecommendationCong Xu, Yunhang He, Jun Wang, Wei ZhangAAAI 2025 · 8 citations
- Reason-to-Rank: Distilling Direct and Comparative Reasoning from Large Language Models for Document RerankingYuelyu Ji, Zhuochun Li, Rui Meng, Daqing HeSIGIR 2025 · 3 citations
- Multi-Location Software Model CompletionAlisa Welter, Christof Tinnes, Sven ApelICSE 2026 · 1 citation
- Aspect-Aware Content-Based Recommendations for Mathematical Research PapersAnkit Satpute, André Greiner-Petter, Noah Gießing, Olaf Teschke et al.SIGIR 2026
Builds on1
Related papers
- Guided Recommendation for Model Fine-TuningHao Li, Charless C. Fowlkes, Hao Yang, Onkar Dabeer et al.CVPR 2023
- MetaGL: Evaluation-Free Selection of Graph Learning Models via Meta-LearningNamyong Park, Ryan A. Rossi, Nesreen K. Ahmed, Christos FaloutsosICLR 2023 · 1 citation
- Quick-Tune: Quickly Learning Which Pretrained Model to Finetune and HowSebastian Pineda-Arango, Fabio Ferreira, Arlind Kadra, Frank Hutter et al.ICLR 2024 · 27 citations
- Efficient Data-specific Model Search for Collaborative FilteringChen Gao, Quanming Yao, Depeng Jin, Yong LiKDD 2021 · 13 citations
- Efficient and Joint Hyperparameter and Architecture Search for Collaborative FilteringYan Wen, Chen Gao, Lingling Yi, Liwei Qiu et al.KDD 2023 · 6 citations
