Test-time Collective Prediction
Celestine Mendler-Dünner, Wenshuo Guo, Stephen Bates, Michael I. Jordan
Abstract
An increasingly common setting in machine learning involves multiple parties, each with their own data, who want to jointly make predictions on future test points. Agents wish to benefit from the collective expertise of the full set of agents to make better predictions than they would individually, but may not be willing to release their data or model parameters. In this work, we explore a decentralized mechanism to make collective predictions at test time, leveraging each agent's pre-trained model without relying on external validation, model retraining, or data pooling. Our approach takes inspiration from the literature in social science on human consensus-making. We analyze our mechanism theoretically, showing that it converges to inverse meansquared-error (MSE) weighting in the large-sample limit. To compute error bars on the collective predictions we propose a decentralized Jackknife procedure that evaluates the sensitivity of our mechanism to a single agent's prediction. Empirically, we demonstrate that our scheme effectively combines models with differing quality across the input space. The proposed consensus prediction achieves significant gains over classical model averaging, and even outperforms weighted averaging schemes that have access to additional validation data.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8dc65a3f-b614-4eed-9bfc-7b7095a18aefCited by top-tier papers2
- Collaborative Learning via Prediction ConsensusDongyang Fan, Celestine Mendler-Dünner, Martin JaggiNeurIPS 2023 · 11 citations
- A Kernel Perspective on Distillation-based Collaborative LearningSejun Park, Kihun Hong, Ganguk HwangNeurIPS 2024 · 3 citations
Builds on1
Related papers
- Decentralized Langevin Dynamics for Bayesian LearningAnjaly Parayil, He Bai, Jemin George, Prudhvi GurramNeurIPS 2020 · 10 citations
- Meta-Learning PAC-Bayes Priors in Model AveragingYimin Huang, Weiran Huang, Liang Li, Zhenguo LiAAAI 2020 · 8 citations
- Synthetic Model Combination: An Instance-wise Approach to Unsupervised Ensemble LearningAlex J. Chan, Mihaela van der SchaarNeurIPS 2022 · 6 citations
- Consensus Control for Decentralized Deep LearningLingjing Kong, Tao Lin, Anastasia Koloskova, Martin Jaggi et al.ICML 2021 · 100 citations
- RelaySum for Decentralized Deep Learning on Heterogeneous DataThijs Vogels, Lie He, Anastasia Koloskova, Sai Praneeth Karimireddy et al.NeurIPS 2021 · 78 citations
