Principled Model Routing for Unknown Mixtures of Source Domains
Christoph Dann, Yishay Mansour, Teodor Vanislavov Marinov, Mehryar Mohri
摘要
The rapid proliferation of domain-specialized machine learning models presents a challenge: while individual models excel in specific domains, their performance varies significantly across diverse applications. This makes selecting the optimal model when faced with an unknown mixture of tasks, especially with limited or no data to estimate the mixture, a difficult problem. We address this challenge by formulating it as a multiple-source domain adaptation (MSA) problem. We introduce a novel, scalable algorithm that effectively routes each input to the best-suited model from a pool of available models. Our approach provides a strong performance guarantee: remarkably, for any mixture domain, the accuracy achieved by the best source model is maintained. This guarantee is established through a theoretical bound on the regret for new domains, expressed as a convex combination of the best regrets in the source domains, plus a concentration term that diminishes as the amount of source data increases. While our primary contributions are theoretical and algorithmic, we also present empirical results demonstrating the effectiveness of our approach.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper14
- Model soups: averaging weights of multiple fine-tuned models improves accuracy without increasing inference timeMitchell Wortsman, Gabriel Ilharco, Samir Yitzhak Gadre, Rebecca Roelofs 等ICML 2022 · 被引用 1,464 次
- Adversarially Trained Actor Critic for Offline Reinforcement LearningChing-An Cheng, Tengyang Xie, Nan Jiang, Alekh AgarwalICML 2022 · 被引用 156 次
- Large Language Model Cascades with Mixture of Thought Representations for Cost-Efficient ReasoningMurong Yue, Jie Zhao, Min Zhang, Liang Du 等ICLR 2024 · 被引用 153 次
- AutoMix: Automatically Mixing Language ModelsPranjal Aggarwal, Aman Madaan, Ankit Anand, Srividya Pranavi Potharaju 等NeurIPS 2024 · 被引用 145 次
- Two-Stage Learning to Defer with Multiple ExpertsAnqi Mao, Christopher Mohri, Mehryar Mohri, Yutao ZhongNeurIPS 2023 · 被引用 98 次
相关 Paper
- A Discriminative Technique for Multiple-Source AdaptationCorinna Cortes, Mehryar Mohri, Ananda Theertha Suresh, Ningshan ZhangICML 2021 · 被引用 15 次
- Aggregating From Multiple Target-Shifted SourcesChangjian Shui, Zijian Li, Jiaqi Li, Christian Gagné 等ICML 2021 · 被引用 36 次
- Unsupervised Multi-Source Domain Adaptation Without Access to Source DataSk Miraj Ahmed, Dripta S. Raychaudhuri, Sujoy Paul, Samet Oymak 等CVPR 2021
- Mixture Weight Estimation and Model Prediction in Multi-source Multi-target Domain AdaptationYuyang Deng, Ilja Kuzborskij, Mehrdad MahdaviNeurIPS 2023 · 被引用 3 次
- Domain Aggregation Networks for Multi-Source Domain AdaptationJunfeng Wen, Russell Greiner, Dale SchuurmansICML 2020 · 被引用 82 次
