RouterRetriever: Routing over a Mixture of Expert Embedding Models
Hyunji Lee, Luca Soldaini, Arman Cohan, Minjoon Seo, Kyle Lo
Abstract
Information retrieval methods often rely on a single embedding model trained on large, general-domain datasets like MSMARCO. While this approach can produce a retriever with reasonable overall performance, they often underperform models trained on domain-specific data when testing on their respective domains. Prior work in information retrieval has tackled this through multi-task training, but the idea of routing over a mixture of domain-specific expert retrievers remains unexplored despite the popularity of such ideas in language model generation research. In this work, we introduce ROUTERRETRIEVER, a retrieval model that leverages a mixture of domain-specific experts by using a routing mechanism to select the most appropriate expert for each query. ROUTERRETRIEVER is lightweight and allows easy addition or removal of experts without additional training. Evaluation on the BEIR benchmark demonstrates that ROUTERRE-TRIEVER outperforms both models trained on MSMARCO (+2.1 absolute nDCG@10) and multi-task models (+3.2). This is achieved by employing our routing mechanism, which surpasses other routing techniques (+1.8 on average) commonly used in language modeling. Furthermore, the benefit generalizes well to other datasets, even in the absence of a specific expert on the dataset. ROUTERRETRIEVER is the first work to demonstrate the advantages of routing over a mixture of domain-specific expert embedding models as an alternative to a single, general-purpose embedding model, especially when retrieving from diverse, specialized domains. Code github/amy-hyunji/RouterRetriever Weights hf.co/amy-hyunji/RouterRetriever
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext ca9a758c-dbbd-4f5c-91ea-1a7036c31bc3Cited by top-tier papers2
- FlexOLMo: Open Language Models for Flexible Data UseWeijia Shi, Akshita Bhagia, Kevin Farhat, Niklas Muennighoff et al.NeurIPS 2025 · 16 citations
- R⌃3AG: Retriever Routing for Retrieval-Augmented GenerationTong Zhao, Yutao Zhu, Yucheng Tian, Zhicheng DouACL 2026 · 1 citation
Builds on12
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- SimCSE: Simple Contrastive Learning of Sentence EmbeddingsTianyu Gao, Xingcheng Yao, Danqi ChenEMNLP 2021 · 2,496 citations
- When Not to Trust Language Models: Investigating Effectiveness of Parametric and Non-Parametric MemoriesAlex Mallen, Akari Asai, Victor Zhong, Rajarshi Das et al.ACL 2023 · 233 citations
- Exploring the Benefits of Training Expert Language Models over Instruction TuningJoel Jang, Seungone Kim, Seonghyeon Ye, Doyoung Kim et al.ICML 2023 · 97 citations
- Learning to Route Among Specialized Experts for Zero-Shot GeneralizationMohammed Muqeeth, Haokun Liu, Yufan Liu, Colin RaffelICML 2024 · 63 citations
Related papers
- Retrv-MoE: Scaling Unified Multimodal Retrieval with Sparse Mixture-of-ExpertsTongxu Lin, Jiayin XiaoKDD 2026
- Glider: Global and Local Instruction-Driven Expert RouterPingzhi Li, Prateek Yadav, Jaehong Yoon, Jie Peng et al.EMNLP 2025 · 1 citation
- Your Mixture-of-Experts LLM Is Secretly an Embedding Model for FreeZiyue Li, Tianyi ZhouICLR 2025
- R2-T2: Re-Routing in Test-Time for Multimodal Mixture-of-ExpertsZhongyang Li, Ziyue Li, Tianyi ZhouICML 2025
- Search-Adaptor: Embedding Customization for Information RetrievalJinsung Yoon, Yanfei Chen, Sercan Ö. Arik, Tomas PfisterACL 2024
