A Multi-Task Embedder For Retrieval Augmented LLMs
Peitian Zhang, Zheng Liu, Shitao Xiao, Zhicheng Dou, Jian-Yun Nie
摘要
LLMs confront inherent limitations in terms of its knowledge, memory, and action. The retrieval augmentation stands as a vital mechanism to address these limitations, which brings in useful information from external sources to augment the LLM. However, existing retrieval methods encounter two pressing issues. On one hand, the general retrievers are not properly optimized for retrieval augmentation hence exhibit limited effectiveness; on the other hand, the task-specific retrievers excel in the targeted retrieval augmentation scenario, while lack the versatility to handle diverse scenarios. In this work, we propose LLM-Embedder for the unified support of diverse retrieval augmentation scenarios. Our method presents three technical contributions. Firstly, we introduce a new reward formulation, namely rank-aware reward. It exploits the ranking position of the desired output among N sampled outputs from the LLM, which leads to fine-grained and robust computation of reward from the LLM's feedback. Secondly, we design a novel distillation objective, called graded distillation. It incorporates both the absolute value and the relative order of the reward for more sufficient utilization of the LLM's feedback. Thirdly, we systematically optimize the multi-task learning, which effectively unifies the multiple retrieval functionalities into one model. In our experiment, LLM-Embedder notably improves the LLM's performances in various downstream tasks, and outperforms both general and taskspecific retrievers with a substantial advantage. et al., 2022). Many of the challenges can be traced 045 back to the inherent limitations of LLMs in terms 046 of knowledge, memory, and action. Specifically, 047 LLMs cannot internalize the vast and constantly 048 changed world knowledge due to their finite and 049 static parameters. LLMs are incapable of memo-050 rizing and utilizing long-term information because 051 of the limited context length. Finally, LLMs re-052 quire manually in-context examples and tools to 053 accomplish complex real-world tasks. 054 Retrieval augmentation stands as a vital mech-055 anism to address these inherent limitations of the 056 LLM. It brings in useful information from exter-057 nal sources, such as knowledge, memory pieces, 058 in-context examples, and tools, which substantially 059 enhances the LLM for the generation of desired 060 outputs (Gao et al., 2023). The embedding model 061 (a.k.a. embedder) is a critical part of retrieval aug-062 mentation, which bridges the LLM's information 063 needs with external sources. The existing embed-064 ding models can be briefly partitioned into two 065 categories. One is the general-purpose embedders, 066 which aim to be universally applicable for various 067
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- MemPoison: Bypassing Selective Memory Mechanisms to Plant Backdoors in LLM AgentsHongtao Wang, Se Yang, Yu Chen, Puzhuo LiuCCS 2026 · 被引用 5 次
- Beyond RAG vs. Long-Context: Learning Distraction-Aware Retrieval for Efficient Knowledge GroundingSeong-Woong Shim, Myunsoo Kim, Jae Hyeon Cho, Byung-Jun LeeICLR 2026 · 被引用 1 次
- NeocorRAG: Less Irrelevant Information, More Explicit Evidence, and More Effective Recall via Evidence ChainsShiyao Peng, Qianhe Zheng, Zhuodi Hao, Zichen Tang 等WWW 2026
- LLM-EDT: Large Language Models Enhanced Cross-domain Sequential Recommendation with Dual-phase TrainingZiwei Liu, Qidong Liu, Wanyu Wang, Yejing Wang 等SIGIR 2026
它引用的顶会 Paper10
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida 等NeurIPS 2022 · 被引用 24,707 次
- WinoGrande: An Adversarial Winograd Schema Challenge at ScaleKeisuke Sakaguchi, Ronan Le Bras, Chandra Bhagavatula, Yejin ChoiAAAI 2020 · 被引用 3,037 次
- PIQA: Reasoning about Physical Commonsense in Natural LanguageYonatan Bisk, Rowan Zellers, Ronan Le Bras, Jianfeng Gao 等AAAI 2020 · 被引用 2,916 次
- Approximate Nearest Neighbor Negative Contrastive Learning for Dense Text RetrievalLee Xiong, Chenyan Xiong, Ye Li, Kwok-Fung Tang 等ICLR 2021 · 被引用 1,547 次
- Compressive Transformers for Long-Range Sequence ModellingJack W. Rae, Anna Potapenko, Siddhant M. Jayakumar, Chloe Hillier 等ICLR 2020 · 被引用 833 次
相关 Paper
- Landmark Embedding: A Chunking-Free Embedding Method For Retrieval Augmented Long-Context Large Language ModelsKun Luo, Zheng Liu, Shitao Xiao, Tong Zhou 等ACL 2024 · 被引用 10 次
- Augmentation-Adapted Retriever Improves Generalization of Language Models as Generic Plug-InZichun Yu, Chenyan Xiong, Shi Yu, Zhiyuan LiuACL 2023 · 被引用 15 次
- Bridging the Preference Gap between Retrievers and LLMsZixuan Ke, Weize Kong, Cheng Li, Mingyang Zhang 等ACL 2024 · 被引用 8 次
- UniLR: Unleashing the Power of LLMs on Multiple Legal Tasks with a Unified Legal RetrieverAng Li, Yiquan Wu, Yifei Liu, Ming Cai 等ACL 2025
- UniRAG: Unified Query Understanding Method for Retrieval Augmented GenerationRui Li, Liyang He, Qi Liu, Zheng Zhang 等ACL 2025
