ATM: Adversarial Tuning Multi-agent System Makes a Robust Retrieval-Augmented Generator
Junda Zhu, Lingyong Yan, Haibo Shi, Dawei Yin, Lei Sha
Abstract
Large language models (LLMs) are proven to benefit a lot from retrieval-augmented generation (RAG) in alleviating hallucinations confronted with knowledge-intensive questions. RAG adopts information retrieval techniques to inject external knowledge from semanticrelevant documents as input contexts. However, since today's Internet is flooded with numerous noisy and fabricating content, it is inevitable that RAG systems are vulnerable to these noises and prone to respond incorrectly. To this end, we propose to optimize the retrieval-augmented GENERATOR with an Adversarial Tuning Multi-agent system (ATM). The ATM steers the GENERATOR to have a robust perspective of useful documents for question answering with the help of an auxiliary ATTACKER agent through adversarially tuning the agents for several iterations. After rounds of multi-agent iterative tuning, the GENERA-TOR can eventually better discriminate useful documents amongst fabrications. The experimental results verify the effectiveness of ATM and we also observe that the GENERATOR can achieve better performance compared to the state-of-the-art baselines. The code is available at https://github.com/chuhac/ATM-RAG .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 90f2b32f-f4ef-466c-bcf6-f94609b510f5Cited by top-tier papers2
- Stable-RAG: Mitigating Retrieval-Permutation-Induced Hallucinations in Retrieval-Augmented GenerationQianchi Zhang, Hainan Zhang, Liang Pang, Hongwei Zheng et al.ACL 2026 · 3 citations
- Less is More: Compact Clue Selection for Efficient Retrieval-Augmented Generation ReasoningQianchi Zhang, Hainan Zhang, Liang Pang, Yongxin Tong et al.WWW 2026
Builds on22
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma et al.NeurIPS 2022 · 22,562 citations
- Retrieval-Augmented Generation for Knowledge-Intensive NLP TasksPatrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni et al.NeurIPS 2020 · 19,162 citations
- Direct Preference Optimization: Your Language Model is Secretly a Reward ModelRafael Rafailov, Archit Sharma, Eric Mitchell, Christopher D. Manning et al.NeurIPS 2023 · 10,924 citations
- SimCSE: Simple Contrastive Learning of Sentence EmbeddingsTianyu Gao, Xingcheng Yao, Danqi ChenEMNLP 2021 · 2,496 citations
Related papers
- Improving Retrieval-Augmented Generation through Multi-Agent Reinforcement LearningYiqun Chen, Lingyong Yan, Weiwei Sun, Xinyu Ma et al.NeurIPS 2025 · 47 citations
- Enhancing Noise Robustness of Retrieval-Augmented Language Models with Adaptive Adversarial TrainingFeiteng Fang, Yuelin Bai, Shiwen Ni, Min Yang et al.ACL 2024 · 18 citations
- MAIN-RAG: Multi-Agent Filtering Retrieval-Augmented GenerationChia-Yuan Chang, Zhimeng Jiang, Vineeth Rakesh, Menghai Pan et al.ACL 2025
- Separate the Wheat from the Chaff: Winnowing Down Divergent Views in Retrieval Augmented GenerationSong Wang, Zihan Chen, Peng Wang, Zhepei Wei et al.EMNLP 2025 · 1 citation
- Open Schrödinger's Closed Box: Identifying Retrieval Augmented Generation in API-Accessible Large Language Model ServicesYukun Jiang, Xinyue Shen, Michael Backes, Zheng Li et al.ACL 2026
