Attention Weights as an Indicator: Analyzing and Improving Document Utilization in Retrieval-Augmented Generation
Jing Jin, Yuhan Song, Wen Luo, Houfeng Wang
Abstract
The generation of Retrieval-Augmented Generation (RAG) models is affected by factors such as the quality and order of input documents, indicating that their ability to utilize documents remains underdeveloped. This ability encompasses not only identifying useful documents from inputs but also minimizing positional bias and filtering irrelevant documents. To achieve this, key challenges include the model's internal estimation of document importance and positional bias. In this paper, we conduct a comprehensive study on the properties of attention weights, examining the impact of factors like aggregation methods, document quality, document position, token type, and so on. Based on our findings, we propose strategies to enhance document utilization from three perspectives: document ranking, placement, and filtering. Comprehensive experiments show that our method outperforms baselines and improves document utilization effectiveness in a trainingfree manner. 1
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d75e5ce7-3905-4eac-8d40-234e8661b563Builds on13
- Self-RAG: Learning to Retrieve, Generate, and Critique through Self-ReflectionAkari Asai, Zeqiu Wu, Yizhong Wang, Avirup Sil et al.ICLR 2024 · 1,798 citations
- Large Language Models Can Be Easily Distracted by Irrelevant ContextFreda Shi, Xinyun Chen, Kanishka Misra, Nathan Scales et al.ICML 2023 · 970 citations
- Making Retrieval-Augmented Language Models Robust to Irrelevant ContextOri Yoran, Tomer Wolfson, Ori Ram, Jonathan BerantICLR 2024 · 361 citations
- The Power of Noise: Redefining Retrieval for RAG SystemsFlorin Cuconasu, Giovanni Trappolini, Federico Siciliano, Simone Filice et al.SIGIR 2024 · 212 citations
- Is ChatGPT Good at Search? Investigating Large Language Models as Re-Ranking AgentsWeiwei Sun, Lingyong Yan, Xinyu Ma, Shuaiqiang Wang et al.EMNLP 2023 · 182 citations
Related papers
- MAIN-RAG: Multi-Agent Filtering Retrieval-Augmented GenerationChia-Yuan Chang, Zhimeng Jiang, Vineeth Rakesh, Menghai Pan et al.ACL 2025
- EAReranker: Efficient Embedding Adequacy Assessment for Retrieval Augmented GenerationDongyang Zeng, Yaping Liu, Wei Zhang, Shuo Zhang et al.NeurIPS 2025
- GRO-RAG: Gradient-aware Re-rank Optimization for Multi-source Retrieval-Augmented GenerationSiyuan Chen, Hang Ding, Xiaoyu Kang, Jiechao GaoICLR 2026
- InfoGain-RAG: Boosting Retrieval-Augmented Generation through Document Information Gain-based Reranking and FilteringZihan Wang, Zihan Liang, Zhou Shao, Yufei Ma et al.EMNLP 2025 · 4 citations
- Long-Context LLMs Meet RAG: Overcoming Challenges for Long Inputs in RAGBowen Jin, Jinsung Yoon, Jiawei Han, Sercan Ö. ArikICLR 2025
