Harnessing Multi-Role Capabilities of Large Language Models for Open-Domain Question Answering
Hongda Sun, Yuxuan Liu, Chengwei Wu, Haiyu Yan, Cheng Tai, Xin Gao, Shuo Shang, Rui Yan
摘要
Open-domain question answering (ODQA) has emerged as a pivotal research spotlight in information systems. Existing methods follow two main paradigms to collect evidence: (1) Theretrieve-then-read paradigm retrieves pertinent documents from an external corpus; and (2) thegenerate-then-read paradigm employs large language models (LLMs) to generate relevant documents. However, neither can fully address multifaceted requirements for evidence. To this end, we propose LLMQA, a generalized framework that formulates the ODQA process into three basic steps: query expansion, document selection, and answer generation, combining the superiority of both retrieval-based and generation-based evidence. Since LLMs exhibit their excellent capabilities to accomplish various tasks, we instruct LLMs to play multiple roles as generators, rerankers, and evaluators within our framework, integrating them to collaborate in the ODQA process. Furthermore, we introduce a novel prompt optimization algorithm to refine role-playing prompts and steer LLMs to produce higher-quality evidence and answers. Extensive experimental results on widely used benchmarks (NQ, WebQ, and TriviaQA) demonstrate that LLMQA achieves the best performance in terms of both answer accuracy and evidence quality, showcasing its potential for advancing ODQA research and applications.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- SheetAgent: Towards a Generalist Agent for Spreadsheet Reasoning and Manipulation via Large Language ModelsYibin Chen, Yifu Yuan, Zeyu Zhang, Yan Zheng 等WWW 2025 · 被引用 13 次
- BiDeV: Bilateral Defusing Verification for Complex Claim Fact-CheckingYuxuan Liu, Hongda Sun, Wenya Guo, Xinyan Xiao 等AAAI 2025 · 被引用 11 次
- Mobile-Bench: An Evaluation Benchmark for LLM-based Mobile AgentsShihan Deng, Weikai Xu, Hongda Sun, Wei Liu 等ACL 2024 · 被引用 10 次
- MobileSteward: Integrating Multiple App-Oriented Agents with Self-Evolution to Automate Cross-App InstructionsYuxuan Liu, Hongda Sun, Wei Liu, Jian Luan 等KDD 2025 · 被引用 4 次
- MockLLM: A Multi-Agent Behavior Collaboration Framework for Online Job Seeking and RecruitingHongda Sun, Hongzhan Lin, Haiyu Yan, Yang Song 等KDD 2025 · 被引用 1 次
它引用的顶会 Paper16
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida 等NeurIPS 2022 · 被引用 24,707 次
- Retrieval-Augmented Generation for Knowledge-Intensive NLP TasksPatrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni 等NeurIPS 2020 · 被引用 19,162 次
- Reflexion: language agents with verbal reinforcement learningNoah Shinn, Federico Cassano, Ashwin Gopinath, Karthik Narasimhan 等NeurIPS 2023 · 被引用 5,828 次
- Self-Refine: Iterative Refinement with Self-FeedbackAman Madaan, Niket Tandon, Prakhar Gupta, Skyler Hallinan 等NeurIPS 2023 · 被引用 4,972 次
相关 Paper
- Generate rather than Retrieve: Large Language Models are Strong Context GeneratorsWenhao Yu, Dan Iter, Shuohang Wang, Yichong Xu 等ICLR 2023 · 被引用 86 次
- SuRe: Summarizing Retrievals using Answer Candidates for Open-domain QA of LLMsJaehyung Kim, Jaehyun Nam, Sangwoo Mo, Jongjin Park 等ICLR 2024 · 被引用 89 次
- Beyond Prompting: An Efficient Embedding Framework for Open-Domain Question AnsweringZhanghao Hu, Hanqi Yan, Qinglin Zhu, Zhenyi Shen 等ACL 2025
- UnitedQA: A Hybrid Approach for Open Domain Question AnsweringHao Cheng, Yelong Shen, Xiaodong Liu, Pengcheng He 等ACL 2021
- Merging Generated and Retrieved Knowledge for Open-Domain QAYunxiang Zhang, Muhammad Khalifa, Lajanugen Logeswaran, Moontae Lee 等EMNLP 2023 · 被引用 12 次
