Variational Open-Domain Question Answering
Valentin Liévin, Andreas Geert Motzfeldt, Ida Riis Jensen, Ole Winther
摘要
Retrieval-augmented models have proven to be effective in natural language processing tasks, yet there remains a lack of research on their optimization using variational inference. We introduce the Variational Open-Domain (VOD) framework for end-to-end training and evaluation of retrieval-augmented models, focusing on open-domain question answering and language modelling. The VOD objective, a self-normalized estimate of the Rényi variational bound, approximates the task marginal likelihood and is evaluated under samples drawn from an auxiliary sampling distribution (cached retriever and/or approximate posterior). It remains tractable, even for retriever distributions defined on large corpora. We demonstrate VOD's versatility by training reader-retriever BERT-sized models on multiple-choice medical exam questions. On the MedMCQA dataset, we outperform the domain-tuned Med-PaLM by +5.3% despite using 2.500 fewer parameters. Our retrieval-augmented BioLinkBERT model scored 62.9% on the MedMCQA and 55.0% on the MedQA-USMLE. Last, we show the effectiveness of our learned retriever component in the context of medical semantic search.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Knowledge-Augmented Reasoning Distillation for Small Language Models in Knowledge-Intensive TasksMinki Kang, Seanie Lee, Jinheon Baek, Kenji Kawaguchi 等NeurIPS 2023 · 被引用 128 次
- Fundamental Capabilities of Large Language Models and their Applications in Domain Scenarios: A SurveyJiawei Li, Yizhe Yang, Yu Bai, Xiaofeng Zhou 等ACL 2024 · 被引用 15 次
- Enhancing Small Medical Learners with Privacy-preserving Contextual PromptingXinlu Zhang, Shiyang Li, Xianjun Yang, Chenxin Tian 等ICLR 2024 · 被引用 13 次
- To Generate or to Retrieve? On the Effectiveness of Artificial Contexts for Medical Open-Domain Question AnsweringGiacomo Frisoni, Alessio Cocchieri, Alex Presepi, Gianluca Moro 等ACL 2024 · 被引用 8 次
- A Statistical Framework for Data-dependent Retrieval-Augmented ModelsSoumya Basu, Ankit Singh Rawat, Manzil ZaheerICML 2024 · 被引用 2 次
它引用的顶会 Paper5
- Retrieval Augmented Language Model Pre-TrainingKelvin Guu, Kenton Lee, Zora Tung, Panupong Pasupat 等ICML 2020 · 被引用 2,937 次
- Improving Language Models by Retrieving from Trillions of TokensSebastian Borgeaud, Arthur Mensch, Jordan Hoffmann, Trevor Cai 等ICML 2022 · 被引用 1,629 次
- End-to-End Training of Multi-Document Reader and Retriever for Open-Domain Question AnsweringDevendra Singh Sachan, Siva Reddy, William L. Hamilton, Chris Dyer 等NeurIPS 2021 · 被引用 197 次
- Hindsight: Posterior-guided training of retrievers for improved open-ended generationAshwin Paranjape, Omar Khattab, Christopher Potts, Matei Zaharia 等ICLR 2022 · 被引用 48 次
- Optimal Variance Control of the Score-Function Gradient Estimator for Importance-Weighted BoundsValentin Liévin, Andrea Dittadi, Anders Christensen, Ole WintherNeurIPS 2020 · 被引用 9 次
相关 Paper
- End-to-End Training of Neural Retrievers for Open-Domain Question AnsweringDevendra Singh Sachan, Mostofa Patwary, Mohammad Shoeybi, Neel Kant 等ACL 2021
- RINK: Reader-Inherited Evidence Reranker for Table-and-Text Open Domain Question AnsweringEunhwan Park, Sung-Min Lee, Daeryong Seo, Seonhoon Kim 等AAAI 2023 · 被引用 4 次
- Retrieval is Accurate GenerationBowen Cao, Deng Cai, Leyang Cui, Xuxin Cheng 等ICLR 2024 · 被引用 13 次
- Making Retrieval-Augmented Language Models Robust to Irrelevant ContextOri Yoran, Tomer Wolfson, Ori Ram, Jonathan BerantICLR 2024 · 被引用 361 次
- Improving Biomedical Information Retrieval with Neural RetrieversMan Luo, Arindam Mitra, Tejas Gokhale, Chitta BaralAAAI 2022 · 被引用 42 次
