Variational Open-Domain Question Answering
Valentin Liévin, Andreas Geert Motzfeldt, Ida Riis Jensen, Ole Winther
Abstract
Retrieval-augmented models have proven to be effective in natural language processing tasks, yet there remains a lack of research on their optimization using variational inference. We introduce the Variational Open-Domain (VOD) framework for end-to-end training and evaluation of retrieval-augmented models, focusing on open-domain question answering and language modelling. The VOD objective, a self-normalized estimate of the Rényi variational bound, approximates the task marginal likelihood and is evaluated under samples drawn from an auxiliary sampling distribution (cached retriever and/or approximate posterior). It remains tractable, even for retriever distributions defined on large corpora. We demonstrate VOD's versatility by training reader-retriever BERT-sized models on multiple-choice medical exam questions. On the MedMCQA dataset, we outperform the domain-tuned Med-PaLM by +5.3% despite using 2.500 fewer parameters. Our retrieval-augmented BioLinkBERT model scored 62.9% on the MedMCQA and 55.0% on the MedQA-USMLE. Last, we show the effectiveness of our learned retriever component in the context of medical semantic search.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers6
- Knowledge-Augmented Reasoning Distillation for Small Language Models in Knowledge-Intensive TasksMinki Kang, Seanie Lee, Jinheon Baek, Kenji Kawaguchi et al.NeurIPS 2023 · 128 citations
- Fundamental Capabilities of Large Language Models and their Applications in Domain Scenarios: A SurveyJiawei Li, Yizhe Yang, Yu Bai, Xiaofeng Zhou et al.ACL 2024 · 15 citations
- Enhancing Small Medical Learners with Privacy-preserving Contextual PromptingXinlu Zhang, Shiyang Li, Xianjun Yang, Chenxin Tian et al.ICLR 2024 · 13 citations
- To Generate or to Retrieve? On the Effectiveness of Artificial Contexts for Medical Open-Domain Question AnsweringGiacomo Frisoni, Alessio Cocchieri, Alex Presepi, Gianluca Moro et al.ACL 2024 · 8 citations
- A Statistical Framework for Data-dependent Retrieval-Augmented ModelsSoumya Basu, Ankit Singh Rawat, Manzil ZaheerICML 2024 · 2 citations
Builds on5
- Retrieval Augmented Language Model Pre-TrainingKelvin Guu, Kenton Lee, Zora Tung, Panupong Pasupat et al.ICML 2020 · 2,937 citations
- Improving Language Models by Retrieving from Trillions of TokensSebastian Borgeaud, Arthur Mensch, Jordan Hoffmann, Trevor Cai et al.ICML 2022 · 1,629 citations
- End-to-End Training of Multi-Document Reader and Retriever for Open-Domain Question AnsweringDevendra Singh Sachan, Siva Reddy, William L. Hamilton, Chris Dyer et al.NeurIPS 2021 · 197 citations
- Hindsight: Posterior-guided training of retrievers for improved open-ended generationAshwin Paranjape, Omar Khattab, Christopher Potts, Matei Zaharia et al.ICLR 2022 · 48 citations
- Optimal Variance Control of the Score-Function Gradient Estimator for Importance-Weighted BoundsValentin Liévin, Andrea Dittadi, Anders Christensen, Ole WintherNeurIPS 2020 · 9 citations
Related papers
- End-to-End Training of Neural Retrievers for Open-Domain Question AnsweringDevendra Singh Sachan, Mostofa Patwary, Mohammad Shoeybi, Neel Kant et al.ACL 2021
- RINK: Reader-Inherited Evidence Reranker for Table-and-Text Open Domain Question AnsweringEunhwan Park, Sung-Min Lee, Daeryong Seo, Seonhoon Kim et al.AAAI 2023 · 4 citations
- Retrieval is Accurate GenerationBowen Cao, Deng Cai, Leyang Cui, Xuxin Cheng et al.ICLR 2024 · 13 citations
- Making Retrieval-Augmented Language Models Robust to Irrelevant ContextOri Yoran, Tomer Wolfson, Ori Ram, Jonathan BerantICLR 2024 · 361 citations
- Improving Biomedical Information Retrieval with Neural RetrieversMan Luo, Arindam Mitra, Tejas Gokhale, Chitta BaralAAAI 2022 · 42 citations
