Will this Question be Answered? Question Filtering via Answer Model Distillation for Efficient Question Answering
Siddhant Garg, Alessandro Moschitti
Abstract
In this paper we propose a novel approach towards improving the efficiency of Question Answering (QA) systems by filtering out questions that will not be answered by them. This is based on an interesting new finding: the answer confidence scores of state-of-the-art QA systems can be approximated well by models solely using the input question text. This enables preemptive filtering of questions that are not answered by the system due to their answer confidence scores being lower than the system threshold. Specifically, we learn Transformer-based question models by distilling Transformer-based answering models. Our experiments on three popular QA datasets and one industrial QA benchmark demonstrate the ability of our question models to approximate the Precision/Recall curves of the target QA system well. These question models, when used as filters, can effectively trade off lower computation cost of QA systems for lower Recall, e.g., reducing computation by 60%, while only losing 3-4% of Recall. Re --2.4 --0.8 --1.6 --2.5 --3.7 --0.8 --2.0 --4.1 --5.0 --6.1 FM * 2 M 1 % Filter 4.1 17.3 20.9 21.0 44.0 1.2 1.8 2.6 3.9 8.8 Re --2.8 --10.3 --9.0 --8.3 --7.4 --0.8 --0.8 --0.4 --2.1 --2.9 ASNQ M 1 Pr 68.2 74.6 77.2 79.6 90.5 75.7 79.6 81.9 84.5 92.9 Re 48.7 41.1 36.1 28.9 10.0 61.1 54.4 49.3 42.7 20.7 FM 1 * 2 M 1 % Filter 7.8 17.8 29.8 54.2 83.8 0.2 10.5 15.6 29.8 66.2
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers5
- The Unreliability of Explanations in Few-shot Prompting for Textual ReasoningXi Ye, Greg DurrettNeurIPS 2022 · 272 citations
- Learning to Reject with a Fixed Predictor: Application to DecontextualizationChristopher Mohri, Daniel Andor, Eunsol Choi, Michael Collins et al.ICLR 2024 · 32 citations
- Knowledge Transfer from Answer Ranking to Answer GenerationMatteo Gabburo, Rik Koncel-Kedziorski, Siddhant Garg, Luca Soldaini et al.EMNLP 2022 · 4 citations
- Learning Answer Generation using Supervision from Automatic Question Answering EvaluatorsMatteo Gabburo, Siddhant Garg, Rik Koncel-Kedziorski, Alessandro MoschittiACL 2023 · 4 citations
- Fairness Beyond Performance: Revealing Reliability Disparities Across Groups in Legal NLPT. Y. S. S. Santosh, Irtiza ChowdhuryACL 2025
Builds on9
- ALBERT: A Lite BERT for Self-supervised Learning of Language RepresentationsZhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel et al.ICLR 2020 · 7,418 citations
- MobileBERT: a Compact Task-Agnostic BERT for Resource-Limited DevicesZhiqing Sun, Hongkun Yu, Xiaodan Song, Renjie Liu et al.ACL 2020 · 660 citations
- ELECTRA: Pre-training Text Encoders as Discriminators Rather Than GeneratorsKevin Clark, Minh-Thang Luong, Quoc V. Le, Christopher D. ManningICLR 2020 · 541 citations
- Learning to Retrieve Reasoning Paths over Wikipedia Graph for Question AnsweringAkari Asai, Kazuma Hashimoto, Hannaneh Hajishirzi, Richard Socher et al.ICLR 2020 · 322 citations
- FastBERT: a Self-distilling BERT with Adaptive Inference TimeWeijie Liu, Peng Zhou, Zhiruo Wang, Zhe Zhao et al.ACL 2020 · 257 citations
Related papers
- DeFormer: Decomposing Pre-trained Transformers for Faster Question AnsweringQingqing Cao, Harsh Trivedi, Aruna Balasubramanian, Niranjan BalasubramanianACL 2020 · 61 citations
- Block-Skim: Efficient Question Answering for TransformerYue Guan, Zhengyi Li, Zhouhan Lin, Yuhao Zhu et al.AAAI 2022 · 33 citations
- PELA: Learning Parameter-Efficient Models with Low-Rank ApproximationYangyang Guo, Guangzhi Wang, Mohan S. KankanhalliCVPR 2024
- The Cascade Transformer: an Application for Efficient Answer Sentence SelectionLuca Soldaini, Alessandro MoschittiACL 2020 · 3 citations
- Dynamic Context Pruning for Efficient and Interpretable Autoregressive TransformersSotiris Anagnostidis, Dario Pavllo, Luca Biggio, Lorenzo Noci et al.NeurIPS 2023 · 95 citations
