Will this Question be Answered? Question Filtering via Answer Model Distillation for Efficient Question Answering
Siddhant Garg, Alessandro Moschitti
摘要
In this paper we propose a novel approach towards improving the efficiency of Question Answering (QA) systems by filtering out questions that will not be answered by them. This is based on an interesting new finding: the answer confidence scores of state-of-the-art QA systems can be approximated well by models solely using the input question text. This enables preemptive filtering of questions that are not answered by the system due to their answer confidence scores being lower than the system threshold. Specifically, we learn Transformer-based question models by distilling Transformer-based answering models. Our experiments on three popular QA datasets and one industrial QA benchmark demonstrate the ability of our question models to approximate the Precision/Recall curves of the target QA system well. These question models, when used as filters, can effectively trade off lower computation cost of QA systems for lower Recall, e.g., reducing computation by 60%, while only losing 3-4% of Recall. Re --2.4 --0.8 --1.6 --2.5 --3.7 --0.8 --2.0 --4.1 --5.0 --6.1 FM * 2 M 1 % Filter 4.1 17.3 20.9 21.0 44.0 1.2 1.8 2.6 3.9 8.8 Re --2.8 --10.3 --9.0 --8.3 --7.4 --0.8 --0.8 --0.4 --2.1 --2.9 ASNQ M 1 Pr 68.2 74.6 77.2 79.6 90.5 75.7 79.6 81.9 84.5 92.9 Re 48.7 41.1 36.1 28.9 10.0 61.1 54.4 49.3 42.7 20.7 FM 1 * 2 M 1 % Filter 7.8 17.8 29.8 54.2 83.8 0.2 10.5 15.6 29.8 66.2
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- The Unreliability of Explanations in Few-shot Prompting for Textual ReasoningXi Ye, Greg DurrettNeurIPS 2022 · 被引用 272 次
- Learning to Reject with a Fixed Predictor: Application to DecontextualizationChristopher Mohri, Daniel Andor, Eunsol Choi, Michael Collins 等ICLR 2024 · 被引用 32 次
- Knowledge Transfer from Answer Ranking to Answer GenerationMatteo Gabburo, Rik Koncel-Kedziorski, Siddhant Garg, Luca Soldaini 等EMNLP 2022 · 被引用 4 次
- Learning Answer Generation using Supervision from Automatic Question Answering EvaluatorsMatteo Gabburo, Siddhant Garg, Rik Koncel-Kedziorski, Alessandro MoschittiACL 2023 · 被引用 4 次
- Fairness Beyond Performance: Revealing Reliability Disparities Across Groups in Legal NLPT. Y. S. S. Santosh, Irtiza ChowdhuryACL 2025
它引用的顶会 Paper9
- ALBERT: A Lite BERT for Self-supervised Learning of Language RepresentationsZhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel 等ICLR 2020 · 被引用 7,418 次
- MobileBERT: a Compact Task-Agnostic BERT for Resource-Limited DevicesZhiqing Sun, Hongkun Yu, Xiaodan Song, Renjie Liu 等ACL 2020 · 被引用 660 次
- ELECTRA: Pre-training Text Encoders as Discriminators Rather Than GeneratorsKevin Clark, Minh-Thang Luong, Quoc V. Le, Christopher D. ManningICLR 2020 · 被引用 541 次
- Learning to Retrieve Reasoning Paths over Wikipedia Graph for Question AnsweringAkari Asai, Kazuma Hashimoto, Hannaneh Hajishirzi, Richard Socher 等ICLR 2020 · 被引用 322 次
- FastBERT: a Self-distilling BERT with Adaptive Inference TimeWeijie Liu, Peng Zhou, Zhiruo Wang, Zhe Zhao 等ACL 2020 · 被引用 257 次
相关 Paper
- DeFormer: Decomposing Pre-trained Transformers for Faster Question AnsweringQingqing Cao, Harsh Trivedi, Aruna Balasubramanian, Niranjan BalasubramanianACL 2020 · 被引用 61 次
- Block-Skim: Efficient Question Answering for TransformerYue Guan, Zhengyi Li, Zhouhan Lin, Yuhao Zhu 等AAAI 2022 · 被引用 33 次
- PELA: Learning Parameter-Efficient Models with Low-Rank ApproximationYangyang Guo, Guangzhi Wang, Mohan S. KankanhalliCVPR 2024
- The Cascade Transformer: an Application for Efficient Answer Sentence SelectionLuca Soldaini, Alessandro MoschittiACL 2020 · 被引用 3 次
- Dynamic Context Pruning for Efficient and Interpretable Autoregressive TransformersSotiris Anagnostidis, Dario Pavllo, Luca Biggio, Lorenzo Noci 等NeurIPS 2023 · 被引用 95 次
