Learning a Cost-Effective Annotation Policy for Question Answering
Bernhard Kratzwald, Stefan Feuerriegel, Huan Sun
Abstract
State-of-the-art question answering (QA) relies upon large amounts of training data for which labeling is time consuming and thus expensive. For this reason, customizing QA systems is challenging. As a remedy, we propose a novel framework for annotating QA datasets that entails learning a cost-effective annotation policy and a semi-supervised annotation scheme. The latter reduces the human effort: it leverages the underlying QA system to suggest potential candidate annotations. Human annotators then simply provide binary feedback on these candidates. Our system is designed such that past annotations continuously improve the future performance and thus overall annotation cost. To the best of our knowledge, this is the first paper to address the problem of annotating questions with minimal annotation cost. We compare our framework against traditional manual annotations in an extensive set of experiments. We find that our approach can reduce up to 21.1% of the annotation cost.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext fb6d8d9d-aa8b-4161-b8b5-2f2cc434251eCited by top-tier papers4
- Learning with Different Amounts of Annotation: From Zero to Many LabelsShujian Zhang, Chengyue Gong, Eunsol ChoiEMNLP 2021 · 21 citations
- Multi-Source Test-Time Adaptation as Dueling Bandits for Extractive Question AnsweringHai Ye, Qizhe Xie, Hwee Tou NgACL 2023 · 2 citations
- QA Domain Adaptation using Hidden Space Augmentation and Self-Supervised Contrastive AdaptationZhenrui Yue, Huimin Zeng, Bernhard Kratzwald, Stefan Feuerriegel et al.EMNLP 2022 · 1 citation
- Simulating Bandit Learning from User Feedback for Extractive Question AnsweringGe Gao, Eunsol Choi, Yoav ArtziACL 2022
Builds on3
- Selective Question Answering under Domain ShiftAmita Kamath, Robin Jia, Percy LiangACL 2020 · 121 citations
- Neural Semantic Parsing in Low-Resource Settings with Back-Translation and Meta-LearningYibo Sun, Duyu Tang, Nan Duan, Yeyun Gong et al.AAAI 2020 · 25 citations
- An Imitation Game for Learning Semantic Parsers from User InteractionZiyu Yao, Yiqi Tang, Wen-tau Yih, Huan Sun et al.EMNLP 2020 · 18 citations
Related papers
- Optimal and Efficient Binary Questioning for Accelerated AnnotationFranco Marchesoni-Acland, Jean-Michel Morel, Josselin Kherroubi, Gabriele FaccioloAAAI 2025
- Continually Improving Extractive QA via Human FeedbackGe Gao, Hung-Ting Chen, Yoav Artzi, Eunsol ChoiEMNLP 2023 · 5 citations
- MCAL: Minimum Cost Human-Machine Active LabelingHang Qiu, Krishna Chintalapudi, Ramesh GovindanICLR 2023
- Instance-wise Supervision-level Optimization in Active LearningShinnosuke Matsuo, Riku Togashi, Ryoma Bise, Seiichi Uchida et al.CVPR 2025
- Discovering Dialogue Slots with Weak SupervisionVojtech Hudecek, Ondrej Dusek, Zhou YuACL 2021
