Distractor Generation in Multiple-Choice Tasks: A Survey of Methods, Datasets, and Evaluation
Elaf Alhazmi, Quan Sheng, Wei Emma Zhang, Munazza Zaib, Ahoud Alhazmi
Abstract
The distractor generation task focuses on generating incorrect but plausible options for objective questions such as fill-in-the-blank and multiple-choice questions. This task is widely utilized in educational settings across various domains and subjects. The effectiveness of these questions in assessments relies on the quality of the distractors, as they challenge examinees to select the correct answer from a set of misleading options. The evolution of artificial intelligence (AI) has transitioned the task from traditional methods to the use of neural networks and pre-trained language models. This shift has established new benchmarks and expanded the use of advanced deep learning methods in generating distractors. This survey explores distractor generation tasks, datasets, methods, and current evaluation metrics for English objective questions, covering both textbased and multi-modal domains. It also evaluates existing AI models and benchmarks and discusses potential future research directions 1 .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers5
- Tailoring Diagnostic Modeling to Individual Learners: Personalized Distractor Generation via MCTS-Guided Reasoning ReconstructionTao Wu, Jingyuan Chen, Wang Lin, Jian Zhan et al.ACL 2026 · 3 citations
- Better Datasets Start from RefineLab: Automatic Optimization for High-Quality Dataset RefinementXiaonan Luo, Yue Huang, Ping He, Xiangliang ZhangAAAI 2026 · 1 citation
- MicroVQA: A Multimodal Reasoning Benchmark for Microscopy-Based Scientific ResearchJames Burgess, Jeffrey J. Nirschl, Laura Bravo-Sánchez, Alejandro Lozano et al.CVPR 2025
- EIFFEL: a novel benchmark to measure bias of English heavy training on French idiomatic expressionsCharlotte Noel, Nicholas Asher, Olivier Gouvert, Farah Benamara et al.ACL 2026
- Question Difficulty Estimation for Large Language Models via Answer Plausibility ScoringJamshid Mozafari, Bhawna Piryani, Adam JatowtACL 2026
Builds on20
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma et al.NeurIPS 2022 · 22,562 citations
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad et al.ACL 2020 · 1,224 citations
- QASC: A Dataset for Question Answering via Sentence CompositionTushar Khot, Peter Clark, Michal Guerquin, Peter Jansen et al.AAAI 2020 · 387 citations
Related papers
- Difficulty-Controllable Cloze Question Distractor GenerationSeokhoon Kang, Yejin Jeon, Seonjeong Hwang, Gary LeeACL 2026
- Knowledge-Driven Distractor Generation for Cloze-Style Multiple Choice QuestionsSiyu Ren, Kenny Q. ZhuAAAI 2021 · 62 citations
- Generating Plausible Distractors for Multiple-Choice Questions via Student Choice PredictionYooseop Lee, Suin Kim, Yohan JoACL 2025
- DiVERT: Distractor Generation with Variational Errors Represented as Text for Math Multiple-choice QuestionsNigel Fernandez, Alexander Scarlatos, Wanyong Feng, Simon Woodhead et al.EMNLP 2024 · 10 citations
- Co-Attention Hierarchical Network: Generating Coherent Long Distractors for Reading ComprehensionXiaorui Zhou, Senlin Luo, Yunfang WuAAAI 2020 · 41 citations
