Which Shortcut Solution Do Question Answering Models Prefer to Learn?
Kazutoshi Shinoda, Saku Sugawara, Akiko Aizawa
Abstract
Question answering (QA) models for reading comprehension tend to exploit spurious correlations in training sets and thus learn shortcut solutions rather than the solutions intended by QA datasets. QA models that have learned shortcut solutions can achieve human-level performance in shortcut examples where shortcuts are valid, but these same behaviors degrade generalization potential on anti-shortcut examples where shortcuts are invalid. Various methods have been proposed to mitigate this problem, but they do not fully take the characteristics of shortcuts themselves into account. We assume that the learnability of shortcuts, i.e., how easy it is to learn a shortcut, is useful to mitigate the problem. Thus, we first examine the learnability of the representative shortcuts on extractive and multiple-choice QA datasets. Behavioral tests using biased training sets reveal that shortcuts that exploit answer positions and word-label correlations are preferentially learned for extractive and multiple-choice QA, respectively. We find that the more learnable a shortcut is, the flatter and deeper the loss landscape is around the shortcut solution in the parameter space. We also find that the availability of the preferred shortcuts tends to make the task easier to perform from an information-theoretic viewpoint. Lastly, we experimentally show that the learnability of shortcuts can be utilized to construct an effective QA training set; the more learnable a shortcut is, the smaller the proportion of anti-shortcut examples required to achieve comparable performance on shortcut and anti-shortcut examples. We claim that the learnability of shortcuts should be considered when designing mitigation methods.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers5
- ToMATO: Verbalizing the Mental States of Role-Playing LLMs for Benchmarking Theory of MindKazutoshi Shinoda, Nobukatsu Hojo, Kyosuke Nishida, Saki Mizuno et al.AAAI 2025 · 10 citations
- Measuring Inductive Biases of In-Context Learning with Underspecified DemonstrationsChenglei Si, Dan Friedman, Nitish Joshi, Shi Feng et al.ACL 2023 · 8 citations
- Navigate Beyond Shortcuts: Debiased Learning through the Lens of Neural CollapseYining Wang, Junjie Sun, Chenyue Wang, Mi Zhang et al.CVPR 2024 · 6 citations
- Debiasing Reward Models via Causally Motivated Inference-Time InterventionKazutoshi Shinoda, Kosuke Nishida, Kyosuke NishidaACL 2026
- Mitigating Shortcut Learning with InterpoLated LearningMichalis Korakakis, Andreas Vlachos, Adrian WellerACL 2025
Builds on11
- ReClor: A Reading Comprehension Dataset Requiring Logical ReasoningWeihao Yu, Zihang Jiang, Yanfei Dong, Jiashi FengICLR 2020 · 325 citations
- InfoBERT: Improving Robustness of Language Models from An Information Theoretic PerspectiveBoxin Wang, Shuohang Wang, Yu Cheng, Zhe Gan et al.ICLR 2021 · 132 citations
- Assessing the Benchmarking Capacity of Machine Reading Comprehension DatasetsSaku Sugawara, Pontus Stenetorp, Kentaro Inui, Akiko AizawaAAAI 2020 · 92 citations
- The Surprising Simplicity of the Early-Time Learning Dynamics of Neural NetworksWei Hu, Lechao Xiao, Ben Adlam, Jeffrey PenningtonNeurIPS 2020 · 77 citations
- Competency Problems: On Finding and Removing Artifacts in Language DataMatt Gardner, William Merrill, Jesse Dodge, Matthew E. Peters et al.EMNLP 2021 · 72 citations
Related papers
- Learning Concept Credible Models for Mitigating ShortcutsJiaxuan Wang, Sarah Jabbour, Maggie Makar, Michael W. Sjoding et al.NeurIPS 2022 · 8 citations
- Beyond Question-Based Biases: Assessing Multimodal Shortcut Learning in Visual Question AnsweringCorentin Dancette, Rémi Cadène, Damien Teney, Matthieu CordICCV 2021 · 95 citations
- Improving the robustness of NLI models with minimax trainingMichalis Korakakis, Andreas VlachosACL 2023 · 4 citations
- COMI: COrrect and MItigate Shortcut Learning Behavior in Deep Neural NetworksLili Zhao, Qi Liu, Linan Yue, Wei Chen et al.SIGIR 2024 · 9 citations
- Multi-source Meta Transfer for Low Resource Multiple-Choice Question AnsweringMing Yan, Hao Zhang, Di Jin, Joey Tianyi ZhouACL 2020 · 21 citations
