To Test Machine Comprehension, Start by Defining Comprehension
Jesse Dunietz, Gregory Burnham, Akash Bharadwaj, Owen Rambow, Jennifer Chu-Carroll, David A. Ferrucci
摘要
Many tasks aim to measure MACHINE READ-ING COMPREHENSION (MRC), often focusing on question types presumed to be difficult. Rarely, however, do task designers start by considering what systems should in fact comprehend. In this paper we make two key contributions. First, we argue that existing approaches do not adequately define comprehension; they are too unsystematic about what content is tested. Second, we present a detailed definition of comprehension-a TEM-PLATE OF UNDERSTANDING-for a widely useful class of texts, namely short narratives. We then conduct an experiment that strongly suggests existing systems are not up to the task of narrative understanding as we define it.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- AESOP: Abstract Encoding of Stories, Objects, and PicturesHareesh Ravi, Kushal Kafle, Scott Cohen, Jonathan Brandt 等ICCV 2021 · 被引用 19 次
- Braid: Weaving Symbolic and Neural Knowledge into Coherent Logical ExplanationsAditya Kalyanpur, Tom Breloff, David A. FerrucciAAAI 2022 · 被引用 18 次
- What's the Meaning of Superhuman Performance in Today's NLU?Simone Tedeschi, Johan Bos, Thierry Declerck, Jan Hajic 等ACL 2023 · 被引用 12 次
- English Machine Reading Comprehension Datasets: A SurveyDaria Dzendzik, Jennifer Foster, Carl VogelEMNLP 2021 · 被引用 8 次
- Situation and Behavior Understanding by Trope Detection on FilmsChen-Hsi Chang, Hung-Ting Su, Juiheng Hsu, Yu-Siang Wang 等WWW 2021 · 被引用 7 次
它引用的顶会 Paper1
相关 Paper
- VisualMRC: Machine Reading Comprehension on Document ImagesRyota Tanaka, Kyosuke Nishida, Sen YoshidaAAAI 2021 · 被引用 201 次
- MMM: Multi-Stage Multi-Task Learning for Multi-Choice Reading ComprehensionDi Jin, Shuyang Gao, Jiun-Yu Kao, Tagyoung Chung 等AAAI 2020 · 被引用 72 次
- Enhancing Multiple-choice Machine Reading Comprehension by Punishing Illogical InterpretationsYiming Ju, Yuanzhe Zhang, Zhixing Tian, Kang Liu 等EMNLP 2021 · 被引用 8 次
- Span Selection Pre-training for Question AnsweringMichael R. Glass, Alfio Gliozzo, Rishav Chakravarti, Anthony Ferritto 等ACL 2020 · 被引用 9 次
- Recurrent Chunking Mechanisms for Long-Text Machine Reading ComprehensionHongyu Gong, Yelong Shen, Dian Yu, Jianshu Chen 等ACL 2020 · 被引用 39 次
