Fantastic Questions and Where to Find Them: FairytaleQA - An Authentic Dataset for Narrative Comprehension
Ying Xu, Dakuo Wang, Mo Yu, Daniel Ritchie, Bingsheng Yao, Tongshuang Wu, Zheng Zhang, Toby Jia-Jun Li, Nora Bradford, Branda Sun, Tran Bao Hoang, Yisi Sang
摘要
Question answering (QA) is a fundamental means to facilitate assessment and training of narrative comprehension skills for both machines and young children, yet there is scarcity of high-quality QA datasets carefully designed to serve this purpose. In particular, existing datasets rarely distinguish fine-grained reading skills, such as the understanding of varying narrative elements. Drawing on the reading education research, we introduce Fairy-taleQA 1 , a dataset focusing on narrative comprehension of kindergarten to eighth-grade students. Generated by educational experts based on an evidence-based theoretical framework, FairytaleQA consists of 10,580 explicit and implicit questions derived from 278 childrenfriendly stories, covering seven types of narrative elements or relations. Our dataset is valuable in two folds: First, we ran existing QA models on our dataset and confirmed that this annotation helps assess models' fine-grained learning skills. Second, the dataset supports question generation (QG) task in the education domain. Through benchmarking with QG models, we show that the QG model trained on FairytaleQA is capable of asking high-quality and more diverse questions.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper23
- FLASK: Fine-grained Language Model Evaluation based on Alignment Skill SetsSeonghyeon Ye, Doyoung Kim, Sungdong Kim, Hyeonbin Hwang 等ICLR 2024 · 被引用 176 次
- It is AI's Turn to Ask Humans a Question: Question-Answer Pair Generation for Children's Story BooksBingsheng Yao, Dakuo Wang, Tongshuang Wu, Zheng Zhang 等ACL 2022 · 被引用 58 次
- Multi-Level Optimal Transport for Universal Cross-Tokenizer Knowledge Distillation on Language ModelsXiao Cui, Mo Zhu, Yulei Qin, Liang Xie 等AAAI 2025 · 被引用 31 次
- Multi-Agent-as-Judge: Aligning LLM-Agent-Based Automated Evaluation with Multi-Dimensional Human EvaluationJiaju Chen, Yuxuan Lu, Xiaojie Wang, Huimin Zeng 等ACL 2026 · 被引用 30 次
- Exploring Parent's Needs for Children-Centered AI to Support Preschoolers' Interactive Storytelling and Reading ActivitiesYuling Sun, Jiaju Chen, Bingsheng Yao, Jiali Liu 等CSCW 2024 · 被引用 27 次
它引用的顶会 Paper2
- It is AI's Turn to Ask Humans a Question: Question-Answer Pair Generation for Children's Story BooksBingsheng Yao, Dakuo Wang, Tongshuang Wu, Zheng Zhang 等ACL 2022 · 被引用 58 次
- Educational Question Generation of Children Storybooks via Question Type Distribution Learning and Event-centric SummarizationZhenjie Zhao, Yufang Hou, Dakuo Wang, Mo Yu 等ACL 2022 · 被引用 50 次
相关 Paper
- Leader-Generator Net: Dividing Skill and Implicitness for Conquering FairytaleQAWei Peng, Wanshui Li, Yue HuSIGIR 2023 · 被引用 6 次
- Diversity Enhanced Narrative Question Generation for StorybooksHokeun Yoon, JinYeong BakEMNLP 2023 · 被引用 5 次
- StorySparkQA: Expert-Annotated QA Pairs with Real-World Knowledge for Children's Story-Based LearningJiaju Chen, Yuxuan Lu, Shao Zhang, Bingsheng Yao 等EMNLP 2024 · 被引用 6 次
- Tell as You Want: Customizing Image Narrative with Knowledge and ThoughtsZiwei Yao, Qian Wang, Ruiping Wang, Xilin ChenAAAI 2026 · 被引用 1 次
- FriendsQA: A New Large-Scale Deep Video Understanding Dataset with Fine-grained Topic Categorization for Story VideosZhengqian Wu, Ruizhe Li, Zijun Xu, Zhongyuan Wang 等AAAI 2025 · 被引用 2 次
