Recurrent Chunking Mechanisms for Long-Text Machine Reading Comprehension
Hongyu Gong, Yelong Shen, Dian Yu, Jianshu Chen, Dong Yu
摘要
In this paper, we study machine reading comprehension (MRC) on long texts, where a model takes as inputs a lengthy document and a question and then extracts a text span from the document as an answer. State-of-the-art models tend to use a pretrained transformer model (e.g., BERT) to encode the joint contextual information of document and question. However, these transformer-based models can only take a fixed-length (e.g., 512) text as its input. To deal with even longer text inputs, previous approaches usually chunk them into equally-spaced segments and predict answers based on each segment independently without considering the information from other segments. As a result, they may form segments that fail to cover the correct answer span or retain insufficient contexts around it, which significantly degrades the performance. Moreover, they are less capable of answering questions that need cross-segment information. We propose to let a model learn to chunk in a more flexible way via reinforcement learning: a model can decide the next segment that it wants to process in either direction. We also employ recurrent mechanisms to enable information to flow across segments. Experiments on three MRC datasets -CoQA, QuAC, and TriviaQA -demonstrate the effectiveness of our proposed recurrent chunking mechanisms: we can obtain segments that are more likely to contain complete answers and at the same time provide sufficient contexts around the ground truth answers for better predictions.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Capturing Global Structural Information in Long Document Question Answering with Compressive Graph Selector NetworkYuxiang Nie, Heyan Huang, Wei Wei, Xian-Ling MaoEMNLP 2022 · 被引用 11 次
- M3DocDep: Multi-modal, Multi-page, Multi-document Dependency Chunking with Large Vision-Language ModelsJoongmin Shin, Jeongbae Park, Jaehyung Seo, Heuiseok LimCVPR 2026
- HiKEY: Hierarchical Multimodal Retrieval for Open-Domain Document Question AnsweringJoongmin Shin, Gyuho Shim, Jeongbae Park, Jaehyung Seo 等ACL 2026
- MultiDocFusion : Hierarchical and Multimodal Chunking Pipeline for Enhanced RAG on Long Industrial DocumentsJoongmin Shin, Chanjun Park, Jeongbae Park, Jaehyung Seo 等EMNLP 2025
相关 Paper
- Span Selection Pre-training for Question AnsweringMichael R. Glass, Alfio Gliozzo, Rishav Chakravarti, Anthony Ferritto 等ACL 2020 · 被引用 9 次
- ReCO: A Large Scale Chinese Reading Comprehension Dataset on OpinionBingning Wang, Ting Yao, Qi Zhang, Jingfang Xu 等AAAI 2020 · 被引用 26 次
- VisualMRC: Machine Reading Comprehension on Document ImagesRyota Tanaka, Kyosuke Nishida, Sen YoshidaAAAI 2021 · 被引用 201 次
- Robust Domain Adaptation for Machine Reading ComprehensionLiang Jiang, Zhenyu Huang, Jia Liu, Zujie Wen 等AAAI 2023 · 被引用 1 次
- Capturing Greater Context for Question GenerationLuu Anh Tuan, Darsh J. Shah, Regina BarzilayAAAI 2020 · 被引用 77 次
