R2F: A General Retrieval, Reading and Fusion Framework for Document-level Natural Language Inference
Hao Wang, Yixin Cao, Yangguang Li, Zhen Huang, Kun Wang, Jing Shao
Abstract
Document-level natural language inference (DOCNLI) is a new challenging task in natural language processing, aiming at judging the entailment relationship between a pair of hypothesis and premise documents. Current datasets and baselines largely follow sentence-level settings, but fail to address the issues raised by longer documents. In this paper, we establish a general solution, named Retrieval, Reading and Fusion (R 2 F) framework, and a new setting, by analyzing the main challenges of DOCNLI: interpretability, long-range dependency, and cross-sentence inference. The basic idea of the framework is to simplify documentlevel task into a set of sentence-level tasks, and improve both performance and interpretability with the power of evidence. For each hypothesis sentence, the framework retrieves evidence sentences from the premise, and reads to estimate its credibility. Then the sentencelevel results are fused to judge the relationship between the documents. For the setting, we contribute complementary evidence and entailment label annotation on hypothesis sentences, for interpretability study. Our experimental results show that R 2 F framework can obtain state-of-the-art performance and is robust for diverse evidence retrieval methods. Moreover, it can give more interpretable prediction results. Our model and code are released at https://github.com/phoenixsecularbird/R2F .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e8f261ea-3833-4d18-8914-97a3f7fd12dbBuilds on11
- SimCSE: Simple Contrastive Learning of Sentence EmbeddingsTianyu Gao, Xingcheng Yao, Danqi ChenEMNLP 2021 · 2,496 citations
- Adversarial NLI: A New Benchmark for Natural Language UnderstandingYixin Nie, Adina Williams, Emily Dinan, Mohit Bansal et al.ACL 2020 · 602 citations
- Extractive Summarization as Text MatchingMing Zhong, Pengfei Liu, Yiran Chen, Danqing Wang et al.ACL 2020 · 410 citations
- CogLTX: Applying BERT to Long TextsMing Ding, Chang Zhou, Hongxia Yang, Jie TangNeurIPS 2020 · 163 citations
- Zoom Out and Observe: News Environment Perception for Fake News DetectionQiang Sheng, Juan Cao, Xueyao Zhang, Rundong Li et al.ACL 2022 · 103 citations
Related papers
- Cross-Document Cross-Lingual NLI via RST-Enhanced Graph Fusion and Interpretability PredictionMengying Yuan, Wenhao Wang, Zixuan Wang, Yujie Huang et al.EMNLP 2025
- DocInfer: Document-level Natural Language Inference using Optimal Evidence SelectionPuneet Mathur, Gautam Kunapuli, Riyaz A. Bhat, Manish Shrivastava et al.EMNLP 2022 · 5 citations
- Revisiting Document-Level Relation Extraction with Context-Guided Link PredictionMonika Jain, Raghava Mutharaju, Ramakanth Kavuluru, Kuldeep SinghAAAI 2024 · 17 citations
- Reasoning with Latent Structure Refinement for Document-Level Relation ExtractionGuoshun Nan, Zhijiang Guo, Ivan Sekulic, Wei LuACL 2020 · 294 citations
- NLI4CT: Multi-Evidence Natural Language Inference for Clinical Trial ReportsMaël Jullien, Marco Valentino, Hannah Frost, Paul O'Regan et al.EMNLP 2023 · 6 citations
