Cross-Document Cross-Lingual NLI via RST-Enhanced Graph Fusion and Interpretability Prediction
Mengying Yuan, Wenhao Wang, Zixuan Wang, Yujie Huang, Kangli Wei, Fei Li, Chong Teng, Donghong Ji
Abstract
Natural Language Inference (NLI) is a fundamental task in natural language processing. While NLI has developed many subdirections such as sentence-level NLI, documentlevel NLI and cross-lingual NLI, Cross-Document Cross-Lingual NLI (CDCL-NLI) remains largely unexplored. In this paper, we propose a novel paradigm: CDCL-NLI, which extends traditional NLI capabilities to multidocument, multilingual scenarios. To support this task, we construct a high-quality CDCL-NLI dataset including 25,410 instances and spanning 26 languages. To address the limitations of previous methods on CDCL-NLI task, we further propose an innovative method that integrates RST-enhanced graph fusion with interpretability-aware prediction. Our approach leverages RST (Rhetorical Structure Theory) within heterogeneous graph neural networks for cross-document context modeling, and employs a structure-aware semantic alignment based on lexical chains for crosslingual understanding. For NLI interpretability, we develop an EDU (Elementary Discourse Unit)-level attribution framework that produces extractive explanations. Extensive experiments demonstrate our approach's superior performance, achieving significant improvements over both conventional NLI models as well as large language models. Our work sheds light on the study of NLI and will bring research interest on cross-document cross-lingual context understanding, hallucination elimination and interpretability inference. Our code and dataset are available at CDCL-NLI-link.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 59416bbe-98a0-40c0-ab5f-9c5919e263a5Cited by top-tier papers3
- HalluCitation Matters: Revealing the Impact of Hallucinated References with 300 Hallucinated Papers in ACL ConferencesYusuke Sakai, Hidetaka Kamigaito, Taro WatanabeACL 2026 · 18 citations
- Beyond Chunking: Discourse-Aware Hierarchical Retrieval for Long Document Question AnsweringHuiyao Chen, Yi Yang, Yinghui Li, Meishan Zhang et al.ACL 2026 · 6 citations
- MCP-Persona: Benchmarking LLM Agents on Real-World Personal Applications via Environment SimulationWenhao Wang, Peizhi Niu, Gongyi Zou, Xiyuan Yang et al.ICML 2026 · 1 citation
Builds on6
- Unsupervised Cross-lingual Representation Learning at ScaleAlexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary et al.ACL 2020 · 539 citations
- Structure-Augmented Text Representation Learning for Efficient Knowledge Graph CompletionBo Wang, Tao Shen, Guodong Long, Tianyi Zhou et al.WWW 2021 · 322 citations
- EvidenceNet: Evidence Fusion Network for Fact VerificationZhendong Chen, Siu Cheung Hui, Fuzhen Zhuang, Lejian Liao et al.WWW 2022 · 32 citations
- Fact or Fiction: Verifying Scientific ClaimsDavid Wadden, Shanchuan Lin, Kyle Lo, Lucy Lu Wang et al.EMNLP 2020 · 6 citations
- DocInfer: Document-level Natural Language Inference using Optimal Evidence SelectionPuneet Mathur, Gautam Kunapuli, Riyaz A. Bhat, Manish Shrivastava et al.EMNLP 2022 · 5 citations
Related papers
- R2F: A General Retrieval, Reading and Fusion Framework for Document-level Natural Language InferenceHao Wang, Yixin Cao, Yangguang Li, Zhen Huang et al.EMNLP 2022
- From Nodes to Narratives: Explaining Graph Neural Networks with LLMs and Graph ContextPeyman Baghershahi, Gregoire Fournier, Pranav Nyati, Sourav MedyaACL 2026 · 9 citations
- A Multilingual Perspective Towards the Evaluation of Attribution Methods in Natural Language InferenceKerem Zaman, Yonatan BelinkovEMNLP 2022 · 7 citations
- Mind the Gap: Cross-Lingual Information Retrieval with Hierarchical Knowledge EnhancementFuwei Zhang, Zhao Zhang, Xiang Ao, Dehong Gao et al.AAAI 2022 · 26 citations
- LIREx: Augmenting Language Inference with Relevant ExplanationsXinyan Zhao, V. G. Vinod VydiswaranAAAI 2021 · 41 citations
