Exploring Self-Distillation Based Relational Reasoning Training for Document-Level Relation Extraction
Liang Zhang, Jinsong Su, Zijun Min, Zhongjian Miao, Qingguo Hu, Biao Fu, Xiaodong Shi, Yidong Chen
摘要
Document-level relation extraction (RE) aims to extract relational triples from a document. One of its primary challenges is to predict implicit relations between entities, which are not explicitly expressed in the document but can usually be extracted through relational reasoning. Previous methods mainly implicitly model relational reasoning through the interaction among entities or entity pairs. However, they suffer from two deficiencies: 1) they often consider only one reasoning pattern, of which coverage on relational triples is limited; 2) they do not explicitly model the process of relational reasoning. In this paper, to deal with the first problem, we propose a document-level RE model with a reasoning module that contains a core unit, the reasoning multi-head self-attention unit. This unit is a variant of the conventional multi-head self-attention and utilizes four attention heads to model four common reasoning patterns, respectively, which can cover more relational triples than previous methods. Then, to address the second issue, we propose a self-distillation training framework, which contains two branches sharing parameters. In the first branch, we first randomly mask some entity pair feature vectors in the document, and then train our reasoning module to infer their relations by exploiting the feature information of other related entity pairs. By doing so, we can explicitly model the process of relational reasoning. However, because the additional masking operation is not used during testing, it causes an input gap between training and testing scenarios, which would hurt the model performance. To reduce this gap, we perform conventional supervised training without masking operation in the second branch and utilize Kullback-Leibler divergence loss to minimize the difference between the predictions of the two branches. Finally, we conduct comprehensive experiments on three benchmark datasets, of which experimental results demonstrate that our model consistently outperforms all competitive baselines. Our source code is available at https://github.com/DeepLearnXMU/DocRE-SD
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Mitigating Catastrophic Forgetting in Large Language Models with Self-Synthesized RehearsalJianheng Huang, Leyang Cui, Ante Wang, Chengyi Yang 等ACL 2024 · 被引用 13 次
- Exploring All-In-One Knowledge Distillation Framework for Neural Machine TranslationZhongjian Miao, Wen Zhang, Jinsong Su, Xiang Li 等EMNLP 2023 · 被引用 5 次
- HyperNetwork-based Decoupling to Improve Model Generalization for Few-Shot Relation ExtractionLiang Zhang, Chulun Zhou, Fandong Meng, Jinsong Su 等EMNLP 2023 · 被引用 3 次
- Multi-Level Cross-Modal Alignment for Speech Relation ExtractionLiang Zhang, Zhen Yang, Biao Fu, Ziyao Lu 等EMNLP 2024 · 被引用 2 次
- LLM-OREF: An Open Relation Extraction Framework Based on Large Language ModelsHongyao Tu, Liang Zhang, Yujie Lin, Xin Lin 等EMNLP 2025 · 被引用 2 次
它引用的顶会 Paper9
- Document-Level Relation Extraction with Adaptive Thresholding and Localized Context PoolingWenxuan Zhou, Kevin Huang, Tengyu Ma, Jing HuangAAAI 2021 · 被引用 360 次
- Reasoning with Latent Structure Refinement for Document-Level Relation ExtractionGuoshun Nan, Zhijiang Guo, Ivan Sekulic, Wei LuACL 2020 · 被引用 294 次
- Double Graph Based Reasoning for Document-level Relation ExtractionShuang Zeng, Runxin Xu, Baobao Chang, Lei LiEMNLP 2020 · 被引用 238 次
- Coreferential Reasoning Learning for Language RepresentationDeming Ye, Yankai Lin, Jiaju Du, Zhenghao Liu 等EMNLP 2020 · 被引用 164 次
- Document-Level Relation Extraction with ReconstructionWang Xu, Kehai Chen, Tiejun ZhaoAAAI 2021 · 被引用 131 次
相关 Paper
- Towards Better Document-level Relation Extraction via Iterative InferenceLiang Zhang, Jinsong Su, Yidong Chen, Zhongjian Miao 等EMNLP 2022 · 被引用 11 次
- Revisiting Document-Level Relation Extraction with Context-Guided Link PredictionMonika Jain, Raghava Mutharaju, Ramakanth Kavuluru, Kuldeep SinghAAAI 2024 · 被引用 17 次
- Global-to-Local Neural Networks for Document-Level Relation ExtractionDifeng Wang, Wei Hu, Ermei Cao, Weijian SunEMNLP 2020 · 被引用 122 次
- Entity-centered Cross-document Relation ExtractionFengqi Wang, Fei Li, Hao Fei, Jingye Li 等EMNLP 2022 · 被引用 49 次
- SRF: Enhancing Document-Level Relation Extraction with a Novel Secondary Reasoning FrameworkFu Zhang, Qi Miao, Jingwei Cheng, Hongsen Yu 等EMNLP 2024 · 被引用 2 次
