GATE: Graph Attention Transformer Encoder for Cross-lingual Relation and Event Extraction
Wasi Uddin Ahmad, Nanyun Peng, Kai-Wei Chang
Abstract
Recent progress in cross-lingual relation and event extraction use graph convolutional networks (GCNs) with universal dependency parses to learn language-agnostic sentence representations such that models trained on one language can be applied to other languages. However, GCNs struggle to model words with long-range dependencies or are not directly connected in the dependency tree. To address these challenges, we propose to utilize the self-attention mechanism where we explicitly fuse structural information to learn the dependencies between words with different syntactic distances. We introduce GATE, a Graph Attention Transformer Encoder, and test its cross-lingual transferability on relation and event extraction tasks. We perform experiments on the ACE05 dataset that includes three typologically different languages: English, Chinese, and Arabic. The evaluation results show that GATE outperforms three recently proposed methods by a large margin. Our detailed analysis reveals that due to the reliance on syntactic dependencies, GATE produces robust representations that facilitate transfer across languages.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 54fb05da-b3f9-42e8-a4f5-78f840c1d117Cited by top-tier papers15
- SGEITL: Scene Graph Enhanced Image-Text Learning for Visual Commonsense ReasoningZhecan Wang, Haoxuan You, Liunian Harold Li, Alireza Zareian et al.AAAI 2022 · 40 citations
- mLUKE: The Power of Entity Representations in Multilingual Pretrained Language ModelsRyokan Ri, Ikuya Yamada, Yoshimasa TsuruokaACL 2022 · 34 citations
- Language Model Priming for Cross-Lingual Event ExtractionSteven Fincke, Shantanu Agarwal, Scott Miller, Elizabeth BoscheeAAAI 2022 · 31 citations
- Improving Zero-Shot Cross-Lingual Transfer Learning via Robust TrainingKuan-Hao Huang, Wasi Uddin Ahmad, Nanyun Peng, Kai-Wei ChangEMNLP 2021 · 29 citations
- AMPERE: AMR-Aware Prefix for Generation-Based Event Argument Extraction ModelI-Hung Hsu, Zhiyu Xie, Kuan-Hao Huang, Prem Natarajan et al.ACL 2023 · 26 citations
Builds on3
- Unsupervised Cross-lingual Representation Learning at ScaleAlexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary et al.ACL 2020 · 539 citations
- Dependency Graph Enhanced Dual-transformer Structure for Aspect-based Sentiment ClassificationHao Tang, Donghong Ji, Chenliang Li, Qiji ZhouACL 2020 · 332 citations
- Domain Knowledge Empowered Structured Neural Net for End-to-End Event Temporal Relation ExtractionRujun Han, Yichao Zhou, Nanyun PengEMNLP 2020 · 38 citations
Related papers
- Dependency Structure-Enhanced Graph Attention Networks for Event DetectionQizhi Wan, Changxuan Wan, Keli Xiao, Kun Lu et al.AAAI 2024 · 7 citations
- Relation Extraction with Convolutional Network over Learnable Syntax-Transport GraphKai Sun, Richong Zhang, Yongyi Mao, Samuel Mensah et al.AAAI 2020 · 57 citations
- ACT: an Attentive Convolutional Transformer for Efficient Text ClassificationPengfei Li, Peixiang Zhong, Kezhi Mao, Dongzhe Wang et al.AAAI 2021 · 47 citations
- Graph Adaptive Semantic Transfer for Cross-domain Sentiment ClassificationKai Zhang, Qi Liu, Zhenya Huang, Mingyue Cheng et al.SIGIR 2022 · 12 citations
- A Closer Look at Graph Transformers: Cross-Aggregation and BeyondJiaming Zhuo, Ziyi Ma, Yintong Lu, Yuwei Liu et al.NeurIPS 2025 · 4 citations
