Unsupervised Graph-Text Mutual Conversion with a Unified Pretrained Language Model
Yi Xu, Shuqian Sheng, Jiexing Qi, Luoyi Fu, Zhouhan Lin, Xinbing Wang, Chenghu Zhou
Abstract
Graph-to-text (G2T) generation and text-tograph (T2G) triple extraction are two essential tasks for knowledge graphs. Existing unsupervised approaches become suitable candidates for jointly learning the two tasks due to their avoidance of using graph-text parallel data. However, they adopt multiple complex modules and still require entity information or relation type for training. To this end, we propose INFINITY, a simple yet effective unsupervised method with a unified pretrained language model that does not introduce external annotation tools or additional parallel information. It achieves fully unsupervised graph-text mutual conversion for the first time. Specifically, INFINITY treats both G2T and T2G as a bidirectional sequence generation task by finetuning only one pretrained seq2seq model. A novel back-translation-based framework is then designed to generate synthetic parallel data automatically. Besides, we investigate the impact of graph linearization and introduce the structure-aware fine-tuning strategy to alleviate possible performance deterioration via retaining structural information in graph sequences. As a fully unsupervised framework, INFINITY is empirically verified to outperform state-ofthe-art baselines for G2T and T2G tasks. Additionally, we also devise a new training setting called cross learning for low-resource unsupervised information extraction. G2T Relation Classification
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e01d987a-aa92-461e-a891-17a1135f708bCited by top-tier papers2
- The Zeno's Paradox of 'Low-Resource' LanguagesHellina Hailu Nigatu, Atnafu Lambebo Tonja, Benjamin Rosman, Thamar Solorio et al.EMNLP 2024 · 10 citations
- Generative Subgraph Retrieval for Knowledge Graph-Grounded Dialog GenerationJinyoung Park, Minseok Joo, Joo-Kyung Kim, Hyunwoo J. KimEMNLP 2024 · 3 citations
Builds on8
- BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and ComprehensionMike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad et al.ACL 2020 · 1,224 citations
- A Novel Cascade Binary Tagging Framework for Relational Triple ExtractionZhepei Wei, Jianlin Su, Yue Wang, Yuan Tian et al.ACL 2020 · 610 citations
- KnowPrompt: Knowledge-aware Prompt-tuning with Synergistic Optimization for Relation ExtractionXiang Chen, Ningyu Zhang, Xin Xie, Shumin Deng et al.WWW 2022 · 488 citations
- Contrastive Triple Extraction with Generative TransformerHongbin Ye, Ningyu Zhang, Shumin Deng, Mosha Chen et al.AAAI 2021 · 146 citations
- A Partition Filter Network for Joint Entity and Relation ExtractionZhiheng Yan, Chong Zhang, Jinlan Fu, Qi Zhang et al.EMNLP 2021 · 142 citations
Related papers
- An Unsupervised Joint System for Text Generation from Knowledge Graphs and Semantic ParsingMartin Schmitt, Sahand Sharifzadeh, Volker Tresp, Hinrich SchützeEMNLP 2020 · 2 citations
- Latent Constraints on Unsupervised Text-Graph Alignment with Information AsymmetryJidong Tian, Wenqing Chen, Yitian Li, Caoyun Fan et al.AAAI 2023
- Learn to Cross-lingual Transfer with Meta Graph Learning Across Heterogeneous LanguagesZheng Li, Mukul Kumar, William Headden, Bing Yin et al.EMNLP 2020 · 26 citations
- UniGTE: Unified Graph-Text Encoding for Zero-Shot Generalization across Graph Tasks and DomainsDuo Wang, Yuan Zuo, Guangyue Lu, Junjie WuNeurIPS 2025 · 9 citations
- Zero-Shot Information Extraction as a Unified Text-to-Triple TranslationChenguang Wang, Xiao Liu, Zui Chen, Haoyun Hong et al.EMNLP 2021 · 37 citations
