Retrofitting Structure-aware Transformer Language Model for End Tasks
Hao Fei, Yafeng Ren, Donghong Ji
Abstract
We consider retrofitting structure-aware Transformer language model for facilitating end tasks by proposing to exploit syntactic distance to encode both the phrasal constituency and dependency connection into the language model. A middle-layer structural learning strategy is leveraged for structure integration, accomplished with main semantic task training under multi-task learning scheme. Experimental results show that the retrofitted structure-aware Transformer language model achieves improved perplexity, meanwhile inducing accurate syntactic phrases. By performing structure-aware fine-tuning, our model achieves significant improvements for both semantic-and syntactic-dependent tasks.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers6
- LasUIE: Unifying Information Extraction with Latent Adaptive Structure-aware Generative Language ModelHao Fei, Shengqiong Wu, Jingye Li, Bobo Li et al.NeurIPS 2022 · 114 citations
- Modeling Hierarchical Structures with Continuous Recursive Neural NetworksJishnu Ray Chowdhury, Cornelia CarageaICML 2021 · 18 citations
- Beam Tree Recursive CellsJishnu Ray Chowdhury, Cornelia CarageaICML 2023 · 7 citations
- Efficient Beam Tree RecursionJishnu Ray Chowdhury, Cornelia CarageaNeurIPS 2023 · 4 citations
- Retrofitting Light-weight Language Models for Emotions using Supervised Contrastive LearningSapan Shah, Sreedhar Reddy, Pushpak BhattacharyyaEMNLP 2023 · 4 citations
Related papers
- GiLT: Augmenting Transformer Language Models with Dependency GraphsTianyu Huang, Yida Zhao, Chuyan Zhou, Kewei TuACL 2026
- Exploiting Syntactic Structure for Better Language Modeling: A Syntactic Distance ApproachWenyu Du, Zhouhan Lin, Yikang Shen, Timothy J. O'Donnell et al.ACL 2020 · 15 citations
- Strengthening Structural Inductive Biases by Pre-training to Perform Syntactic TransformationsMatthias Lindemann, Alexander Koller, Ivan TitovEMNLP 2024
- Syntax-Enhanced Pre-trained ModelZenan Xu, Daya Guo, Duyu Tang, Qinliang Su et al.ACL 2021
- Dependency-based Mixture Language ModelsZhixian Yang, Xiaojun WanACL 2022 · 3 citations
