Semi-Supervised Learning on Meta Structure: Multi-Task Tagging and Parsing in Low-Resource Scenarios
Kyungtae Lim, Jay Yoon Lee, Jaime G. Carbonell, Thierry Poibeau
Abstract
Multi-view learning makes use of diverse models arising from multiple sources of input or different feature subsets for the same task. For example, a given natural language processing task can combine evidence from models arising from character, morpheme, lexical, or phrasal views. The most common strategy with multi-view learning, especially popular in the neural network community, is to unify multiple representations into one unified vector through concatenation, averaging, or pooling, and then build a single-view model on top of the unified representation. As an alternative, we examine whether building one model per view and then unifying the different models can lead to improvements, especially in low-resource scenarios. More specifically, taking inspiration from co-training methods, we propose a semi-supervised learning approach based on multi-view models through consensus promotion, and investigate whether this improves overall performance. To test the multi-view hypothesis, we use moderately low-resource scenarios for nine languages and test the performance of the joint model for part-of-speech tagging and dependency parsing. The proposed model shows significant improvements across the test cases, with average gains of -0.9 ∼ +9.3 labeled attachment score (LAS) points. We also investigate the effect of unlabeled data on the proposed model by varying the amount of training data and by using different domains of unlabeled data.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 64709c7b-92b1-417a-9620-e415f1d4cf6fCited by top-tier papers2
- CoMRes: Semi-Supervised Time Series Forecasting Utilizing Consensus Promotion of Multi-ResolutionYunju Cho, Jay-Yoon LeeICLR 2025
- Revisiting Tri-training of Dependency ParsersJoachim Wagner, Jennifer FosterEMNLP 2021
Related papers
- Multi-View Cross-Lingual Structured Prediction with Minimum SupervisionZechuan Hu, Yong Jiang, Nguyen Bach, Tao Wang et al.ACL 2021
- Graph-Based Multilingual Label Propagation for Low-Resource Part-of-Speech TaggingAyyoob Imani, Silvia Severini, Masoud Jalili Sabet, François Yvon et al.EMNLP 2022 · 8 citations
- Unsupervised Cross-Lingual Part-of-Speech Tagging for Truly Low-Resource ScenariosRamy Eskander, Smaranda Muresan, Michael CollinsEMNLP 2020 · 16 citations
- Weakly Supervised POS Taggers Perform Poorly on Truly Low-Resource LanguagesKatharina Kann, Ophélie Lacroix, Anders SøgaardAAAI 2020 · 21 citations
- Enhancing Low-Resource Relation Representations through Multi-View DecouplingChenghao Fan, Wei Wei, Xiaoye Qu, Zhenyi Lu et al.AAAI 2024 · 10 citations
