Semi-Supervised Learning on Meta Structure: Multi-Task Tagging and Parsing in Low-Resource Scenarios
Kyungtae Lim, Jay Yoon Lee, Jaime G. Carbonell, Thierry Poibeau
摘要
Multi-view learning makes use of diverse models arising from multiple sources of input or different feature subsets for the same task. For example, a given natural language processing task can combine evidence from models arising from character, morpheme, lexical, or phrasal views. The most common strategy with multi-view learning, especially popular in the neural network community, is to unify multiple representations into one unified vector through concatenation, averaging, or pooling, and then build a single-view model on top of the unified representation. As an alternative, we examine whether building one model per view and then unifying the different models can lead to improvements, especially in low-resource scenarios. More specifically, taking inspiration from co-training methods, we propose a semi-supervised learning approach based on multi-view models through consensus promotion, and investigate whether this improves overall performance. To test the multi-view hypothesis, we use moderately low-resource scenarios for nine languages and test the performance of the joint model for part-of-speech tagging and dependency parsing. The proposed model shows significant improvements across the test cases, with average gains of -0.9 ∼ +9.3 labeled attachment score (LAS) points. We also investigate the effect of unlabeled data on the proposed model by varying the amount of training data and by using different domains of unlabeled data.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- CoMRes: Semi-Supervised Time Series Forecasting Utilizing Consensus Promotion of Multi-ResolutionYunju Cho, Jay-Yoon LeeICLR 2025
- Revisiting Tri-training of Dependency ParsersJoachim Wagner, Jennifer FosterEMNLP 2021
相关 Paper
- Multi-View Cross-Lingual Structured Prediction with Minimum SupervisionZechuan Hu, Yong Jiang, Nguyen Bach, Tao Wang 等ACL 2021
- Graph-Based Multilingual Label Propagation for Low-Resource Part-of-Speech TaggingAyyoob Imani, Silvia Severini, Masoud Jalili Sabet, François Yvon 等EMNLP 2022 · 被引用 8 次
- Unsupervised Cross-Lingual Part-of-Speech Tagging for Truly Low-Resource ScenariosRamy Eskander, Smaranda Muresan, Michael CollinsEMNLP 2020 · 被引用 16 次
- Weakly Supervised POS Taggers Perform Poorly on Truly Low-Resource LanguagesKatharina Kann, Ophélie Lacroix, Anders SøgaardAAAI 2020 · 被引用 21 次
- Enhancing Low-Resource Relation Representations through Multi-View DecouplingChenghao Fan, Wei Wei, Xiaoye Qu, Zhenyi Lu 等AAAI 2024 · 被引用 10 次
