Meta Self-training for Few-shot Neural Sequence Labeling
Yaqing Wang, Subhabrata Mukherjee, Haoda Chu, Yuancheng Tu, Ming Wu, Jing Gao, Ahmed Hassan Awadallah
Abstract
Neural sequence labeling is widely adopted for many Natural Language Processing (NLP) tasks, such as Named Entity Recognition (NER) and slot tagging for dialog systems and semantic parsing. Recent advances with large-scale pre-trained language models have shown remarkable success in these tasks when fine-tuned on large amounts of task-specific labeled data. However, obtaining such large-scale labeled training data is not only costly, but also may not be feasible in many sensitive user applications due to data access and privacy constraints. This is exacerbated for sequence labeling tasks requiring such annotations at token-level. In this work, we develop techniques to address the label scarcity challenge for neural sequence labeling models. Specifically, we propose a meta self-training framework which leverages very few manually annotated labels for training neural sequence models. While self-training serves as an effective mechanism to learn from large amounts of unlabeled data via iterative knowledge exchange -- meta-learning helps in adaptive sample re-weighting to mitigate error propagation from noisy pseudo-labels. Extensive experiments on six benchmark datasets including two for massive multilingual NER and four slot tagging datasets for task-oriented dialog systems demonstrate the effectiveness of our method. With only 10 labeled examples for each class in each task, the proposed method achieves 10% improvement over state-of-the-art methods demonstrating its effectiveness for limited training labels regime.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a372b62e-af1f-47a9-b4e8-cabcd6e9bbceCited by top-tier papers5
- Neighborhood-Regularized Self-Training for Learning with Few LabelsRan Xu, Yue Yu, Hejie Cui, Xuan Kan et al.AAAI 2023 · 29 citations
- Self-Training for Label-Efficient Information Extraction from Semi-Structured Web-PagesRitesh Sarkhel, Binxuan Huang, Colin Lockard, Prashant ShiralkarVLDB 2023 · 11 citations
- Few-Shot Joint Multimodal Entity-Relation Extraction via Knowledge-Enhanced Cross-modal Prompt ModelLi Yuan, Yi Cai, Junsheng HuangACM MM 2024 · 9 citations
- CLUR: Uncertainty Estimation for Few-Shot Text Classification with Contrastive LearningJianfeng He, Xuchao Zhang, Shuo Lei, Abdulaziz Alhamadani et al.KDD 2023 · 4 citations
- KnowDA: All-in-One Knowledge Mixture Model for Data Augmentation in Low-Resource NLPYufei Wang, Jiayi Zheng, Can Xu, Xiubo Geng et al.ICLR 2023 · 2 citations
Builds on6
- Revisiting Self-Training for Neural Sequence GenerationJunxian He, Jiatao Gu, Jiajun Shen, Marc'Aurelio RanzatoICLR 2020 · 294 citations
- Understanding Self-Training for Gradual Domain AdaptationAnanya Kumar, Tengyu Ma, Percy LiangICML 2020 · 266 citations
- BOND: BERT-Assisted Open-Domain Named Entity Recognition with Distant SupervisionChen Liang, Yue Yu, Haoming Jiang, Siawpeng Er et al.KDD 2020 · 118 citations
- SeqVAT: Virtual Adversarial Training for Semi-Supervised Sequence LabelingLuoxin Chen, Weitong Ruan, Xinyue Liu, Jianhua LuACL 2020 · 118 citations
- Don't Stop Pretraining: Adapt Language Models to Domains and TasksSuchin Gururangan, Ana Marasovic, Swabha Swayamdipta, Kyle Lo et al.ACL 2020 · 93 citations
Related papers
- MetaTS: Meta Teacher-Student Network for Multilingual Sequence Labeling with Minimal SupervisionZheng Li, Danqing Zhang, Tianyu Cao, Ying Wei et al.EMNLP 2021 · 7 citations
- Uncertainty-Aware Self-Training for Low-Resource Neural Sequence LabelingJianing Wang, Chengyu Wang, Jun Huang, Ming Gao et al.AAAI 2023 · 5 citations
- Few-Shot Named Entity Recognition: An Empirical Baseline StudyJiaxin Huang, Chunyuan Li, Krishan Subudhi, Damien Jose et al.EMNLP 2021 · 97 citations
- Self-training Improves Pre-training for Few-shot Learning in Task-oriented Dialog SystemsFei Mi, Wanhao Zhou, Lingjing Kong, Fengyu Cai et al.EMNLP 2021 · 18 citations
- Augmented Natural Language for Generative Sequence LabelingBen Athiwaratkun, Cícero Nogueira dos Santos, Jason Krone, Bing XiangEMNLP 2020 · 54 citations
