Meta Fine-Tuning Neural Language Models for Multi-Domain Text Mining
Chengyu Wang, Minghui Qiu, Jun Huang, Xiaofeng He
摘要
Pre-trained neural language models bring significant improvement for various NLP tasks, by fine-tuning the models on task-specific training sets. During fine-tuning, the parameters are initialized from pre-trained models directly, which ignores how the learning process of similar NLP tasks in different domains is correlated and mutually reinforced. In this paper, we propose an effective learning procedure named Meta Fine-Tuning (MFT), serving as a meta-learner to solve a group of similar NLP tasks for neural language models. Instead of simply multi-task training over all the datasets, MFT only learns from typical instances of various domains to acquire highly transferable knowledge. It further encourages the language model to encode domaininvariant representations by optimizing a series of novel domain corruption loss functions. After MFT, the model can be fine-tuned for each domain with better parameter initialization and higher generalization ability. We implement MFT upon BERT to solve several multi-domain text mining tasks. Experimental results confirm the effectiveness of MFT and its usefulness for few-shot learning. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- SAIL: Self-Augmented Graph Contrastive LearningLu Yu, Shichao Pei, Lizhong Ding, Jun Zhou 等AAAI 2022 · 被引用 47 次
- HRKD: Hierarchical Relational Knowledge Distillation for Cross-domain Language Model CompressionChenhe Dong, Yaliang Li, Ying Shen, Minghui QiuEMNLP 2021 · 被引用 6 次
- Prompt-based Distribution Alignment for Domain Generalization in Text ClassificationChen Jia, Yue ZhangEMNLP 2022 · 被引用 4 次
- Meta Distant Transfer Learning for Pre-trained Language ModelsChengyu Wang, Haojie Pan, Minghui Qiu, Jun Huang 等EMNLP 2021 · 被引用 3 次
- Learning Knowledge-Enhanced Contextual Language Representations for Domain Natural Language UnderstandingTaolin Zhang, Ruyao Xu, Chengyu Wang, Zhongjie Duan 等EMNLP 2023 · 被引用 1 次
它引用的顶会 Paper4
- ALBERT: A Lite BERT for Self-supervised Learning of Language RepresentationsZhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel 等ICLR 2020 · 被引用 7,418 次
- StructBERT: Incorporating Language Structures into Pre-training for Deep Language UnderstandingWei Wang, Bin Bi, Ming Yan, Chen Wu 等ICLR 2020 · 被引用 297 次
- Automated Relational Meta-learningHuaxiu Yao, Xian Wu, Zhiqiang Tao, Yaliang Li 等ICLR 2020 · 被引用 102 次
- KEML: A Knowledge-Enriched Meta-Learning Framework for Lexical Relation ClassificationChengyu Wang, Minghui Qiu, Jun Huang, Xiaofeng HeAAAI 2021 · 被引用 16 次
相关 Paper
- Learning to Initialize: Can Meta Learning Improve Cross-task Generalization in Prompt Tuning?Chengwei Qin, Shafiq R. Joty, Qian Li, Ruochen ZhaoACL 2023 · 被引用 8 次
- TransPrompt: Towards an Automatic Transferable Prompting Framework for Few-shot Text ClassificationChengyu Wang, Jianing Wang, Minghui Qiu, Jun Huang 等EMNLP 2021 · 被引用 39 次
- Self-Supervised Meta-Learning for Few-Shot Natural Language Classification TasksTrapit Bansal, Rishikesh Jha, Tsendsuren Munkhdalai, Andrew McCallumEMNLP 2020 · 被引用 9 次
- Meta-Learning for Fast Cross-Lingual Adaptation in Dependency ParsingAnna Langedijk, Verna Dankers, Phillip Lippe, Sander Bos 等ACL 2022 · 被引用 17 次
- Parameter-efficient Multi-task Fine-tuning for Transformers via Shared HypernetworksRabeeh Karimi Mahabadi, Sebastian Ruder, Mostafa Dehghani, James HendersonACL 2021
