Towards Robust Low-Resource Fine-Tuning with Multi-View Compressed Representations
Linlin Liu, Xingxuan Li, Megh Thakkar, Xin Li, Shafiq Joty, Luo Si, Lidong Bing
摘要
Due to the huge amount of parameters, finetuning of pretrained language models (PLMs) is prone to overfitting in the low resource scenarios. In this work, we present a novel method that operates on the hidden representations of a PLM to reduce overfitting. During fine-tuning, our method inserts random autoencoders between the hidden layers of a PLM, which transform activations from the previous layers into multi-view compressed representations before feeding them into the upper layers. The autoencoders are plugged out after fine-tuning, so our method does not add extra parameters or increase computation cost during inference. Our method demonstrates promising performance improvement across a wide range of sequenceand token-level low-resource NLP tasks. Our code is available at https://github.com/DAMO-NLP-SG/MVCR . * Equal contribution, order decided by coin flip. Linlin Liu and Xingxuan Li are under the Joint Ph.D. Program between
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Mining Useful General Data for Low-Resource Domain AdaptationPingjie Wang, Hongcheng Liu, Yusheng Liao, Ziqing Fan 等ICML 2026 · 被引用 3 次
- IM-BERT: Enhancing Robustness of BERT through the Implicit Euler MethodMihyeon Kim, Juhyoung Park, YoungBin KimEMNLP 2024
它引用的顶会 Paper15
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- XTREME: A Massively Multilingual Multi-task Benchmark for Evaluating Cross-lingual GeneralisationJunjie Hu, Sebastian Ruder, Aditya Siddhant, Graham Neubig 等ICML 2020 · 被引用 1,132 次
- LUKE: Deep Contextualized Entity Representations with Entity-aware Self-attentionIkuya Yamada, Akari Asai, Hiroyuki Shindo, Hideaki Takeda 等EMNLP 2020 · 被引用 562 次
- Unsupervised Cross-lingual Representation Learning at ScaleAlexis Conneau, Kartikay Khandelwal, Naman Goyal, Vishrav Chaudhary 等ACL 2020 · 被引用 539 次
- On the Effectiveness of Parameter-Efficient Fine-TuningZihao Fu, Haoran Yang, Anthony Man-Cho So, Wai Lam 等AAAI 2023 · 被引用 234 次
相关 Paper
- MixPHM: Redundancy-Aware Parameter-Efficient Tuning for Low-Resource Visual Question AnsweringJingjing Jiang, Nanning ZhengCVPR 2023
- Variational Information Bottleneck for Effective Low-Resource Fine-TuningRabeeh Karimi Mahabadi, Yonatan Belinkov, James HendersonICLR 2021 · 被引用 88 次
- On the Effectiveness of Adapter-based Tuning for Pretrained Language Model AdaptationRuidan He, Linlin Liu, Hai Ye, Qingyu Tan 等ACL 2021
- Prototypical Fine-Tuning: Towards Robust Performance under Varying Data SizesYiqiao Jin, Xiting Wang, Yaru Hao, Yizhou Sun 等AAAI 2023 · 被引用 15 次
- Enhancing Low-Resource Relation Representations through Multi-View DecouplingChenghao Fan, Wei Wei, Xiaoye Qu, Zhenyi Lu 等AAAI 2024 · 被引用 10 次
