Prototypical Fine-Tuning: Towards Robust Performance under Varying Data Sizes
Yiqiao Jin, Xiting Wang, Yaru Hao, Yizhou Sun, Xing Xie
Abstract
In this paper, we move towards combining large parametric models with non-parametric prototypical networks. We propose prototypical fine-tuning, a novel prototypical framework for fine-tuning pretrained language models (LM), which automatically learns a bias to improve predictive performance for varying data sizes, especially low-resource settings. Our prototypical fine-tuning approach can automatically adjust the model capacity according to the number of data points and the model's inherent attributes. Moreover, we propose four principles for effective prototype fine-tuning towards the optimal solution. Experimental results across various datasets show that our work achieves significant performance improvements under various low-resource settings, as well as comparable and usually better performances in high-resource scenarios.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3bd66824-ef5b-4468-bc5c-423a0e568d87Cited by top-tier papers6
- Better to Ask in English: Cross-Lingual Evaluation of Large Language Models for Healthcare QueriesYiqiao Jin, Mohit Chandra, Gaurav Verma, Yibo Hu et al.WWW 2024 · 126 citations
- Continual Learning on Dynamic Graphs via Parameter IsolationPeiyan Zhang, Yuchen Yan, Chaozhuo Li, Senzhang Wang et al.SIGIR 2023 · 45 citations
- TMac: Temporal Multi-Modal Graph Learning for Acoustic Event ClassificationMeng Liu, Ke Liang, Dayu Hu, Hao Yu et al.ACM MM 2023 · 34 citations
- Predicting Information Pathways Across Online CommunitiesYiqiao Jin, Yeon-Chang Lee, Kartik Sharma, Meng Ye et al.KDD 2023 · 18 citations
- ProtoTS: Learning Hierarchical Prototypes for Explainable Time Series ForecastingZiheng Peng, Shijie Ren, Xinyue Gu, Linxiao Yang et al.ICLR 2026 · 2 citations
Builds on10
- Multimodal Few-Shot Learning with Frozen Language ModelsMaria Tsimpoukelli, Jacob Menick, Serkan Cabi, S. M. Ali Eslami et al.NeurIPS 2021 · 1,020 citations
- Supervised Contrastive Learning for Pre-trained Language Model Fine-tuningBeliz Gunel, Jingfei Du, Alexis Conneau, Veselin StoyanovICLR 2021 · 595 citations
- True Few-Shot Learning with Language ModelsEthan Perez, Douwe Kiela, Kyunghyun ChoNeurIPS 2021 · 547 citations
- Why Do Pretrained Language Models Help in Downstream Tasks? An Analysis of Head and Prompt TuningColin Wei, Sang Michael Xie, Tengyu MaNeurIPS 2021 · 119 citations
- Towards Fine-Grained Reasoning for Fake News DetectionYiqiao Jin, Xiting Wang, Ruichao Yang, Yizhou Sun et al.AAAI 2022 · 89 citations
Related papers
- Breaking Physical and Linguistic Borders: Multilingual Federated Prompt Tuning for Low-Resource LanguagesWanru Zhao, Yihong Chen, Royson Lee, Xinchi Qiu et al.ICLR 2024 · 21 citations
- Memorisation versus Generalisation in Pre-trained Language ModelsMichael Tänzer, Sebastian Ruder, Marek ReiACL 2022 · 59 citations
- Prototype-based HyperAdapter for Sample-Efficient Multi-task TuningHao Zhao, Jie Fu, Zhaofeng HeEMNLP 2023 · 3 citations
- When Scaling Meets LLM Finetuning: The Effect of Data, Model and Finetuning MethodBiao Zhang, Zhongtao Liu, Colin Cherry, Orhan FiratICLR 2024 · 271 citations
- Meta-learning via Language Model In-context TuningYanda Chen, Ruiqi Zhong, Sheng Zha, George Karypis et al.ACL 2022
