TART: Improved Few-shot Text Classification Using Task-Adaptive Reference Transformation
Shuo Lei, Xuchao Zhang, Jianfeng He, Fanglan Chen, Chang-Tien Lu
Abstract
Meta-learning has emerged as a trending technique to tackle few-shot text classification and achieve state-of-the-art performance. However, the performance of existing approaches heavily depends on the inter-class variance of the support set. As a result, it can perform well on tasks when the semantics of sampled classes are distinct while failing to differentiate classes with similar semantics. In this paper, we propose a novel Task-Adaptive Reference Transformation (TART) network, aiming to enhance the generalization by transforming the class prototypes to per-class fixed reference points in task-adaptive metric spaces. To further maximize divergence between transformed prototypes in task-adaptive metric spaces, TART introduces a discriminative reference regularization among transformed prototypes. Extensive experiments are conducted on four benchmark datasets and our method demonstrates clear superiority over the stateof-the-art models in all the datasets. In particular, our model surpasses the state-of-the-art method by 7.4% and 5.4% in 1-shot and 5-shot classification on the 20 Newsgroups dataset, respectively. Our code is available at https: //github.com/slei109/TART Class Testing Sample Task 1: Support class: 1,2,3,4 Task 2: Support class: 1,2,3,5 MLADA Ours MLADA Ours 1 Animal photos of the week: baby tiger goes for a swim. 1 1 1 1 2 Twitter helps confirm X-shaped bulge at Center of Milky Way. 4 2 2 2 3 Toronto van attack suspect's Facebook post praised misogynist mass killer. 4 3 2 3 4 Apple just solved one of the iphone's most harmful features. 2 4 --5 Apple fritter season is here, and so are the recipes you'll need.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 4aadd169-0738-4f52-aa96-81ee4b99a0d3Cited by top-tier papers1
Ask how each one uses itBuilds on5
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Few-shot Text Classification with Distributional SignaturesYujia Bao, Menghua Wu, Shiyu Chang, Regina BarzilayICLR 2020 · 183 citations
- ContrastNet: A Contrastive Learning Framework for Few-Shot Text ClassificationJunfan Chen, Richong Zhang, Yongyi Mao, Jie XuAAAI 2022 · 100 citations
- TransPrompt: Towards an Automatic Transferable Prompting Framework for Few-shot Text ClassificationChengyu Wang, Jianing Wang, Minghui Qiu, Jun Huang et al.EMNLP 2021 · 39 citations
- Making Pre-trained Language Models Better Few-shot LearnersTianyu Gao, Adam Fisch, Danqi ChenACL 2021
Related papers
- Cross-Domain Few-Shot Classification via Learned Feature-Wise TransformationHung-Yu Tseng, Hsin-Ying Lee, Jia-Bin Huang, Ming-Hsuan YangICLR 2020 · 467 citations
- Topological Transduction for Hybrid Few-shot LearningJiayi Chen, Aidong ZhangWWW 2022 · 2 citations
- XtarNet: Learning to Extract Task-Adaptive Representation for Incremental Few-Shot LearningSung Whan Yoon, Do-Yeon Kim, Jun Seo, Jaekyun MoonICML 2020 · 49 citations
- Learn to Adapt for Generalized Zero-Shot Text ClassificationYiwen Zhang, Caixia Yuan, Xiaojie Wang, Ziwei Bai et al.ACL 2022 · 21 citations
- Boosting Few-Shot Learning With Adaptive Margin LossAoxue Li, Weiran Huang, Xu Lan, Jiashi Feng et al.CVPR 2020
