SemSup-XC: Semantic Supervision for Zero and Few-shot Extreme Classification
Pranjal Aggarwal, Ameet Deshpande, Karthik R. Narasimhan
摘要
Extreme classification (XC) involves predicting over large numbers of classes (thousands to millions), with real-world applications like news article classification and e-commerce product tagging. The zero-shot version of this task requires generalization to novel classes without additional supervision. In this paper, we develop SemSup-XC, a model that achieves state-of-the-art zero-shot and few-shot performance on three XC datasets derived from legal, e-commerce, and Wikipedia data. To develop SemSup-XC, we use automatically collected semantic class descriptions to represent classes and facilitate generalization through a novel hybrid matching module that matches input instances to class descriptions using a combination of semantic and lexical similarity. Trained with contrastive learning, SemSup-XC significantly outperforms baselines and establishes state-of-the-art performance on all three datasets considered, gaining up to 12 precision points on zero-shot and more than 10 precision points on one-shot tests, with similar gains for recall@10. Our ablation studies highlight the relative importance of our hybrid matching module and automatically collected class descriptions.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Enabling Few-Shot Learning with PID Control: A Layer Adaptive OptimizerLe Yu, Xinde Li, Pengfei Zhang, Zhentong Zhang 等ICML 2024 · 被引用 3 次
- On the Necessity of World Knowledge for Mitigating Missing Labels in Extreme ClassificationJatin Prakash, Anirudh Buvanesh, Bishal Santra, Deepak Saini 等KDD 2025 · 被引用 1 次
- Extreme Meta-Classification for Large-Scale Zero-Shot RetrievalSachin Yadav, Deepak Saini, Anirudh Buvanesh, Bhawna Paliwal 等KDD 2024 · 被引用 1 次
- Incubating Text Classifiers Following User Instruction with Nothing but LLMLetian Peng, Zilong Wang, Jingbo ShangEMNLP 2024
它引用的顶会 Paper5
- Deduplicating Training Data Makes Language Models BetterKatherine Lee, Daphne Ippolito, Andrew Nystrom, Chiyuan Zhang 等ACL 2022 · 被引用 844 次
- Fast Multi-Resolution Transformer Fine-tuning for Extreme Multi-label Text ClassificationJiong Zhang, Wei-Cheng Chang, Hsiang-Fu Yu, Inderjit S. DhillonNeurIPS 2021 · 被引用 147 次
- SiameseXML: Siamese Networks meet Extreme Classifiers with 100M LabelsKunal Dahiya, Ananye Agarwal, Deepak Saini, Gururaj K 等ICML 2021 · 被引用 61 次
- Metadata-Induced Contrastive Learning for Zero-Shot Multi-Label Text ClassificationYu Zhang, Zhihong Shen, Chieh-Han Wu, Boya Xie 等WWW 2022 · 被引用 34 次
- Generalized Zero-Shot Extreme Multi-label LearningNilesh Gupta, Sakina Bohra, Yashoteja Prabhu, Saurabh Purohit 等KDD 2021 · 被引用 23 次
相关 Paper
- PESCO: Prompt-enhanced Self Contrastive Learning for Zero-shot Text ClassificationYau-Shian Wang, Ta-Chung Chi, Ruohong Zhang, Yiming YangACL 2023 · 被引用 17 次
- Transferable Contrastive Network for Generalized Zero-Shot LearningHuajie Jiang, Ruiping Wang, Shiguang Shan, Xilin ChenICCV 2019 · 被引用 200 次
- Learning with Fantasy: Semantic-Aware Virtual Contrastive Constraint for Few-Shot Class-Incremental LearningZeyin Song, Yifan Zhao, Yujun Shi, Peixi Peng 等CVPR 2023
- Semantic matching for text classification with complex class descriptionsBrian de Silva, Kuan-Wen Huang, Gwang Lee, Karen Hovsepian 等EMNLP 2023 · 被引用 1 次
- Contrastive Embedding for Generalized Zero-Shot LearningZongyan Han, Zhenyong Fu, Shuo Chen, Jian YangCVPR 2021
