Fast Multi-Resolution Transformer Fine-tuning for Extreme Multi-label Text Classification
Jiong Zhang, Wei-Cheng Chang, Hsiang-Fu Yu, Inderjit S. Dhillon
摘要
Extreme multi-label text classification (XMC) seeks to find relevant labels from an extreme large label collection for a given text input. Many real-world applications can be formulated as XMC problems, such as recommendation systems, document tagging and semantic search. Recently, transformer based XMC methods, such as X-Transformer and LightXML, have shown significant improvement over other XMC methods. Despite leveraging pre-trained transformer models for text representation, the fine-tuning procedure of transformer models on large label space still has lengthy computational time even with powerful GPUs. In this paper, we propose a novel recursive approach, XR-Transformer to accelerate the procedure through recursively fine-tuning transformer models on a series of multi-resolution objectives related to the original XMC objective function. Empirical results show that XR-Transformer takes significantly less training time compared to other transformerbased XMC models while yielding better state-of-the-art results. In particular, on the public Amazon-3M dataset with 3 million labels, XR-Transformer is not only 20x faster than X-Transformer but also improves the Precision@1 from 51% to 54%. Our code is publicly available at https://github.com/amzn/pecos .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper25
- Node Feature Extraction by Self-Supervised Multi-scale Neighborhood PredictionEli Chien, Wei-Cheng Chang, Cho-Jui Hsieh, Hsiang-Fu Yu 等ICLR 2022 · 被引用 185 次
- Nuances are the Key: Unlocking ChatGPT to Find Failure-Inducing Tests with Differential PromptingTsz On Li, Wenxi Zong, Yibo Wang, Haoye Tian 等ASE 2023 · 被引用 51 次
- The Effect of Metadata on Scientific Literature Tagging: A Cross-Field Cross-Model StudyYu Zhang, Bowen Jin, Qi Zhu, Yu Meng 等WWW 2023 · 被引用 27 次
- CascadeXML: Rethinking Transformers for End-to-end Multi-resolution Training in Extreme Multi-label ClassificationSiddhant Kharbanda, Atmadeep Banerjee, Erik Schultheis, Rohit BabbarNeurIPS 2022 · 被引用 26 次
- Large Language Model Meets Graph Neural Network in Knowledge DistillationShengxiang Hu, Guobing Zou, Song Yang, Shiyi Lin 等AAAI 2025 · 被引用 19 次
它引用的顶会 Paper6
- Accelerating Large-Scale Inference with Anisotropic Vector QuantizationRuiqi Guo, Philip Sun, Erik Lindgren, Quan Geng 等ICML 2020 · 被引用 539 次
- Pre-training Tasks for Embedding-based Large-scale RetrievalWei-Cheng Chang, Felix X. Yu, Yin-Wen Chang, Yiming Yang 等ICLR 2020 · 被引用 325 次
- LightXML: Transformer with Dynamic Negative Sampling for High-Performance Extreme Multi-label Text ClassificationTing Jiang, Deqing Wang, Leilei Sun, Huayi Yang 等AAAI 2021 · 被引用 170 次
- ECLARE: Extreme Classification with Label Graph CorrelationsAnshul Mittal, Noveen Sachdeva, Sheshansh Agrawal, Sumeet Agarwal 等WWW 2021 · 被引用 71 次
- Pretrained Generalized Autoregressive Model with Adaptive Probabilistic Label Clusters for Extreme Multi-label Text ClassificationHui Ye, Zhiyu Chen, Da-Han Wang, Brian D. DavisonICML 2020 · 被引用 57 次
相关 Paper
- PINA: Leveraging Side Information in eXtreme Multi-label Classification via Predicted Instance Neighborhood AggregationEli Chien, Jiong Zhang, Cho-Jui Hsieh, Jyun-Yu Jiang 等ICML 2023 · 被引用 11 次
- ELIAS: End-to-End Learning to Index and Search in Large Output SpacesNilesh Gupta, Patrick H. Chen, Hsiang-Fu Yu, Cho-Jui Hsieh 等NeurIPS 2022 · 被引用 19 次
- Follow the Path: Hierarchy-Aware Extreme Multi-Label Completion for Semantic Text TaggingNatalia Ostapuk, Julien Audiffren, Ljiljana Dolamic, Alain Mermoud 等WWW 2024 · 被引用 5 次
- Correlation Networks for Extreme Multi-label Text ClassificationGuangxu Xun, Kishlay Jha, Jianhui Sun, Aidong ZhangKDD 2020 · 被引用 60 次
- Gandalf: Learning Label-label Correlations in Extreme Multi-label Classification via Label FeaturesSiddhant Kharbanda, Devaansh Gupta, Erik Schultheis, Atmadeep Banerjee 等KDD 2024 · 被引用 6 次
