ECLARE: Extreme Classification with Label Graph Correlations
Anshul Mittal, Noveen Sachdeva, Sheshansh Agrawal, Sumeet Agarwal, Purushottam Kar, Manik Varma
Abstract
Deep extreme classification (XC) seeks to train deep architectures that can tag a data point with its most relevant subset of labels from an extremely large label set. The core utility of XC comes from predicting labels that are rarely seen during training. Such rare labels hold the key to personalized recommendations that can delight and surprise a user. However, the large number of rare labels and small amount of training data per rare label offer significant statistical and computational challenges. State-of-the-art deep XC methods attempt to remedy this by incorporating textual descriptions of labels but do not adequately address the problem. This paper presents ECLARE, a scalable deep learning architecture that incorporates not only label text, but also label correlations, to offer accurate real-time predictions within a few milliseconds. Core contributions of ECLARE include a frugal architecture and scalable techniques to train deep models along with label correlation graphs at the scale of millions of labels. In particular, ECLARE offers predictions that are 2-14% more accurate on both publicly available benchmark datasets as well as proprietary datasets for a related products recommendation task sourced from the Bing search engine. Code for ECLARE is available at https://github.com/Extreme-classification/ECLARE CCS CONCEPTS • Computing methodologies → Machine learning; Supervised learning by classification.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers24
- Fast Multi-Resolution Transformer Fine-tuning for Extreme Multi-label Text ClassificationJiong Zhang, Wei-Cheng Chang, Hsiang-Fu Yu, Inderjit S. DhillonNeurIPS 2021 · 147 citations
- SiameseXML: Siamese Networks meet Extreme Classifiers with 100M LabelsKunal Dahiya, Ananye Agarwal, Deepak Saini, Gururaj K et al.ICML 2021 · 61 citations
- Infinite Recommendation Networks: A Data-Centric ApproachNoveen Sachdeva, Mehak Preet Dhaliwal, Carole-Jean Wu, Julian J. McAuleyNeurIPS 2022 · 37 citations
- Metadata-Induced Contrastive Learning for Zero-Shot Multi-Label Text ClassificationYu Zhang, Zhihong Shen, Chieh-Han Wu, Boya Xie et al.WWW 2022 · 34 citations
- The Effect of Metadata on Scientific Literature Tagging: A Cross-Field Cross-Model StudyYu Zhang, Bowen Jin, Qi Zhu, Yu Meng et al.WWW 2023 · 27 citations
Builds on2
- LightGCN: Simplifying and Powering Graph Convolution Network for RecommendationXiangnan He, Kuan Deng, Xiang Wang, Yan Li et al.SIGIR 2020 · 4,448 citations
- Multi-Label Patent Categorization with Non-Local Attention-Based Graph Convolutional NetworkPingjie Tang, Meng Jiang, Bryan (Ning) Xia, Jed W. Pitera et al.AAAI 2020 · 52 citations
Related papers
- Deep Encoders with Auxiliary Parameters for Extreme ClassificationKunal Dahiya, Sachin Yadav, Sushant Sondhi, Deepak Saini et al.KDD 2023 · 6 citations
- Correlation Networks for Extreme Multi-label Text ClassificationGuangxu Xun, Kishlay Jha, Jianhui Sun, Aidong ZhangKDD 2020 · 60 citations
- GalaXC: Graph Neural Networks with Labelwise Attention for Extreme ClassificationDeepak Saini, Arnav Kumar Jain, Kushal Dave, Jian Jiao et al.WWW 2021 · 49 citations
- ELIAS: End-to-End Learning to Index and Search in Large Output SpacesNilesh Gupta, Patrick H. Chen, Hsiang-Fu Yu, Cho-Jui Hsieh et al.NeurIPS 2022 · 19 citations
- ENMC: Extreme Near-Memory Classification via Approximate ScreeningLiu Liu, Jilan Lin, Zheng Qu, Yufei Ding et al.MICRO 2021 · 13 citations
