Hybrid Sharing for Multi-Label Image Classification
Zihao Yin, Chen Gan, Kelei He, Yang Gao, Junfeng Zhang
摘要
Existing multi-label classification methods have long suffered from label heterogeneity, where learning a label obscures another. By modeling multi-label classification as a multi-task problem, this issue can be regarded as a negative transfer, which indicates challenges to achieve simultaneously satisfied performance across multiple tasks. In this work, we propose the Hybrid Sharing Query (HSQ), a transformer-based model that introduces the mixture-of-experts architecture to image multi-label classification. HSQ is designed to leverage label correlations while mitigating heterogeneity effectively. To this end, HSQ is incorporated with a fusion expert framework that enables it to optimally combine the strengths of task-specialized experts with shared experts, ultimately enhancing multi-label classification performance across most labels. Extensive experiments are conducted on two benchmark datasets, with the results demonstrating that the proposed method achieves state-of-the-art performance and yields simultaneous improvements across most labels. The code is available at this URL.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- MAT-Agent: Adaptive Multi-Agent Training OptimizationJusheng Zhang, Kaitong Cai, Yijia Fan, Ningyuan Liu 等NeurIPS 2025 · 被引用 46 次
- Correlative and Discriminative Label Grouping for Multi-Label Visual Prompt TuningLei-Lei Ma, Shuo Xu, Ming-Kun Xie, Lei Wang 等CVPR 2025
它引用的顶会 Paper12
- Swin Transformer: Hierarchical Vision Transformer using Shifted WindowsZe Liu, Yutong Lin, Yue Cao, Han Hu 等ICCV 2021 · 被引用 31,683 次
- A ConvNet for the 2020sZhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer 等CVPR 2022 · 被引用 6,782 次
- CvT: Introducing Convolutions to Vision TransformersHaiping Wu, Bin Xiao, Noel Codella, Mengchen Liu 等ICCV 2021 · 被引用 2,397 次
- Mixture-of-Experts with Expert Choice RoutingYanqi Zhou, Tao Lei, Hanxiao Liu, Nan Du 等NeurIPS 2022 · 被引用 933 次
- Learning Semantic-Specific Graph Representation for Multi-Label Image RecognitionTianshui Chen, Muxin Xu, Xiaolu Hui, Hefeng Wu 等ICCV 2019 · 被引用 347 次
相关 Paper
- Language-Guided Transformer for Federated Multi-Label ClassificationI-Jieh Liu, Ci-Siang Lin, Fu-En Yang, Yu-Chiang Frank WangAAAI 2024 · 被引用 16 次
- Mod-Squad: Designing Mixtures of Experts As Modular Multi-Task LearnersZitian Chen, Yikang Shen, Mingyu Ding, Zhenfang Chen 等CVPR 2023
- View-Category Interactive Sharing Transformer for Incomplete Multi-View Multi-Label LearningShilong Ou, Zhe Xue, Yawen Li, Meiyu Liang 等CVPR 2024 · 被引用 11 次
- HSVLT: Hierarchical Scale-Aware Vision-Language Transformer for Multi-Label Image ClassificationShuyi Ouyang, Hongyi Wang, Ziwei Niu, Zhenjia Bai 等ACM MM 2023 · 被引用 5 次
- General Multi-Label Image Classification With TransformersJack Lanchantin, Tianlu Wang, Vicente Ordonez, Yanjun QiCVPR 2021
