(ML)2P-Encoder: On Exploration of Channel-Class Correlation for Multi-Label Zero-Shot Learning
Ziming Liu, Song Guo, Xiaocheng Lu, Jingcai Guo, Jiewei Zhang, Yue Zeng, Fushuo Huo
Abstract
Recent studies usually approach multi-label zeroshot learning (MLZSL) with visual-semantic mapping on spatial-class correlation, which can be computationally costly, and worse still, fails to capture fine-grained classspecific semantics. We observe that different channels may usually have different sensitivities on classes, which can correspond to specific semantics. Such an intrinsic channelclass correlation suggests a potential alternative for the more accurate and class-harmonious feature representations. In this paper, our interest is to fully explore the power of channel-class correlation as the unique base for MLZSL. Specifically, we propose a light yet efficient Multi-Label Multi-Layer Perceptron-based Encoder, dubbed (ML) 2 P-Encoder, to extract and preserve channel-wise semantics. We reorganize the generated feature maps into several groups, of which each of them can be trained independently with (ML) 2 P-Encoder. On top of that, a global groupwise attention module is further designed to build the multilabel specific class relationships among different classes, which eventually fulfills a novel Channel-Class Correlation MLZSL framework (C 3 -MLZSL) 1 . Extensive experiments on large-scale MLZSL benchmarks including NUS-WIDE and Open-Images-V4 demonstrate the superiority of our model against other representative state-of-the-art models.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9b410f76-adcf-4550-95d4-8ab5ca8c5f80Cited by top-tier papers3
- Epsilon: Exploring Comprehensive Visual-Semantic Projection for Multi-Label Zero-Shot LearningZiming Liu, Jingcai Guo, Song Guo, Xiaocheng LuAAAI 2025 · 6 citations
- Text as Any-Modality for Zero-Shot Classification by Consistent Prompt TuningXiangyu Wu, Feng Yu, Yang Yang, Jianfeng LuACM MM 2025
- DART: Dual Adaptive Refinement Transfer for Open-Vocabulary Multi-Label RecognitionHaijing Liu, Tao Pu, Hefeng Wu, Keze Wang et al.ACM MM 2025
Builds on5
- Discriminative Region-based Multi-Label Zero-Shot LearningSanath Narayan, Akshita Gupta, Salman H. Khan, Fahad Shahbaz Khan et al.ICCV 2021 · 62 citations
- Learning Modality-Invariant Latent Representations for Generalized Zero-shot LearningJingjing Li, Mengmeng Jing, Lei Zhu, Zhengming Ding et al.ACM MM 2020 · 35 citations
- Mitigating Generation Shifts for Generalized Zero-Shot LearningZhi Chen, Yadan Luo, Sen Wang, Ruihong Qiu et al.ACM MM 2021 · 30 citations
- Generalized Zero-Shot Learning using Generated Proxy Unseen Samples and Entropy SeparationOmkar Gune, Biplab Banerjee, Subhasis Chaudhuri, Fabio CuzzolinACM MM 2020 · 15 citations
- A Shared Multi-Attention Framework for Multi-Label Zero-Shot LearningDat Huynh, Ehsan ElhamifarCVPR 2020
Related papers
- Semantic Feature Extraction for Generalized Zero-Shot LearningJunhan Kim, Kyuhong Shim, Byonghyo ShimAAAI 2022 · 46 citations
- Rethinking Zero-Shot Learning: A Conditional Visual Classification PerspectiveKai Li, Martin Renqiang Min, Yun FuICCV 2019 · 151 citations
- Class Semantic Attribute Perception Guided Zero-Shot LearningQin Yue, Junbiao Cui, Jianqing Liang, Liang BaiAAAI 2025 · 1 citation
- Rethinking Zero-Shot Video Classification: End-to-End Training for Realistic ApplicationsBiagio Brattoli, Joseph Tighe, Fedor Zhdanov, Pietro Perona et al.CVPR 2020
- Causal Visual-semantic Correlation for Zero-shot LearningShuhuang Chen, Dingjie Fu, Shiming Chen, Shuo Ye et al.ACM MM 2024 · 11 citations
