Class Normalization for (Continual)? Generalized Zero-Shot Learning
Ivan Skorokhodov, Mohamed Elhoseiny
Abstract
Normalization techniques have proved to be a crucial ingredient of successful training in a traditional supervised learning regime. However, in the zero-shot learning (ZSL) world, these ideas have received only marginal attention. This work studies normalization in ZSL scenario from both theoretical and practical perspectives. First, we give a theoretical explanation to two popular tricks used in zero-shot learning: normalize+scale and attributes normalization and show that they help training by preserving variance during a forward pass. Next, we demonstrate that they are insufficient to normalize a deep ZSL model and propose Class Normalization (CN): a normalization scheme, which alleviates this issue both provably and in practice. Third, we show that ZSL models typically have more irregular loss surface compared to traditional classifiers and that the proposed method partially remedies this problem. Then, we test our approach on 4 standard ZSL datasets and outperform sophisticated modern SotA with a simple MLP optimized without any bells and whistles and having ≈50 times faster training speed. Finally, we generalize ZSL to a broader problem -continual ZSL, and introduce some principled metrics and rigorous baselines for this new setup. The project page is located at https://universome.github.io/class-norm .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers11
- En-Compactness: Self-Distillation Embedding & Contrastive Generation for Generalized Zero-Shot LearningXia Kong, Zuodong Gao, Xiaofan Li, Ming Hong et al.CVPR 2022 · 70 citations
- Distinguishing Unseen from Seen for Generalized Zero-shot LearningHongzu Su, Jingjing Li, Zhi Chen, Lei Zhu et al.CVPR 2022 · 40 citations
- Evolving Semantic Prototype Improves Generative Zero-Shot LearningShiming Chen, Wenjin Hou, Ziming Hong, Xiaohan Ding et al.ICML 2023 · 33 citations
- Generalized Lightness Adaptation with Channel Selective NormalizationMingde Yao, Jie Huang, Xin Jin, Ruikang Xu et al.ICCV 2023 · 22 citations
- Unseen Classes at a Later Time? No ProblemHari Chandana Kuchibhotla, Sumitra S. Malagi, Shivam Chandhok, Vineeth N. BalasubramanianCVPR 2022 · 9 citations
Builds on12
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Rethinking Zero-Shot Learning: A Conditional Visual Classification PerspectiveKai Li, Martin Renqiang Min, Yun FuICCV 2019 · 151 citations
- Provable Benefit of Orthogonal Initialization in Optimizing Deep Linear NetworksWei Hu, Lechao Xiao, Jeffrey PenningtonICLR 2020 · 136 citations
- Meta-Learning for Generalized Zero-Shot LearningVinay Kumar Verma, Dhanajit Brahma, Piyush RaiAAAI 2020 · 112 citations
- Principled Weight Initialization for HypernetworksOscar Chang, Lampros Flokas, Hod LipsonICLR 2020 · 87 citations
Related papers
- Rethinking Zero-Shot Video Classification: End-to-End Training for Realistic ApplicationsBiagio Brattoli, Joseph Tighe, Fedor Zhdanov, Pietro Perona et al.CVPR 2020
- Z-Score Normalization, Hubness, and Few-Shot LearningNanyi Fei, Yizhao Gao, Zhiwu Lu, Tao XiangICCV 2021 · 158 citations
- Hierarchical Normalization for Robust Monocular Depth EstimationChi Zhang, Wei Yin, Billzb Wang, Gang Yu et al.NeurIPS 2022 · 73 citations
- Semantic Feature Extraction for Generalized Zero-Shot LearningJunhan Kim, Kyuhong Shim, Byonghyo ShimAAAI 2022 · 46 citations
- Deconstructed Generation-Based Zero-Shot ModelDubing Chen, Yuming Shen, Haofeng Zhang, Philip H. S. TorrAAAI 2023 · 8 citations
