The Devil Is in Gradient Entanglement: Energy-Aware Gradient Coordinator for Robust Generalized Category Discovery
Haiyang Zheng, Nan Pu, Yaqi Cai, Teng Long, Wenjing Li, Nicu Sebe, Zhun Zhong
摘要
Generalized Category Discovery (GCD) aims to categorize unlabeled samples that may belong to either known or unknown categories by leveraging the knowledge from labeled data. Most previous methods jointly optimize supervised and unsupervised objectives and achieve promising results. However, inherent optimization interference still limits their ability to improve further. Through quantitative analysis, we identify a key issue, i.e. , gradient entanglement , which 1) distorts supervised gradients and weakens discrimination among known classes, and 2) induces representation-subspace overlap between known and novel classes, reducing the separability of novel categories. To address this issue, we propose the Energy-Aware Gradient Coordinator (EAGC), a plug-and-play gradient-level module that explicitly regulates the optimization process. EAGC comprises two components: Anchor-based Gradient Alignment (AGA) and Energy-aware Elastic Projection (EEP). AGA introduces a reference model to anchor the gradient directions of labeled samples, preserving the discriminative structure of known classes against the interference of unlabeled gradients. EEP softly projects unlabeled gradients onto the complement of the known-class subspace and derives an energy-based coefficient to adaptively scale the projection for each unlabeled sample according to its degree of alignment with the known subspace, thereby reducing subspace overlap without suppressing unlabeled samples that likely belong to known classes. EAGC can be seamlessly integrated with both parametric and non-parametric GCD methods. Experiments show that EAGC consistently boosts existing approaches and establishes new state-of-the-art results on multiple GCD benchmarks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper55
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
- Gradient Surgery for Multi-Task LearningTianhe Yu, Saurabh Kumar, Abhishek Gupta, Sergey Levine 等NeurIPS 2020 · 被引用 2,261 次
- Open-Set Recognition: A Good Closed-Set Classifier is All You NeedSagar Vaze, Kai Han, Andrea Vedaldi, Andrew ZissermanICLR 2022 · 被引用 594 次
- Gradient Projection Memory for Continual LearningGobinda Saha, Isha Garg, Kaushik RoyICLR 2021 · 被引用 409 次
- Learning to Discover Novel Visual Categories via Deep Transfer ClusteringKai Han, Andrea Vedaldi, Andrew ZissermanICCV 2019 · 被引用 378 次
相关 Paper
- Consistent Supervised-Unsupervised Alignment for Generalized Category DiscoveryJizhou Han, Shaokun Wang, Yuhang He, Chenhao Ding 等NeurIPS 2025 · 被引用 7 次
- Parametric Classification for Generalized Category Discovery: A Baseline StudyXin Wen, Bingchen Zhao, Xiaojuan QiICCV 2023 · 被引用 152 次
- CURE: Consistency-under-Unified Semantic Regularization for Generalized Category DiscoveryYuwei Bian, Shidong Wang, Haofeng ZhangICML 2026
- Collaborative Cloud-edge Generalized Category DiscoveryYingbing Liu, Fei Ma, Yanan Wu, Xinxin Zuo 等ACM MM 2025
- Prior-Constrained Association Learning for Fine-Grained Generalized Category DiscoveryMenglin Wang, Zhun Zhong, Xiaojin GongAAAI 2025 · 被引用 4 次
