Generalized Category Discovery
Sagar Vaze, Kai Han, Andrea Vedaldi, Andrew Zisserman
Abstract
In this paper, we consider a highly general image recognition setting wherein, given a labelled and unlabelled set of images, the task is to categorize all images in the unlabelled set. Here, the unlabelled images may come from labelled classes or from novel ones. Existing recognition methods are not able to deal with this setting, because they make several restrictive assumptions, such as the unlabelled instances only coming from known – or unknown – classes, and the number of unknown classes being known a-priori. We address the more unconstrained setting, naming it ‘Generalized Category Discovery’, and challenge all these assumptions. We first establish strong baselines by taking state-of-the-art algorithms from novel category discovery and adapting them for this task. Next, we propose the use of vision transformers with contrastive representation learning for this open-world setting. We then introduce a simple yet effective semi-supervised k-means method to cluster the unlabelled data into seen and unseen classes automatically, substantially outperforming the baselines. Finally, we also propose a new approach to estimate the number of classes in the unlabelled data. We thoroughly evaluate our approach on public datasets for generic object classification and on fine-grained datasets, leveraging the recent Semantic Shift Benchmark suite. Code: https://www.robots.ox.ac.uk/ vgg/research/gcd
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 235103ca-410a-4742-8fe9-aee85a5c4b7dCited by top-tier papers120
- Parametric Classification for Generalized Category Discovery: A Baseline StudyXin Wen, Bingchen Zhao, Xiaojuan QiICCV 2023 · 152 citations
- Learning Semi-supervised Gaussian Mixture Models for Generalized Category DiscoveryBingchen Zhao, Xin Wen, Kai HanICCV 2023 · 109 citations
- No Representation Rules Them All in Category DiscoverySagar Vaze, Andrea Vedaldi, Andrew ZissermanNeurIPS 2023 · 79 citations
- Generalized Category Discovery with Decoupled Prototypical NetworkWenbin An, Feng Tian, Qinghua Zheng, Wei Ding et al.AAAI 2023 · 68 citations
- BioCLIP 2: Emergent Properties from Scaling Hierarchical Contrastive LearningJianyang Gu, Sam Stevens, Elizabeth G. Campolongo, Matthew J. Thompson et al.NeurIPS 2025 · 60 citations
Builds on17
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- Supervised Contrastive LearningPrannay Khosla, Piotr Teterwak, Chen Wang, Aaron Sarna et al.NeurIPS 2020 · 7,049 citations
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal et al.NeurIPS 2020 · 5,249 citations
- FixMatch: Simplifying Semi-Supervised Learning with Consistency and ConfidenceKihyuk Sohn, David Berthelot, Nicholas Carlini, Zizhao Zhang et al.NeurIPS 2020 · 5,129 citations
Related papers
- PromptCAL: Contrastive Affinity Learning via Auxiliary Prompts for Generalized Novel Category DiscoverySheng Zhang, Salman H. Khan, Zhiqiang Shen, Muzammal Naseer et al.CVPR 2023
- Dynamic Conceptional Contrastive Learning for Generalized Category DiscoveryNan Pu, Zhun Zhong, Nicu SebeCVPR 2023
- Collaborative Cloud-edge Generalized Category DiscoveryYingbing Liu, Fei Ma, Yanan Wu, Xinxin Zuo et al.ACM MM 2025
- Contrastive Mean-Shift Learning for Generalized Category DiscoverySua Choi, Dahyun Kang, Minsu ChoCVPR 2024
- ALLGCD: Leveraging All Unlabeled Data for Generalized Category DiscoveryXinzi Cao, Ke Chen, Feidiao Yang, Xiawu Zheng et al.ICCV 2025 · 2 citations
