HASSOD: Hierarchical Adaptive Self-Supervised Object Detection
Shengcao Cao, Dhiraj Joshi, Liangyan Gui, Yu-Xiong Wang
摘要
The human visual perception system demonstrates exceptional capabilities in learning without explicit supervision and understanding the part-to-whole composition of objects. Drawing inspiration from these two abilities, we propose Hierarchical Adaptive Self-Supervised Object Detection (HASSOD), a novel approach that learns to detect objects and understand their compositions without human supervision. HASSOD employs a hierarchical adaptive clustering strategy to group regions into object masks based on self-supervised visual representations, adaptively determining the number of objects per image. Furthermore, HASSOD identifies the hierarchical levels of objects in terms of composition, by analyzing coverage relations between masks and constructing tree structures. This additional self-supervised learning task leads to improved detection performance and enhanced interpretability. Lastly, we abandon the inefficient multi-round self-training process utilized in prior methods and instead adapt the Mean Teacher framework from semi-supervised learning, which leads to a smoother and more efficient training process. Through extensive experiments on prevalent image datasets, we demonstrate the superiority of HASSOD over existing methods, thereby advancing the state of the art in self-supervised object detection. Notably, we improve Mask AR from 20.2 to 22.5 on LVIS, and from 17.0 to 26.0 on SA-1B. Project page: https://HASSOD-NeurIPS23.github.io.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper6
- Segment Anything without SupervisionXudong Wang, Jingfeng Yang, Trevor DarrellNeurIPS 2024 · 被引用 36 次
- DetKDS: Knowledge Distillation Search for Object DetectorsLujun Li, Yufan Bao, Peijie Dong, Chuanguang Yang 等ICML 2024 · 被引用 35 次
- DiPEx: Dispersing Prompt Expansion for Class-Agnostic Object DetectionJia Syuen Lim, Zhuoxiao Chen, Zhi Chen, Mahsa Baktashmotlagh 等NeurIPS 2024 · 被引用 19 次
- SOHES: Self-supervised Open-world Hierarchical Entity SegmentationShengcao Cao, Jiuxiang Gu, Jason Kuen, Hao Tan 等ICLR 2024 · 被引用 3 次
- Scene-Centric Unsupervised Panoptic SegmentationOliver Hahn, Christoph Reich, Nikita Araslanov, Daniel Cremers 等CVPR 2025
它引用的顶会 Paper10
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao 等ICCV 2023 · 被引用 13,211 次
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou 等ICCV 2021 · 被引用 8,921 次
- Objects365: A Large-Scale, High-Quality Dataset for Object DetectionShuai Shao, Zeming Li, Tianyuan Zhang, Chao Peng 等ICCV 2019 · 被引用 1,018 次
- Unbiased Teacher for Semi-Supervised Object DetectionYen-Cheng Liu, Chih-Yao Ma, Zijian He, Chia-Wen Kuo 等ICLR 2021 · 被引用 603 次
- Large-Scale Unsupervised Object DiscoveryHuy V. Vo, Elena Sizikova, Cordelia Schmid, Patrick Pérez 等NeurIPS 2021 · 被引用 63 次
相关 Paper
- Self-Supervised Object Detection from Egocentric VideosPeri Akiva, Jing Huang, Kevin J. Liang, Rama Kovvuri 等ICCV 2023 · 被引用 9 次
- Unsupervised Semantic Segmentation with Self-supervised Object-centric RepresentationsAndrii Zadaianchuk, Matthäus Kleindessner, Yi Zhu, Francesco Locatello 等ICLR 2023 · 被引用 16 次
- Adaptive Hierarchical Representation Learning for Long-Tailed Object DetectionBanghuai LiCVPR 2022 · 被引用 16 次
- Self-Supervised Object Detection via Generative Image SynthesisSiva Karthik Mustikovela, Shalini De Mello, Aayush Prakash, Umar Iqbal 等ICCV 2021 · 被引用 3 次
- Open-Vocabulary Object Detection via Language HierarchyJiaxing Huang, Jingyi Zhang, Kai Jiang, Shijian LuNeurIPS 2024 · 被引用 16 次
