Hierarchical Compact Clustering Attention (COCA) for Unsupervised Object-Centric Learning
Can Küçüksözen, Yücel Yemez
2025Year
Abstract
Figure 1 . Compactness scores obtained for each pixel in the scene, across four different datasets. The transition from bright yellow to deep purple signifies decreasing compactness. To obtain these scores, a trained COCA-Net encoder is used to generate object masks. Each object mask is then broadcasted to pixels based on the pixel-object assignments. This operation associates every pixel with a copy of its object's mask. Finally, compactness scores for each pixel's mask are calculated via Eq. 3.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Builds on27
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- Emerging Properties in Self-Supervised Vision TransformersMathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou et al.ICCV 2021 · 8,921 citations
- Object-Centric Learning with Slot AttentionFrancesco Locatello, Dirk Weissenborn, Thomas Unterthiner, Aravindh Mahendran et al.NeurIPS 2020 · 1,275 citations
- Scaling Vision Transformers to 22 Billion ParametersMostafa Dehghani, Josip Djolonga, Basil Mustafa, Piotr Padlewski et al.ICML 2023 · 848 citations
- Vision GNN: An Image is Worth Graph of NodesKai Han, Yunhe Wang, Jianyuan Guo, Yehui Tang et al.NeurIPS 2022 · 668 citations
Related papers
- COAT: Measuring Object Compositionality in Emergent RepresentationsSirui Xie, Ari S. Morcos, Song-Chun Zhu, Ramakrishna VedantamICML 2022 · 10 citations
- Scene EssenceJiayan Qiu, Yiding Yang, Xinchao Wang, Dacheng TaoCVPR 2021
- The Making and Breaking of CamouflageHala Lamdouar, Weidi Xie, Andrew ZissermanICCV 2023 · 19 citations
- L-CoIns: Language-based Colorization With Instance AwarenessZheng Chang, Shuchen Weng, Peixuan Zhang, Yu Li et al.CVPR 2023
- Simultaneously Localize, Segment and Rank the Camouflaged ObjectsYunqiu Lv, Jing Zhang, Yuchao Dai, Aixuan Li et al.CVPR 2021
