GENESIS-V2: Inferring Unordered Object Representations without Iterative Refinement
Martin Engelcke, Oiwi Parker Jones, Ingmar Posner
摘要
Advances in unsupervised learning of object-representations have culminated in the development of a broad range of methods for unsupervised object segmentation and interpretable object-centric scene generation. These methods, however, are limited to simulated and real-world datasets with limited visual complexity. Moreover, object representations are often inferred using RNNs which do not scale well to large images or iterative refinement which avoids imposing an unnatural ordering on objects in an image but requires the a priori initialisation of a fixed number of object representations. In contrast to established paradigms, this work proposes an embedding-based approach in which embeddings of pixels are clustered in a differentiable fashion using a stochastic stick-breaking process. Similar to iterative refinement, this clustering procedure also leads to randomly ordered object representations, but without the need of initialising a fixed number of clusters a priori. This is used to develop a new model, GENESIS-v2, which can infer a variable number of object representations without using RNNs or iterative refinement. We show that GENESIS-v2 performs strongly in comparison to recent baselines in terms of unsupervised image segmentation and object-centric scene generation on established synthetic datasets as well as more complex real-world datasets.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper41
- Object-Centric Slot DiffusionJindong Jiang, Fei Deng, Gautam Singh, Sungjin AhnNeurIPS 2023 · 被引用 106 次
- SlotDiffusion: Object-Centric Generative Modeling with Diffusion ModelsZiyi Wu, Jingyu Hu, Wuyue Lu, Igor Gilitschenski 等NeurIPS 2023 · 被引用 106 次
- Object-Centric Learning for Real-World Videos by Predicting Temporal Feature SimilaritiesAndrii Zadaianchuk, Maximilian Seitzer, Georg MartiusNeurIPS 2023 · 被引用 104 次
- An Investigation into Pre-Training Object-Centric Representations for Reinforcement LearningJaesik Yoon, Yi-Fu Wu, Heechul Bae, Sungjin AhnICML 2023 · 被引用 59 次
- Provably Learning Object-Centric RepresentationsJack Brady, Roland S. Zimmermann, Yash Sharma, Bernhard Schölkopf 等ICML 2023 · 被引用 55 次
它引用的顶会 Paper11
- Object-Centric Learning with Slot AttentionFrancesco Locatello, Dirk Weissenborn, Thomas Unterthiner, Aravindh Mahendran 等NeurIPS 2020 · 被引用 1,275 次
- NVAE: A Deep Hierarchical Variational AutoencoderArash Vahdat, Jan KautzNeurIPS 2020 · 被引用 1,141 次
- GENESIS: Generative Scene Inference and Sampling with Object-Centric Latent RepresentationsMartin Engelcke, Adam R. Kosiorek, Oiwi Parker Jones, Ingmar PosnerICLR 2020 · 被引用 334 次
- SPACE: Unsupervised Object-Oriented Scene Representation via Spatial Attention and DecompositionZhixuan Lin, Yi-Fu Wu, Skand Vishwanath Peri, Weihao Sun 等ICLR 2020 · 被引用 276 次
- SCALOR: Generative World Models with Scalable Object RepresentationsJindong Jiang, Sepehr Janghorbani, Gerard de Melo, Sungjin AhnICLR 2020 · 被引用 152 次
相关 Paper
- Robust and Controllable Object-Centric Learning through Energy-based ModelsRuixiang Zhang, Tong Che, Boris Ivanovic, Renhao Wang 等ICLR 2023
- Unsupervised Causal Generative Understanding of ImagesTitas Anciukevicius, Patrick Fox-Roberts, Edward Rosten, Paul HendersonNeurIPS 2022 · 被引用 6 次
- SegSort: Segmentation by Discriminative Sorting of SegmentsJyh-Jing Hwang, Stella X. Yu, Jianbo Shi, Maxwell D. Collins 等ICCV 2019 · 被引用 160 次
- Learning the Superpixel in a Non-Iterative and Lifelong MannerLei Zhu, Qi She, Bin Zhang, Yanye Lu 等CVPR 2021
- Bridging the Gap to Real-World Object-Centric LearningMaximilian Seitzer, Max Horn, Andrii Zadaianchuk, Dominik Zietlow 等ICLR 2023 · 被引用 31 次
