Unsupervised Layered Image Decomposition into Object Prototypes
Tom Monnier, Elliot Vincent, Jean Ponce, Mathieu Aubry
摘要
We present an unsupervised learning framework for decomposing images into layers of automatically discovered object models. Contrary to recent approaches that model image layers with autoencoder networks, we represent them as explicit transformations of a small set of prototypical images. Our model has three main components: (i) a set of object prototypes in the form of learnable images with a transparency channel, which we refer to as sprites; (ii) differentiable parametric functions predicting occlusions and transformation parameters necessary to instantiate the sprites in a given image; (iii) a layered image formation model with occlusion for compositing these instances into complete images including background. By jointly learning the sprites and occlusion/transformation predictors to reconstruct images, our approach not only yields accurate layered image decompositions, but also identifies object categories and instance parameters. We first validate our approach by providing results on par with the state of the art on standard multiobject synthetic benchmarks (Tetrominoes, Multi-dSprites, CLEVR6). We then demonstrate the applicability of our model to real images in tasks that include clustering (SVHN, GTSRB), cosegmentation (Weizmann Horse) and object discovery from unfiltered social network images. To the best of our knowledge, our approach is the first layered image decomposition algorithm that learns an explicit and shared concept of object type, and is robust enough to be applied to real images.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper23
- Unsupervised Discovery of Object Radiance FieldsHong-Xing Yu, Leonidas J. Guibas, Jiajun WuICLR 2022 · 被引用 132 次
- Deep Spectral Methods: A Surprisingly Strong Baseline for Unsupervised Semantic Segmentation and LocalizationLuke Melas-Kyriazi, Christian Rupprecht, Iro Laina, Andrea VedaldiCVPR 2022 · 被引用 132 次
- Large-Scale Unsupervised Object DiscoveryHuy V. Vo, Elena Sizikova, Cordelia Schmid, Patrick Pérez 等NeurIPS 2021 · 被引用 63 次
- Invariant Slot Attention: Object Discovery with Slot-Centric Reference FramesOndrej Biza, Sjoerd van Steenkiste, Mehdi S. M. Sajjadi, Gamaleldin Fathy Elsayed 等ICML 2023 · 被引用 54 次
- Differentiable Blocks World: Qualitative 3D Decomposition by Rendering PrimitivesTom Monnier, Jake Austin, Angjoo Kanazawa, Alexei A. Efros 等NeurIPS 2023 · 被引用 50 次
它引用的顶会 Paper8
- Object-Centric Learning with Slot AttentionFrancesco Locatello, Dirk Weissenborn, Thomas Unterthiner, Aravindh Mahendran 等NeurIPS 2020 · 被引用 1,275 次
- Invariant Information Clustering for Unsupervised Image Classification and SegmentationXu Ji, Andrea Vedaldi, João F. HenriquesICCV 2019 · 被引用 956 次
- GENESIS: Generative Scene Inference and Sampling with Object-Centric Latent RepresentationsMartin Engelcke, Adam R. Kosiorek, Oiwi Parker Jones, Ingmar PosnerICLR 2020 · 被引用 334 次
- SPACE: Unsupervised Object-Oriented Scene Representation via Spatial Attention and DecompositionZhixuan Lin, Yi-Fu Wu, Skand Vishwanath Peri, Weihao Sun 等ICLR 2020 · 被引用 276 次
- Group-Wise Deep Object Co-Segmentation With Co-Attention Recurrent Neural NetworkBo Li, Zhengxing Sun, Qian Li, Yunjie Wu 等ICCV 2019 · 被引用 62 次
相关 Paper
- Deformable Sprites for Unsupervised Video DecompositionVickie Ye, Zhengqi Li, Richard Tucker, Angjoo Kanazawa 等CVPR 2022 · 被引用 45 次
- MarioNette: Self-Supervised Sprite LearningDmitriy Smirnov, Michaël Gharbi, Matthew Fisher, Vitor Guizilini 等NeurIPS 2021 · 被引用 47 次
- Generative Scene Graph NetworksFei Deng, Zhuo Zhi, Donghun Lee, Sungjin AhnICLR 2021 · 被引用 10 次
- unMORE: Unsupervised Multi-Object Segmentation via Center-Boundary ReasoningYafei Yang, Zihui Zhang, Bo YangICML 2025
- From Inpainting to Layer Decomposition: Repurposing Generative Inpainting Models for Image Layer DecompositionJingxi Chen, Yixiao Zhang, Xiaoye Qian, Zongxia Li 等CVPR 2026 · 被引用 5 次
