InSeGAN: A Generative Approach to Segmenting Identical Instances in Depth Images
Anoop Cherian, Gonçalo Dias Pais, Siddarth Jain, Tim K. Marks, Alan Sullivan
摘要
In this paper, we present InSeGAN, an unsupervised 3D generative adversarial network (GAN) for segmenting (nearly) identical instances of rigid objects in depth images. Using an analysis-by-synthesis approach, we design a novel GAN architecture to synthesize a multiple-instance depth image with independent control over each instance. InSeGAN takes in a set of code vectors (e.g., random noise vectors), each encoding the 3D pose of an object that is represented by a learned implicit object template. The generator has two distinct modules. The first module, the instance feature generator, uses each encoded pose to transform the implicit template into a feature map representation of each object instance. The second module, the depth image renderer, aggregates all of the single-instance feature maps output by the first module and generates a multiple-instance depth image. A discriminator distinguishes the generated multiple-instance depth images from the distribution of true depth images. To use our model for instance segmentation, we propose an instance pose encoder that learns to take in a generated depth image and reproduce the pose code vectors for all of the object instances. To evaluate our approach, we introduce a new synthetic dataset, "Insta-10," consisting of 100,000 depth images, each with 5 instances of an object from one of 10 classes. Our experiments on Insta-10, as well as on real-world noisy depth images, show that InSeGAN achieves state-of-the-art performance, often outperforming prior methods by large margins.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper4
- Object-Centric Learning with Slot AttentionFrancesco Locatello, Dirk Weissenborn, Thomas Unterthiner, Aravindh Mahendran 等NeurIPS 2020 · 被引用 1,275 次
- Escaping Plato's Cave: 3D Shape From Adversarial RenderingPhilipp Henzler, Niloy J. Mitra, Tobias RitschelICCV 2019 · 被引用 254 次
- HoloGAN: Unsupervised Learning of 3D Representations From Natural ImagesThu Nguyen-Phuoc, Chuan Li, Lucas Theis, Christian Richardt 等ICCV 2019 · 被引用 98 次
- Towards Unsupervised Learning of Generative Models for 3D Controllable Image SynthesisYiyi Liao, Katja Schwarz, Lars M. Mescheder, Andreas GeigerCVPR 2020
相关 Paper
- View Independent Generative Adversarial Network for Novel View SynthesisXiaogang Xu, Ying-Cong Chen, Jiaya JiaICCV 2019 · 被引用 43 次
- Unsupervised K-modal styled content generationOmry Sendik, Dani Lischinski, Daniel Cohen-OrSIGGRAPH 2020 · 被引用 6 次
- Intrinsic-Extrinsic Preserved GANs for Unsupervised 3D Pose TransferHaoyu Chen, Hao Tang, Henglin Shi, Wei Peng 等ICCV 2021 · 被引用 33 次
- Collaging Class-specific GANs for Semantic Image SynthesisYuheng Li, Yijun Li, Jingwan Lu, Eli Shechtman 等ICCV 2021 · 被引用 36 次
- UnScene3D: Unsupervised 3D Instance Segmentation for Indoor ScenesDávid Rozenberszki, Or Litany, Angela DaiCVPR 2024 · 被引用 25 次
