Canonical Capsules: Self-Supervised Capsules in Canonical Pose
Weiwei Sun, Andrea Tagliasacchi, Boyang Deng, Sara Sabour, Soroosh Yazdani, Geoffrey E. Hinton, Kwang Moo Yi
摘要
We propose an unsupervised capsule architecture for 3D point clouds. We compute capsule decompositions of objects through permutation-equivariant attention, and self-supervise the process by training with pairs of randomly rotated objects. Our key idea is to aggregate the attention masks into semantic keypoints, and use these to supervise a decomposition that satisfies the capsule invariance/equivariance properties. This not only enables the training of a semantically consistent decomposition, but also allows us to learn a canonicalization operation that enables object-centric reasoning. In doing so, we require neither classification labels nor manually-aligned training datasets to train. Yet, by learning an object-centric representation in an unsupervised manner, our method outperforms the state-of-the-art on 3D point cloud reconstruction, registration, and unsupervised classification. We will release the code and dataset to reproduce our results as soon as the paper is published.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper13
- A Canonicalization Perspective on Invariant and Equivariant LearningGeorge Ma, Yifei Wang, Derek Lim, Stefanie Jegelka 等NeurIPS 2024 · 被引用 38 次
- Laplacian Canonization: A Minimalist Approach to Sign and Basis Invariant Spectral EmbeddingGeorge Ma, Yifei Wang, Yisen WangNeurIPS 2023 · 被引用 31 次
- The Devil is in the Pose: Ambiguity-free 3D Rotation-invariant Learning via Pose-aware ConvolutionRonghan Chen, Yang CongCVPR 2022 · 被引用 26 次
- Unsupervised Object Representation Learning using Translation and Rotation Group Equivariant VAEAlireza Nasiri, Tristan BeplerNeurIPS 2022 · 被引用 18 次
- PaRot: Patch-Wise Rotation-Invariant Network via Feature Disentanglement and Pose RestorationDingxin Zhang, Jianhui Yu, Chaoyi Zhang, Weidong CaiAAAI 2023 · 被引用 17 次
它引用的顶会 Paper13
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Deep Closest Point: Learning Representations for Point Cloud RegistrationYue Wang, Justin SolomonICCV 2019 · 被引用 1,026 次
- Learning Shape Templates With Structured Implicit FunctionsKyle Genova, Forrester Cole, Daniel Vlasic, Aaron Sarna 等ICCV 2019 · 被引用 427 次
- Vector Neurons: A General Framework for SO(3)-Equivariant NetworksCongyue Deng, Or Litany, Yueqi Duan, Adrien Poulenard 等ICCV 2021 · 被引用 411 次
- C3DPO: Canonical 3D Pose Networks for Non-Rigid Structure From MotionDavid Novotný, Nikhila Ravi, Benjamin Graham, Natalia Neverova 等ICCV 2019 · 被引用 126 次
相关 Paper
- ToThePoint: Efficient Contrastive Learning of 3D Point Clouds via RecyclingXinglin Li, Jiajing Chen, Jinhui Ouyang, Hanhui Deng 等CVPR 2023
- ConDor: Self-Supervised Canonicalization of 3D Pose for Partial ShapesRahul Sajnani, Adrien Poulenard, Jivitesh Jain, Radhika Dua 等CVPR 2022 · 被引用 28 次
- Leveraging SE(3) Equivariance for Self-supervised Category-Level Object Pose Estimation from Point CloudsXiaolong Li, Yijia Weng, Li Yi, Leonidas J. Guibas 等NeurIPS 2021 · 被引用 61 次
- Equicaps: Predictor-Free Pose-Aware Pre-Trained Capsule NetworksAthinoulla Konstantinou, Georgios Leontidis, Mamatha Thota, Aiden DurrantICCV 2025
- ShellNet: Efficient Point Cloud Convolutional Neural Networks Using Concentric Shells StatisticsZhiyuan Zhang, Binh-Son Hua, Sai-Kit YeungICCV 2019 · 被引用 400 次
