3D Shape Variational Autoencoder Latent Disentanglement via Mini-Batch Feature Swapping for Bodies and Faces
Simone Foti, Bongjin Koo, Danail Stoyanov, Matthew J. Clarkson
Abstract
Learning a disentangled, interpretable, and structured latent representation in 3D generative models of faces and bodies is still an open problem. The problem is particularly acute when control over identity features is required. In this paper, we propose an intuitive yet effective self-supervised approach to train a 3D shape variational autoencoder (VAE) which encourages a disentangled latent representation of identity features. Curating the mini-batch generation by swapping arbitrary features across different shapes allows to define a loss function leveraging known differences and similarities in the latent representations. Experimental results conducted on 3D meshes show that state-of-the-art methods for latent disentanglement are not able to disentangle identity features of faces and bodies. Our proposed method properly decouples the generation of such features while maintaining good representation and reconstruction capabilities. Our code and pretrained models are available at github.com/simofoti/3DVAE-SwapDisentangled.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9d299a07-e6a6-4d20-992d-a46bec002777Cited by top-tier papers5
- Taming Video Models for 3D and 4D Generation via Zero-Shot Camera ControlChenxi Song, Yanming Yang, Tong Zhao, Ruibo Li et al.CVPR 2026 · 17 citations
- TUVF: Learning Generalizable Texture UV Radiance FieldsAn-Chieh Cheng, Xueting Li, Sifei Liu, Xiaolong WangICLR 2024 · 9 citations
- Parallelised Differentiable Straightest Geodesics for 3D MeshesHippolyte Verninas, Caner Korkmaz, Stefanos Zafeiriou, Tolga Birdal et al.CVPR 2026 · 4 citations
- Locally Adaptive Neural 3D Morphable ModelsMichail Tarasiou, Rolandos Alexandros Potamias, Eimear O' Sullivan, Stylianos Ploumpis et al.CVPR 2024 · 2 citations
- Representing 3D Faces with Learnable B-Spline VolumesPrashanth Chandran, Daoye Wang, Timo BolkartCVPR 2026
Builds on7
- PointFlow: 3D Point Cloud Generation With Continuous Normalizing FlowsGuandao Yang, Xun Huang, Zekun Hao, Ming-Yu Liu et al.ICCV 2019 · 794 citations
- Fully Convolutional Mesh Autoencoder using Efficient Spatially Varying KernelsYi Zhou, Chenglei Wu, Zimo Li, Chen Cao et al.NeurIPS 2020 · 98 citations
- Unsupervised Model Selection for Variational Disentangled Representation LearningSunny Duan, Loic Matthey, Andre Saraiva, Nick Watters et al.ICLR 2020 · 87 citations
- Geometric Disentanglement for Generative Latent Shape ModelsTristan Aumentado-Armstrong, Stavros Tsogkas, Allan D. Jepson, Sven J. DickinsonICCV 2019 · 61 citations
- A Decoupled 3D Facial Shape Model by Adversarial TrainingVictoria Fernández Abrevaya, Adnane Boukhayma, Stefanie Wuhrer, Edmond BoyerICCV 2019 · 36 citations
Related papers
- EditVAE: Unsupervised Parts-Aware Controllable 3D Point Cloud Shape GenerationShidi Li, Miaomiao Liu, Christian WalderAAAI 2022 · 35 citations
- 3D-Aware Face SwappingYixuan Li, Chao Ma, Yichao Yan, Wenhan Zhu et al.CVPR 2023
- Swapping Autoencoder for Deep Image ManipulationTaesung Park, Jun-Yan Zhu, Oliver Wang, Jingwan Lu et al.NeurIPS 2020 · 376 citations
- Guided Variational Autoencoder for Disentanglement LearningZheng Ding, Yifan Xu, Weijian Xu, Gaurav Parmar et al.CVPR 2020
- Information Bottleneck Disentanglement for Identity SwappingGege Gao, Huaibo Huang, Chaoyou Fu, Zhaoyang Li et al.CVPR 2021
