Gromov-Wasserstein Autoencoders
Nao Nakagawa, Ren Togo, Takahiro Ogawa, Miki Haseyama
Abstract
Variational Autoencoder (VAE)-based generative models offer flexible representation learning by incorporating meta-priors, general premises considered beneficial for downstream tasks. However, the incorporated meta-priors often involve ad-hoc model deviations from the original likelihood architecture, causing undesirable changes in their training. In this paper, we propose a novel representation learning method, Gromov-Wasserstein Autoencoders (GWAE), which directly matches the latent and data distributions using the variational autoencoding scheme. Instead of likelihood-based objectives, GWAE models minimize the Gromov-Wasserstein (GW) metric between the trainable prior and given data distributions. The GW metric measures the distance structure-oriented discrepancy between distributions even with different dimensionalities, which provides a direct measure between the latent and data spaces. By restricting the prior family, we can introduce meta-priors into the latent space without changing their objective. The empirical comparisons with VAE-based models show that GWAE models work in two prominent meta-priors, disentanglement and clustering, with their GW objective unchanged.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6a120f74-b2dc-4a01-9de5-aa76abdc47c3Cited by top-tier papers1
Ask how each one uses itBuilds on11
- NVAE: A Deep Hierarchical Variational AutoencoderArash Vahdat, Jan KautzNeurIPS 2020 · 1,141 citations
- Self-labelling via simultaneous clustering and representation learningYuki Markus Asano, Christian Rupprecht, Andrea VedaldiICLR 2020 · 873 citations
- Simple and Effective VAE Training with Calibrated DecodersOleh Rybkin, Kostas Daniilidis, Sergey LevineICML 2021 · 119 citations
- A Contrastive Learning Approach for Training Variational Autoencoder PriorsJyoti Aneja, Alexander G. Schwing, Jan Kautz, Arash VahdatNeurIPS 2021 · 112 citations
- Theory and Evaluation Metrics for Learning Disentangled RepresentationsKien Do, Truyen TranICLR 2020 · 107 citations
Related papers
- Learning Autoencoders with Relational RegularizationHongteng Xu, Dixin Luo, Ricardo Henao, Svati Shah et al.ICML 2020 · 47 citations
- Guided Variational Autoencoder for Disentanglement LearningZheng Ding, Yifan Xu, Weijian Xu, Gaurav Parmar et al.CVPR 2020
- Improving Relational Regularized Autoencoders with Spherical Sliced Fused Gromov WassersteinKhai Nguyen, Son Nguyen, Nhat Ho, Tung Pham et al.ICLR 2021 · 21 citations
- Structure-Centric Graph Foundation Model via Geometric BasesXiaodong He, Haolan He, Ruiyi Fang, Ming Sun et al.ICML 2026 · 1 citation
- Model Selection for Bayesian AutoencodersBa-Hien Tran, Simone Rossi, Dimitrios Milios, Pietro Michiardi et al.NeurIPS 2021 · 15 citations
