Flow Factorized Representation Learning
Yue Song, Andy Keller, Nicu Sebe, Max Welling
Abstract
A prominent goal of representation learning research is to achieve representations which are factorized in a useful manner with respect to the ground truth factors of variation. The fields of disentangled and equivariant representation learning have approached this ideal from a range of complimentary perspectives; however, to date, most approaches have proven to either be ill-specified or insufficiently flexible to effectively separate all realistic factors of interest in a learned latent space. In this work, we propose an alternative viewpoint on such structured representation learning which we call Flow Factorized Representation Learning, and demonstrate it to learn both more efficient and more usefully structured representations than existing frameworks. Specifically, we introduce a generative model which specifies a distinct set of latent probability paths that define different input transformations. Each latent flow is generated by the gradient field of a learned potential following dynamic optimal transport. Our novel setup brings new understandings to both disentanglement and equivariance. We show that our model achieves higher likelihoods on standard representation learning benchmarks while simultaneously being closer to approximately equivariant models. Furthermore, we demonstrate that the transformations learned by our model are flexibly composable and can also extrapolate to new data, implying a degree of robustness and generalizability approaching the ultimate goal of usefully factorized representation learning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e3e48aac-894f-4626-93c4-c2be665841d9Cited by top-tier papers4
- Transferring disentangled representations: bridging the gap between synthetic and real imagesJacopo Dapueto, Nicoletta Noceti, Francesca OdoneNeurIPS 2024 · 3 citations
- Unsupervised Region-Based Image Editing of Denoising Diffusion ModelsZixiang Li, Yue Song, Renshuai Tao, Xiaohong Jia et al.AAAI 2025 · 1 citation
- Partially Observed Trajectory Inference using Optimal Transport and a Dynamics PriorAnming Gu, Edward Chien, Kristjan H. GreenewaldICLR 2025
- CoGe-GCD: Reframing Generalized Category Discovery with Compositional GeneralizationLuyao Tang, Jiewei Zheng, Kunze Huang, Chaoqi Chen et al.ICML 2026
Builds on39
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- GANSpace: Discovering Interpretable GAN ControlsErik Härkönen, Aaron Hertzmann, Jaakko Lehtinen, Sylvain ParisNeurIPS 2020 · 1,049 citations
- Equivariant Diffusion for Molecule Generation in 3DEmiel Hoogeboom, Victor Garcia Satorras, Clément Vignac, Max WellingICML 2022 · 865 citations
- Unsupervised Discovery of Interpretable Directions in the GAN Latent SpaceAndrey Voynov, Artem BabenkoICML 2020 · 459 citations
- On the "steerability" of generative adversarial networksAli Jahanian, Lucy Chai, Phillip IsolaICLR 2020 · 421 citations
Related papers
- Latent Traversals in Generative Models as Potential FlowsYue Song, T. Anderson Keller, Nicu Sebe, Max WellingICML 2023 · 15 citations
- Equivariant Latent Alignment via Flow Matching under Group SymmetriesSunghyun Kim, Jaehoon Hahm, Jeongwoo Shin, Joonseok LeeICML 2026
- Learning Disentangled Representation by Exploiting Pretrained Generative Models: A Contrastive Learning ViewXuanchi Ren, Tao Yang, Yuwang Wang, Wenjun ZengICLR 2022 · 54 citations
- Commutative Lie Group VAE for Disentanglement LearningXinqi Zhu, Chang Xu, Dacheng TaoICML 2021 · 35 citations
- On the Relation between Rectified Flows and Optimal TransportJohannes Hertrich, Antonin Chambolle, Julie DelonNeurIPS 2025 · 15 citations
