Flow Factorized Representation Learning
Yue Song, Andy Keller, Nicu Sebe, Max Welling
摘要
A prominent goal of representation learning research is to achieve representations which are factorized in a useful manner with respect to the ground truth factors of variation. The fields of disentangled and equivariant representation learning have approached this ideal from a range of complimentary perspectives; however, to date, most approaches have proven to either be ill-specified or insufficiently flexible to effectively separate all realistic factors of interest in a learned latent space. In this work, we propose an alternative viewpoint on such structured representation learning which we call Flow Factorized Representation Learning, and demonstrate it to learn both more efficient and more usefully structured representations than existing frameworks. Specifically, we introduce a generative model which specifies a distinct set of latent probability paths that define different input transformations. Each latent flow is generated by the gradient field of a learned potential following dynamic optimal transport. Our novel setup brings new understandings to both disentanglement and equivariance. We show that our model achieves higher likelihoods on standard representation learning benchmarks while simultaneously being closer to approximately equivariant models. Furthermore, we demonstrate that the transformations learned by our model are flexibly composable and can also extrapolate to new data, implying a degree of robustness and generalizability approaching the ultimate goal of usefully factorized representation learning.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Transferring disentangled representations: bridging the gap between synthetic and real imagesJacopo Dapueto, Nicoletta Noceti, Francesca OdoneNeurIPS 2024 · 被引用 3 次
- Unsupervised Region-Based Image Editing of Denoising Diffusion ModelsZixiang Li, Yue Song, Renshuai Tao, Xiaohong Jia 等AAAI 2025 · 被引用 1 次
- Partially Observed Trajectory Inference using Optimal Transport and a Dynamics PriorAnming Gu, Edward Chien, Kristjan H. GreenewaldICLR 2025
- CoGe-GCD: Reframing Generalized Category Discovery with Compositional GeneralizationLuyao Tang, Jiewei Zheng, Kunze Huang, Chaoqi Chen 等ICML 2026
它引用的顶会 Paper39
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- GANSpace: Discovering Interpretable GAN ControlsErik Härkönen, Aaron Hertzmann, Jaakko Lehtinen, Sylvain ParisNeurIPS 2020 · 被引用 1,049 次
- Equivariant Diffusion for Molecule Generation in 3DEmiel Hoogeboom, Victor Garcia Satorras, Clément Vignac, Max WellingICML 2022 · 被引用 865 次
- Unsupervised Discovery of Interpretable Directions in the GAN Latent SpaceAndrey Voynov, Artem BabenkoICML 2020 · 被引用 459 次
- On the "steerability" of generative adversarial networksAli Jahanian, Lucy Chai, Phillip IsolaICLR 2020 · 被引用 421 次
相关 Paper
- Latent Traversals in Generative Models as Potential FlowsYue Song, T. Anderson Keller, Nicu Sebe, Max WellingICML 2023 · 被引用 15 次
- Equivariant Latent Alignment via Flow Matching under Group SymmetriesSunghyun Kim, Jaehoon Hahm, Jeongwoo Shin, Joonseok LeeICML 2026
- Learning Disentangled Representation by Exploiting Pretrained Generative Models: A Contrastive Learning ViewXuanchi Ren, Tao Yang, Yuwang Wang, Wenjun ZengICLR 2022 · 被引用 54 次
- Commutative Lie Group VAE for Disentanglement LearningXinqi Zhu, Chang Xu, Dacheng TaoICML 2021 · 被引用 35 次
- On the Relation between Rectified Flows and Optimal TransportJohannes Hertrich, Antonin Chambolle, Julie DelonNeurIPS 2025 · 被引用 15 次
