Multifactor Sequential Disentanglement via Structured Koopman Autoencoders
Nimrod Berman, Ilan Naiman, Omri Azencot
Abstract
Disentangling complex data to its latent factors of variation is a fundamental task in representation learning. Existing work on sequential disentanglement mostly provides two factor representations, i.e., it separates the data to time-varying and time-invariant factors. In contrast, we consider multifactor disentanglement in which multiple (more than two) semantic disentangled components are generated. Key to our approach is a strong inductive bias where we assume that the underlying dynamics can be represented linearly in the latent space. Under this assumption, it becomes natural to exploit the recently introduced Koopman autoencoder models. However, disentangled representations are not guaranteed in Koopman approaches, and thus we propose a novel spectral loss term which leads to structured Koopman matrices and disentanglement. Overall, we propose a simple and easy to code new deep model that is fully unsupervised and it supports multifactor disentanglement. We showcase new disentangling abilities such as swapping of individual static factors between characters, and an incremental swap of disentangled factors from the source to the target. Moreover, we evaluate our method extensively on two factor standard benchmark tasks where we significantly improve over competing unsupervised approaches, and we perform competitively in comparison to weakly- and self-supervised state-of-the-art approaches. The code is available at https://github.com/azencot-group/SKD.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f3785fe9-8196-498a-bd4c-bdadcaa78880Cited by top-tier papers15
- Koopa: Learning Non-stationary Time Series Dynamics with Koopman PredictorsYong Liu, Chenyu Li, Jianmin Wang, Mingsheng LongNeurIPS 2023 · 284 citations
- Utilizing Image Transforms and Diffusion Models for Generative Modeling of Short and Long Time SeriesIlan Naiman, Nimrod Berman, Itai Pemper, Idan Arbiv et al.NeurIPS 2024 · 69 citations
- Generative Modeling of Regular and Irregular Time Series Data via Koopman VAEsIlan Naiman, N. Benjamin Erichson, Pu Ren, Michael W. Mahoney et al.ICLR 2024 · 49 citations
- Temporally Disentangled Representation Learning under Unknown NonstationarityXiangchen Song, Weiran Yao, Yewen Fan, Xinshuai Dong et al.NeurIPS 2023 · 36 citations
- CaRiNG: Learning Temporal Causal Representation under Non-Invertible Generation ProcessGuangyi Chen, Yifan Shen, Zhenhao Chen, Xiangchen Song et al.ICML 2024 · 22 citations
Builds on10
- Forecasting Sequential Data Using Consistent Koopman AutoencodersOmri Azencot, N. Benjamin Erichson, Vanessa Lin, Michael W. MahoneyICML 2020 · 203 citations
- Learning Compositional Koopman Operators for Model-Based ControlYunzhu Li, Hao He, Jiajun Wu, Dina Katabi et al.ICLR 2020 · 135 citations
- Contrastively Disentangled Sequential Variational AutoencoderJunwen Bai, Weiran Wang, Carla P. GomesNeurIPS 2021 · 60 citations
- DeSKO: Stability-Assured Robust Control with a Deep Stochastic Koopman OperatorMinghao Han, Jacob Euler-Rolle, Robert K. KatzschmannICLR 2022 · 53 citations
- An Operator Theoretic View On Pruning Deep Neural NetworksWilliam T. Redman, Maria Fonoberova, Ryan Mohr, Yannis G. Kevrekidis et al.ICLR 2022 · 21 citations
Related papers
- S3VAE: Self-Supervised Sequential VAE for Representation Disentanglement and Data GenerationYizhe Zhu, Martin Renqiang Min, Asim Kadav, Hans Peter GrafCVPR 2020
- Sequential Disentanglement by Extracting Static Information From A Single Sequence ElementNimrod Berman, Ilan Naiman, Idan Arbiv, Gal Fadlon et al.ICML 2024 · 9 citations
- Disentangled Recurrent Wasserstein AutoencoderJun Han, Martin Renqiang Min, Ligong Han, Li Erran Li et al.ICLR 2021 · 37 citations
- VDSM: Unsupervised Video Disentanglement With State-Space Modeling and Deep Mixtures of ExpertsMatthew J. Vowels, Necati Cihan Camgöz, Richard BowdenCVPR 2021
- Sample and Predict Your Latent: Modality-free Sequential Disentanglement via Contrastive EstimationIlan Naiman, Nimrod Berman, Omri AzencotICML 2023 · 12 citations
