Sequential Disentanglement by Extracting Static Information From A Single Sequence Element
Nimrod Berman, Ilan Naiman, Idan Arbiv, Gal Fadlon, Omri Azencot
摘要
One of the fundamental representation learning tasks is unsupervised sequential disentanglement, where latent codes of inputs are decomposed to a single static factor and a sequence of dynamic factors. To extract this latent information, existing methods condition the static and dynamic codes on the entire input sequence. Unfortunately, these models often suffer from information leakage, i.e., the dynamic vectors encode both static and dynamic information, or vice versa, leading to a non-disentangled representation. Attempts to alleviate this problem via reducing the dynamic dimension and auxiliary loss terms gain only partial success. Instead, we propose a novel and simple architecture that mitigates information leakage by offering a simple and effective subtraction inductive bias while conditioning on a single sample. Remarkably, the resulting variational framework is simpler in terms of required loss terms, hyperparameters, and data augmentation. We evaluate our method on multiple data-modality benchmarks including general time series, video, and audio, and we show beyond state-of-the-art results on generation and prediction tasks in comparison to several strong baselines. Code is at GitHub.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Dyn-O: Building Structured World Models with Object-Centric RepresentationsZizhao Wang, Kaixin Wang, Li Zhao, Peter Stone 等NeurIPS 2025 · 被引用 15 次
- DiffSDA: Unsupervised Diffusion Sequential Disentanglement Across ModalitiesHedi Zisling, Ilan Naiman, Nimrod Berman, Supasorn Suwajanakorn 等ICLR 2026 · 被引用 2 次
- Bitrate-Controlled Diffusion for Disentangling Motion and Content in VideoXiao Li, Qi Chen, Xiulian Peng, Kai Yu 等ICCV 2025 · 被引用 1 次
它引用的顶会 Paper12
- Informer: Beyond Efficient Transformer for Long Sequence Time-Series ForecastingHaoyi Zhou, Shanghang Zhang, Jieqi Peng, Shuai Zhang 等AAAI 2021 · 被引用 7,289 次
- A Good Image Generator Is What You Need for High-Resolution Video SynthesisYu Tian, Jian Ren, Menglei Chai, Kyle Olszewski 等ICLR 2021 · 被引用 208 次
- Disentangling Factors of Variations Using Few LabelsFrancesco Locatello, Michael Tschannen, Stefan Bauer, Gunnar Rätsch 等ICLR 2020 · 被引用 77 次
- Contrastively Disentangled Sequential Variational AutoencoderJunwen Bai, Weiran Wang, Carla P. GomesNeurIPS 2021 · 被引用 60 次
- Generative Modeling of Regular and Irregular Time Series Data via Koopman VAEsIlan Naiman, N. Benjamin Erichson, Pu Ren, Michael W. Mahoney 等ICLR 2024 · 被引用 49 次
相关 Paper
- S3VAE: Self-Supervised Sequential VAE for Representation Disentanglement and Data GenerationYizhe Zhu, Martin Renqiang Min, Asim Kadav, Hans Peter GrafCVPR 2020
- Sample and Predict Your Latent: Modality-free Sequential Disentanglement via Contrastive EstimationIlan Naiman, Nimrod Berman, Omri AzencotICML 2023 · 被引用 12 次
- Disentangled Recurrent Wasserstein AutoencoderJun Han, Martin Renqiang Min, Ligong Han, Li Erran Li 等ICLR 2021 · 被引用 37 次
- VDSM: Unsupervised Video Disentanglement With State-Space Modeling and Deep Mixtures of ExpertsMatthew J. Vowels, Necati Cihan Camgöz, Richard BowdenCVPR 2021
- Multifactor Sequential Disentanglement via Structured Koopman AutoencodersNimrod Berman, Ilan Naiman, Omri AzencotICLR 2023 · 被引用 4 次
