CausalVAE: Disentangled Representation Learning via Neural Structural Causal Models
Mengyue Yang, Furui Liu, Zhitang Chen, Xinwei Shen, Jianye Hao, Jun Wang
Abstract
Learning disentanglement aims at finding a low dimensional representation which consists of multiple explanatory and generative factors of the observational data. The framework of variational autoencoder (VAE) is commonly used to disentangle independent factors from observations. However, in real scenarios, factors with semantics are not necessarily independent. Instead, there might be an underlying causal structure which renders these factors dependent. We thus propose a new VAE based framework named CausalVAE, which includes a Causal Layer to transform independent exogenous factors into causal endogenous ones that correspond to causally related concepts in data. We further analyze the model identifiabitily, showing that the proposed model learned from observations recovers the true one up to a certain degree by providing supervision signals (e.g. feature labels). Experiments are conducted on various datasets, including synthetic and real word benchmark CelebA. Results show that the causal representations learned by CausalVAE are semantically interpretable, and their causal relationship as a Directed Acyclic Graph (DAG) is identified with good accuracy. Furthermore, we demonstrate that the proposed CausalVAE model is able to generate counterfactual data through "do-operation" to the causal factors.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 686643ca-ccd0-49d1-aeb2-fa2d06cde3dbCited by top-tier papers78
- Weakly supervised causal representation learningJohann Brehmer, Pim de Haan, Phillip Lippe, Taco S. CohenNeurIPS 2022 · 196 citations
- DECAF: Generating Fair Synthetic Data Using Causally-Aware Generative NetworksBoris van Breugel, Trent Kyono, Jeroen Berrevoets, Mihaela van der SchaarNeurIPS 2021 · 174 citations
- DiBS: Differentiable Bayesian Structure LearningLars Lorch, Jonas Rothfuss, Bernhard Schölkopf, Andreas KrauseNeurIPS 2021 · 144 citations
- CITRIS: Causal Identifiability from Temporal Intervened SequencesPhillip Lippe, Sara Magliacane, Sindy Löwe, Yuki M. Asano et al.ICML 2022 · 136 citations
- Nonparametric Identifiability of Causal Representations from Unknown InterventionsJulius von Kügelgen, Michel Besserve, Wendong Liang, Luigi Gresele et al.NeurIPS 2023 · 127 citations
Builds on3
- Causal Discovery with Reinforcement LearningShengyu Zhu, Ignavier Ng, Zhitang ChenICLR 2020 · 285 citations
- Counterfactuals uncover the modular structure of deep generative modelsMichel Besserve, Arash Mehrjou, Rémy Sun, Bernhard SchölkopfICLR 2020 · 109 citations
- Disentangling Factors of Variations Using Few LabelsFrancesco Locatello, Michael Tschannen, Stefan Bauer, Gunnar Rätsch et al.ICLR 2020 · 77 citations
Related papers
- Counterfactual Fairness with Disentangled Causal Effect Variational AutoencoderHyemi Kim, Seungjae Shin, JoonHo Jang, Kyungwoo Song et al.AAAI 2021 · 72 citations
- A Critical Look at the Consistency of Causal Estimation with Deep Latent Variable ModelsSeveri Rissanen, Pekka MarttinenNeurIPS 2021 · 38 citations
- Causal Representation Learning via Counterfactual InterventionXiutian Li, Siqi Sun, Rui FengAAAI 2024 · 13 citations
- Counterfactual Generative Modeling with Variational Causal InferenceYulun Wu, Louie McConnell, Claudia IriondoICLR 2025
- Identifiability Guarantees for Causal Disentanglement from Soft InterventionsJiaqi Zhang, Kristjan H. Greenewald, Chandler Squires, Akash Srivastava et al.NeurIPS 2023 · 120 citations
