Self-Distilled Disentangled Learning for Counterfactual Prediction
Xinshu Li, Mingming Gong, Lina Yao
Abstract
The advancements in disentangled representation learning significantly enhance the accuracy of counterfactual predictions by granting precise control over instrumental variables, confounders, and adjustable variables. An appealing method for achieving the independent separation of these factors is mutual information minimization, a task that presents challenges in numerous machine learning scenarios, especially within high-dimensional spaces. To circumvent this challenge, we propose the Self-Distilled Disentanglement framework, referred to as 𝑆𝐷 2 . Grounded in information theory, it ensures theoretically sound independent disentangled representations without intricate mutual information estimator designs for high-dimensional representations. Our comprehensive experiments, conducted on both synthetic and real-world datasets, confirms the effectiveness of our approach in facilitating counterfactual inference in the presence of both observed and unobserved confounders.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f2ce898b-d025-45ed-8768-f97b7212fa23Cited by top-tier papers2
- Beyond Whole Dialogue Modeling: Contextual Disentanglement for Conversational RecommendationGuojia An, Jie Zou, Jiwei Wei, Chaoning Zhang et al.SIGIR 2025 · 11 citations
- Causality-aligned Prompt Learning via Diffusion-based Counterfactual GenerationXinshu Li, Ruoyu Wang, Erdun Gao, Mingming Gong et al.ACM MM 2025 · 3 citations
Builds on10
- CLUB: A Contrastive Log-ratio Upper Bound of Mutual InformationPengyu Cheng, Weituo Hao, Shuyang Dai, Jiachang Liu et al.ICML 2020 · 512 citations
- Learning Disentangled Representations for CounterFactual RegressionNegar Hassanpour, Russell GreinerICLR 2020 · 176 citations
- Treatment Effect Estimation with Disentangled Latent FactorsWeijia Zhang, Lin Liu, Jiuyong LiAAAI 2021 · 115 citations
- Dual Instrumental Variable RegressionKrikamol Muandet, Arash Mehrjou, Si Kai Lee, Anant RajNeurIPS 2020 · 87 citations
- Learning Deep Features in Instrumental Variable RegressionLiyuan Xu, Yutian Chen, Siddarth Srinivasan, Nando de Freitas et al.ICLR 2021 · 85 citations
Related papers
- An Information Criterion for Controlled Disentanglement of Multimodal DataChenyu Wang, Sharut Gupta, Xinyi Zhang, Sana Tonekaboni et al.ICLR 2025
- FADES: Fair Disentanglement with Sensitive RelevanceTaeuk Jang, Xiaoqian WangCVPR 2024
- C-Disentanglement: Discovering Causally-Independent Generative Factors under an Inductive Bias of ConfounderXiaoyu Liu, Jiaxin Yuan, Bang An, Yuancheng Xu et al.NeurIPS 2023 · 13 citations
- Bounds on Representation-Induced Confounding Bias for Treatment Effect EstimationValentyn Melnychuk, Dennis Frauen, Stefan FeuerriegelICLR 2024 · 23 citations
- DisUnknown: Distilling Unknown Factors for Disentanglement LearningSitao Xiang, Yuming Gu, Pengda Xiang, Menglei Chai et al.ICCV 2021 · 6 citations
