Identification of Causal Structure in the Presence of Missing Data with Additive Noise Model
Jie Qiao, Zhengming Chen, Jianhua Yu, Ruichu Cai, Zhifeng Hao
Abstract
Missing data are an unavoidable complication frequently encountered in many causal discovery tasks. When a missing process depends on the missing values themselves (known as self-masking missingness), the recovery of the joint distribution becomes unattainable, and detecting the presence of such self-masking missingness remains a perplexing challenge. Consequently, due to the inability to reconstruct the original distribution and to discern the underlying missingness mechanism, simply applying existing causal discovery methods would lead to wrong conclusions. In this work, we found that the recent advances additive noise model has the potential for learning causal structure under the existence of the self-masking missingness. With this observation, we aim to investigate the identification problem of learning causal structure from missing data under an additive noise model with different missingness mechanisms, where the `no self-masking missingness' assumption can be eliminated appropriately. Specifically, we first elegantly extend the scope of identifiability of causal skeleton to the case with weak self-masking missingness (i.e., no other variable could be the cause of self-masking indicators except itself). We further provide the sufficient and necessary identification conditions of the causal direction under additive noise model and show that the causal structure can be identified up to an IN-equivalent pattern. We finally propose a practical algorithm based on the above theoretical results on learning the causal skeleton and causal direction. Extensive experiments on synthetic and real data demonstrate the efficiency and effectiveness of the proposed algorithms.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 584df7e4-4102-4642-bc76-b2f594303857Cited by top-tier papers2
- On the Identifiability of Poisson Branching Structural Causal Model Using Probability Generating FunctionYu Xiang, Jie Qiao, Zefeng Liang, Zihuai Zeng et al.NeurIPS 2024 · 3 citations
- Causal Discovery for Irregularly Time Series with Consistency GuaranteesWeihong Li, Baohong Li, Anpeng Wu, Zhihan Li et al.ICML 2026
Builds on3
- Full Law Identification in Graphical Models of Missing Data: Completeness ResultsRazieh Nabi, Rohit Bhattacharya, Ilya ShpitserICML 2020 · 60 citations
- Identifiable Generative models for Missing Not at Random Data ImputationChao Ma, Cheng ZhangNeurIPS 2021 · 56 citations
- MissDAG: Causal Discovery in the Presence of Missing Data with Continuous Additive Noise ModelsErdun Gao, Ignavier Ng, Mingming Gong, Li Shen et al.NeurIPS 2022 · 36 citations
Related papers
- Identification of Linear Latent Variable Model with Arbitrary DistributionZhengming Chen, Feng Xie, Jie Qiao, Zhifeng Hao et al.AAAI 2022 · 24 citations
- Strong and Weak Identifiability of Optimization-based Causal Discovery in Non-linear Additive Noise ModelsMingjia Li, Hong Qian, Tian-Zuo Wang, Shujun Li et al.ICML 2025
- Score Matching Enables Causal Discovery of Nonlinear Additive Noise ModelsPaul Rolland, Volkan Cevher, Matthäus Kleindessner, Chris Russell et al.ICML 2022 · 123 citations
- Causal Representation Learning Made Identifiable by Grouping of Observational VariablesHiroshi Morioka, Aapo HyvärinenICML 2024 · 26 citations
- Causal Discovery from Subsampled Time Series with Proxy VariablesMingzhou Liu, Xinwei Sun, Lingjing Hu, Yizhou WangNeurIPS 2023 · 14 citations
