Optimal Transport for Structure Learning Under Missing Data
Vy Vo, He Zhao, Trung Le, Edwin V. Bonilla, Dinh Phung
Abstract
Causal discovery in the presence of missing data introduces a chicken-and-egg dilemma. While the goal is to recover the true causal structure, robust imputation requires considering the dependencies or, preferably, causal relations among variables. Merely filling in missing values with existing imputation methods and subsequently applying structure learning on the complete data is empirically shown to be sub-optimal. To address this problem, we propose a score-based algorithm for learning causal structures from missing data based on optimal transport. This optimal transport viewpoint diverges from existing score-based approaches that are dominantly based on expectation maximization. We formulate structure learning as a density fitting problem, where the goal is to find the causal model that induces a distribution of minimum Wasserstein distance with the observed data distribution. Our framework is shown to recover the true causal graphs more effectively than competing methods in most simulations and real-data settings. Empirical evidence also shows the superior scalability of our approach, along with the flexibility to incorporate any off-the-shelf causal discovery methods for complete data.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c4cd61a1-1dca-485b-a447-553be3c1ded3Cited by top-tier papers7
- Distribution Alignment Optimization through Neural Collapse for Long-tailed ClassificationJintong Gao, He Zhao, Dandan Guo, Hongyuan ZhaICML 2024 · 27 citations
- Neural Topic Modeling with Large Language Models in the LoopXiaohao Yang, He Zhao, Weijie Xu, Yuanyuan Qi et al.ACL 2025 · 13 citations
- Parameter Estimation in DAGs from Incomplete Data via Optimal TransportVy Vo, Trung Le, Long Tung Vuong, He Zhao et al.ICML 2024 · 5 citations
- Ordering-based Causal Discovery via Generalized Score MatchingVy Vo, Trung Le, He Zhao, Edwin V. Bonilla et al.KDD 2026 · 1 citation
- Robust Simulation-Based Inference under Missing Data via Neural ProcessesYogesh Verma, Ayush Bharti, Vikas GargICLR 2025
Builds on27
- Gradient-Based Neural DAG LearningSébastien Lachapelle, Philippe Brouillard, Tristan Deleu, Simon Lacoste-JulienICLR 2020 · 337 citations
- On the Role of Sparsity and DAG Constraints for Learning Linear DAGsIgnavier Ng, AmirEmad Ghassami, Kun ZhangNeurIPS 2020 · 306 citations
- Causal Discovery with Reinforcement LearningShengyu Zhu, Ignavier Ng, Zhitang ChenICLR 2020 · 285 citations
- Handling Missing Data with Graph Representation LearningJiaxuan You, Xiaobai Ma, Daisy Yi Ding, Mykel J. Kochenderfer et al.NeurIPS 2020 · 274 citations
- Missing Data Imputation using Optimal TransportBoris Muzellec, Julie Josse, Claire Boyer, Marco CuturiICML 2020 · 179 citations
Related papers
- MissDAG: Causal Discovery in the Presence of Missing Data with Continuous Additive Noise ModelsErdun Gao, Ignavier Ng, Mingming Gong, Li Shen et al.NeurIPS 2022 · 36 citations
- MissScore: High-Order Score Estimation in the Presence of Missing DataWenqin Liu, Haoze Hou, Erdun Gao, Biwei Huang et al.ICML 2025
- Score Matching Enables Causal Discovery of Nonlinear Additive Noise ModelsPaul Rolland, Volkan Cevher, Matthäus Kleindessner, Chris Russell et al.ICML 2022 · 123 citations
- Learning to Induce Causal StructureNan Rosemary Ke, Silvia Chiappa, Jane X. Wang, Jörg Bornschein et al.ICLR 2023 · 17 citations
- Causal Discovery for Irregularly Time Series with Consistency GuaranteesWeihong Li, Baohong Li, Anpeng Wu, Zhihan Li et al.ICML 2026
