CASTLE: Regularization via Auxiliary Causal Graph Discovery
Trent Kyono, Yao Zhang, Mihaela van der Schaar
Abstract
Regularization improves generalization of supervised models to out-of-sample data. Prior works have shown that prediction in the causal direction (effect from cause) results in lower testing error than the anti-causal direction. However, existing regularization methods are agnostic of causality. We introduce Causal Structure Learning (CASTLE) regularization and propose to regularize a neural network by jointly learning the causal relationships between variables. CASTLE learns the causal directed acyclical graph (DAG) as an adjacency matrix embedded in the neural network's input layers, thereby facilitating the discovery of optimal predictors. Furthermore, CASTLE efficiently reconstructs only the features in the causal DAG that have a causal neighbor, whereas reconstruction-based regularizers suboptimally reconstruct all input features. We provide a theoretical generalization bound for our approach and conduct experiments on a plethora of synthetic and real publicly available datasets demonstrating that CASTLE consistently leads to better out-of-sample predictions as compared to other popular benchmark regularizers.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 455d25e4-bfc3-469b-827b-acec6860b487Cited by top-tier papers17
- DAGMA: Learning DAGs via M-matrices and a Log-Determinant Acyclicity CharacterizationKevin Bello, Bryon Aragam, Pradeep RavikumarNeurIPS 2022 · 222 citations
- TabPFN: A Transformer That Solves Small Tabular Classification Problems in a SecondNoah Hollmann, Samuel Müller, Katharina Eggensperger, Frank HutterICLR 2023 · 96 citations
- CausPref: Causal Preference Learning for Out-of-Distribution RecommendationYue He, Zimu Wang, Peng Cui, Hao Zou et al.WWW 2022 · 64 citations
- DARING: Differentiable Causal Discovery with Residual IndependenceYue He, Peng Cui, Zheyan Shen, Renzhe Xu et al.KDD 2021 · 28 citations
- Simultaneous Missing Value Imputation and Structure Learning with GroupsPablo Morales-Alvarez, Wenbo Gong, Angus Lamb, Simon Woodhead et al.NeurIPS 2022 · 24 citations
Builds on2
Related papers
- Matching Learned Causal Effects of Neural Networks with Domain PriorsSai Srinivas Kancheti, Abbavaram Gowtham Reddy, Vineeth N. Balasubramanian, Amit SharmaICML 2022 · 17 citations
- Towards Learning and Explaining Indirect Causal Effects in Neural NetworksAbbavaram Gowtham Reddy, Saketh Bachu, Harsharaj Pathak, Benin Godfrey L et al.AAAI 2024 · 3 citations
- Returning The Favour: When Regression Benefits From Probabilistic Causal KnowledgeShahine Bouabid, Jake Fawkes, Dino SejdinovicICML 2023
- Learning Large DAGs is Harder than you Think: Many Losses are Minimal for the Wrong DAGJonas Seng, Matej Zecevic, Devendra Singh Dhami, Kristian KerstingICLR 2024 · 9 citations
- Deciphering Spatio-Temporal Graph Forecasting: A Causal Lens and TreatmentYutong Xia, Yuxuan Liang, Haomin Wen, Xu Liu et al.NeurIPS 2023 · 110 citations
