CASTLE: Regularization via Auxiliary Causal Graph Discovery
Trent Kyono, Yao Zhang, Mihaela van der Schaar
摘要
Regularization improves generalization of supervised models to out-of-sample data. Prior works have shown that prediction in the causal direction (effect from cause) results in lower testing error than the anti-causal direction. However, existing regularization methods are agnostic of causality. We introduce Causal Structure Learning (CASTLE) regularization and propose to regularize a neural network by jointly learning the causal relationships between variables. CASTLE learns the causal directed acyclical graph (DAG) as an adjacency matrix embedded in the neural network's input layers, thereby facilitating the discovery of optimal predictors. Furthermore, CASTLE efficiently reconstructs only the features in the causal DAG that have a causal neighbor, whereas reconstruction-based regularizers suboptimally reconstruct all input features. We provide a theoretical generalization bound for our approach and conduct experiments on a plethora of synthetic and real publicly available datasets demonstrating that CASTLE consistently leads to better out-of-sample predictions as compared to other popular benchmark regularizers.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper17
- DAGMA: Learning DAGs via M-matrices and a Log-Determinant Acyclicity CharacterizationKevin Bello, Bryon Aragam, Pradeep RavikumarNeurIPS 2022 · 被引用 222 次
- TabPFN: A Transformer That Solves Small Tabular Classification Problems in a SecondNoah Hollmann, Samuel Müller, Katharina Eggensperger, Frank HutterICLR 2023 · 被引用 96 次
- CausPref: Causal Preference Learning for Out-of-Distribution RecommendationYue He, Zimu Wang, Peng Cui, Hao Zou 等WWW 2022 · 被引用 64 次
- DARING: Differentiable Causal Discovery with Residual IndependenceYue He, Peng Cui, Zheyan Shen, Renzhe Xu 等KDD 2021 · 被引用 28 次
- Simultaneous Missing Value Imputation and Structure Learning with GroupsPablo Morales-Alvarez, Wenbo Gong, Angus Lamb, Simon Woodhead 等NeurIPS 2022 · 被引用 24 次
它引用的顶会 Paper2
相关 Paper
- Matching Learned Causal Effects of Neural Networks with Domain PriorsSai Srinivas Kancheti, Abbavaram Gowtham Reddy, Vineeth N. Balasubramanian, Amit SharmaICML 2022 · 被引用 17 次
- Towards Learning and Explaining Indirect Causal Effects in Neural NetworksAbbavaram Gowtham Reddy, Saketh Bachu, Harsharaj Pathak, Benin Godfrey L 等AAAI 2024 · 被引用 3 次
- Returning The Favour: When Regression Benefits From Probabilistic Causal KnowledgeShahine Bouabid, Jake Fawkes, Dino SejdinovicICML 2023
- Learning Large DAGs is Harder than you Think: Many Losses are Minimal for the Wrong DAGJonas Seng, Matej Zecevic, Devendra Singh Dhami, Kristian KerstingICLR 2024 · 被引用 9 次
- Deciphering Spatio-Temporal Graph Forecasting: A Causal Lens and TreatmentYutong Xia, Yuxuan Liang, Haomin Wen, Xu Liu 等NeurIPS 2023 · 被引用 110 次
