Spiked-CFR: Causal Representation Learning from LLMs via Wasserstein Projection Pursuit
Fan Wang, Hengyu Yue, Yu Bowen, Weiming Liu, Zongxin Yang, Xuyun Zhang, Xiaolin Zheng, Chaochao Chen, Shuiguang Deng
摘要
Estimating treatment effects from observational text is increasingly practical with Large Language Models (LLMs). However, applying causal representation learning directly to high-dimensional LLM embeddings faces a fundamental barrier: empirical Wasserstein matching suffers from the curse of dimensionality, rendering standard generalization guarantees effectively vacuous. We propose SPIKED-CFR, a framework bridging this gap by assuming a Spiked Structure, where treatment selection bias is assumed to manifest primarily as a low-dimensional treated--control discrepancy in the semantic representation. We develop Wasserstein Projection Pursuit, a minimax objective that adversarially learns an orthogonal projection on the Stiefel manifold to identify and balance only this subspace while preserving prognostic information. Under a spiked structure, we show the projected discrepancy can be estimated at a rate governed by the intrinsic dimension , and we derive a tighter PEHE generalization bound that depends on rather than the ambient embedding dimension. Experiments on four semi-synthetic benchmarks and four real-world clinical benchmarks demonstrate improved accuracy and robustness over strong baselines.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper12
- Learning Disentangled Representations for CounterFactual RegressionNegar Hassanpour, Russell GreinerICLR 2020 · 被引用 176 次
- Projection Robust Wasserstein Distance and Riemannian OptimizationTianyi Lin, Chenyou Fan, Nhat Ho, Marco Cuturi 等NeurIPS 2020 · 被引用 84 次
- FOOGD: Federated Collaboration for Both Out-of-distribution Generalization and DetectionXinting Liao, Weiming Liu, Pengyang Zhou, Fengyuan Yu 等NeurIPS 2024 · 被引用 24 次
- End-To-End Causal Effect Estimation from Unstructured Natural Language DataNikita Dhawan, Leonardo Cotta, Karen Ullrich, Rahul G. Krishnan 等NeurIPS 2024 · 被引用 24 次
- CE-RCFR: Robust Counterfactual Regression for Consensus-Enabled Treatment Effect EstimationFan Wang, Chaochao Chen, Weiming Liu, Tianhao Fan 等KDD 2024 · 被引用 18 次
相关 Paper
- Partial Identification under High-Dimensional Potential Outcomes and Confounders via Optimal TransportYunfeng Wang, Zhiheng Zhang, Zijun GaoICML 2026
- Proximity Matters: Local Proximity Enhanced Balancing for Treatment Effect EstimationHao Wang, Zhichao Chen, Zhaoran Liu, Xu Chen 等KDD 2025 · 被引用 4 次
- The Paradox of Outcome Optimization: A Causal Information-Theoretic Bound on Reasoning Shortcuts in LLMsZihan Chen, Yiming Zhang, Wenxiang Geng, Zenghui Ding 等ACL 2026
- Learning Treatment Representations for Downstream Instrumental Variable RegressionShiangyi Lin, Hui Lan, Vasilis SyrgkanisICML 2026
- LLM-Driven Treatment Effect Estimation Under Inference Time Text ConfoundingYuchen Ma, Dennis Frauen, Jonas Schweisthal, Stefan FeuerriegelNeurIPS 2025 · 被引用 7 次
