Adjustment for Confounding using Pre-Trained Representations
Rickmer Schulte, David Rügamer, Thomas Nagler
Abstract
There is growing interest in extending average treatment effect (ATE) estimation to incorporate non-tabular data, such as images and text, which may act as sources of confounding. Neglecting these effects risks biased results and flawed scientific conclusions. However, incorporating nontabular data necessitates sophisticated feature extractors, often in combination with ideas of transfer learning. In this work, we investigate how latent features from pre-trained neural networks can be leveraged to adjust for sources of confounding. We formalize conditions under which these latent features enable valid adjustment and statistical inference in ATE estimation, demonstrating results along the example of double machine learning. We discuss critical challenges inherent to latent feature learning and downstream parameter estimation arising from the high dimensionality and non-identifiability of representations. Common structural assumptions for obtaining fast convergence rates with additive or sparse linear models are shown to be unrealistic for latent features. We argue, however, that neural networks are largely insensitive to these issues. In particular, we show that neural networks can achieve fast convergence rates by adapting to intrinsic notions of sparsity and dimension of the learning problem.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e692a49a-ecc5-440c-b3d5-921105de954dCited by top-tier papers1
Ask how each one uses itBuilds on6
- The Intrinsic Dimension of Images and Its Impact on LearningPhillip Pope, Chen Zhu, Ahmed Abdelkader, Micah Goldblum et al.ICLR 2021 · 381 citations
- Proximal Causal Learning with Kernels: Two-Stage Estimation and Moment RestrictionAfsaneh Mastouri, Yuchen Zhu, Limor Gultchin, Anna Korba et al.ICML 2021 · 78 citations
- RieszNet and ForestRiesz: Automatic Debiased Machine Learning with Neural Nets and Random ForestsVictor Chernozhukov, Whitney Newey, Victor Quintas-Martinez, Vasilis SyrgkanisICML 2022 · 61 citations
- End-To-End Causal Effect Estimation from Unstructured Natural Language DataNikita Dhawan, Leonardo Cotta, Karen Ullrich, Rahul G. Krishnan et al.NeurIPS 2024 · 24 citations
- The Effect of Intrinsic Dataset Properties on Generalization: Unraveling Learning Differences Between Natural and Medical ImagesNicholas Konz, Maciej A. MazurowskiICLR 2024 · 15 citations
Related papers
- A Neural Mean Embedding Approach for Back-door and Front-door AdjustmentLiyuan Xu, Arthur GrettonICLR 2023
- Bounds on Representation-Induced Confounding Bias for Treatment Effect EstimationValentyn Melnychuk, Dennis Frauen, Stefan FeuerriegelICLR 2024 · 23 citations
- Deep Multi-Modal Structural Equations For Causal Effect Estimation With Unstructured ProxiesShachi Deshpande, Kaiwen Wang, Dhruv Sreenivas, Zheng Li et al.NeurIPS 2022 · 15 citations
- Coordinated Double Machine LearningNitai Fingerhut, Matteo Sesia, Yaniv RomanoICML 2022 · 5 citations
- LLM-Driven Treatment Effect Estimation Under Inference Time Text ConfoundingYuchen Ma, Dennis Frauen, Jonas Schweisthal, Stefan FeuerriegelNeurIPS 2025 · 7 citations
