Provably Efficient Neural Estimation of Structural Equation Models: An Adversarial Approach
Luofeng Liao, You-Lin Chen, Zhuoran Yang, Bo Dai, Mladen Kolar, Zhaoran Wang
摘要
Structural equation models (SEMs) are widely used in sciences, ranging from economics to psychology, to uncover causal relationships underlying a complex system under consideration and estimate structural parameters of interest. We study estimation in a class of generalized SEMs where the object of interest is defined as the solution to a linear operator equation. We formulate the linear operator equation as a min-max game, where both players are parameterized by neural networks (NNs), and learn the parameters of these neural networks using the stochastic gradient descent. We consider both 2-layer and multi-layer NNs with ReLU activation functions and prove global convergence in an overparametrized regime, where the number of neurons is diverging. The results are established using techniques from online learning and local linearization of NNs, and improve in several aspects the current state-of-the-art. For the first time we provide a tractable estimation procedure for SEMs based on NNs with provable convergence and without the need for sample splitting. 34th Conference on Neural Information Processing Systems (NeurIPS 2020),
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper7
- Deep Proxy Causal Learning and its Application to Confounded Bandit Policy EvaluationLiyuan Xu, Heishiro Kanagawa, Arthur GrettonNeurIPS 2021 · 被引用 52 次
- Learning Causal Models from Conditional Moment Restrictions by Importance WeightingMasahiro Kato, Masaaki Imaizumi, Kenichiro McAlinn, Shota Yasui 等ICLR 2022 · 被引用 7 次
- Demystifying Spectral Feature Learning for Instrumental Variable RegressionDimitri Meunier, Antoine Moulin, Jakub Wornbard, Vladimir Kostic 等NeurIPS 2025 · 被引用 5 次
- Targeted Sequential Indirect Experiment DesignElisabeth Ailer, Niclas Dern, Jason S. Hartford, Niki KilbertusNeurIPS 2024 · 被引用 4 次
- Stochastic Optimization Algorithms for Instrumental Variable Regression with Streaming DataXuxing Chen, Abhishek Roy, Yifan Hu, Krishnakumar BalasubramanianNeurIPS 2024 · 被引用 4 次
它引用的顶会 Paper3
- Neural Policy Gradient Methods: Global Optimality and Rates of ConvergenceLingxiao Wang, Qi Cai, Zhuoran Yang, Zhaoran WangICLR 2020 · 被引用 270 次
- Minimax Estimation of Conditional Moment ModelsNishanth Dikkala, Greg Lewis, Lester Mackey, Vasilis SyrgkanisNeurIPS 2020 · 被引用 125 次
- A Finite-Time Analysis of Q-Learning with Neural Network Function ApproximationPan Xu, Quanquan GuICML 2020 · 被引用 79 次
相关 Paper
- A global convergence theory for deep ReLU implicit networks via over-parameterizationTianxiang Gao, Hailiang Liu, Jia Liu, Hridesh Rajan 等ICLR 2022 · 被引用 21 次
- Solving Neural Min-Max Games: The Role of Architecture, Initialization & DynamicsDeep Patel, Emmanouil-Vasileios Vlatakis-GkaragkounisNeurIPS 2025 · 被引用 1 次
- On Learnability via Gradient Method for Two-Layer ReLU Neural Networks in Teacher-Student SettingShunta Akiyama, Taiji SuzukiICML 2021 · 被引用 16 次
- Optimal Rates for Averaged Stochastic Gradient Descent under Neural Tangent Kernel RegimeAtsushi Nitanda, Taiji SuzukiICLR 2021 · 被引用 49 次
- DNA-SE: Towards Deep Neural-Nets Assisted Semiparametric EstimationQinshuo Liu, Zixin Wang, Xi-An Li, Xinyao Ji 等ICML 2024
