Stochastic Online Instrumental Variable Regression: Regrets for Endogeneity and Bandit Feedback
Riccardo Della Vecchia, Debabrota Basu
Abstract
Endogeneity, i.e. the dependence of noise and covariates, is a common phenomenon in real data due to omitted variables, strategic behaviours, measurement errors etc. In contrast, the existing analyses of stochastic online linear regression with unbounded noise and linear bandits depend heavily on exogeneity, i.e. the independence of noise and covariates. Motivated by this gap, we study the over-and just-identified Instrumental Variable (IV) regression, specifically Two-Stage Least Squares, for stochastic online learning, and propose to use an online variant of Two-Stage Least Squares, namely O2SLS. We show that O2SLS achieves O(d x d z log 2 T ) identification and O(γ √ d z T ) oracle regret after T interactions, where d x and d z are the dimensions of covariates and IVs, and γ is the bias due to endogeneity. For γ = 0, i.e. under exogeneity, O2SLS exhibits O(d 2 x log 2 T ) oracle regret, which is of the same order as that of the stochastic online ridge. Then, we leverage O2SLS as an oracle to design OFUL-IV, a stochastic linear bandit algorithm to tackle endogeneity. OFUL-IVyields O( √ d x d z T ) regret that matches the regret lower bound under exogeneity. For different datasets with endogeneity, we experimentally show efficiencies of O2SLS and OFUL-IV.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 55e62e74-448c-4985-9adf-beabcdfa6cf1Cited by top-tier papers6
- Stochastic Optimization Algorithms for Instrumental Variable Regression with Streaming DataXuxing Chen, Abhishek Roy, Yifan Hu, Krishnakumar BalasubramanianNeurIPS 2024 · 4 citations
- A Unifying View of Coverage in Linear Off-policy EvaluationPhilip Amortila, Audrey Huang, Akshay Krishnamurthy, Nan JiangICLR 2026 · 2 citations
- Generator-Mediated Bandits: Thompson Sampling for GenAI-Powered Adaptive InterventionsMarc Brooks, Gabriel Durham, Kihyuk Hong, Ambuj TewariNeurIPS 2025 · 1 citation
- Differentially Private Two-Stage Gradient Descent for Instrumental Variable RegressionHaodong Liang, Yanhao Jin, Krishna Balasubramanian, Lifeng LaiICLR 2026
- Efficient Adaptive Experimentation with NoncomplianceMiruna Oprescu, Brian Cho, Nathan KallusNeurIPS 2025
Builds on4
- Beyond UCB: Optimal and Efficient Contextual Bandits with Regression OraclesDylan J. Foster, Alexander RakhlinICML 2020 · 241 citations
- Strategic Instrumental Variable Regression: Recovering Causal Relationships From Strategic ResponsesKeegan Harris, Dung Daniel T. Ngo, Logan Stapleton, Hoda Heidari et al.ICML 2022 · 37 citations
- Leveraging Good Representations in Linear Contextual BanditsMatteo Papini, Andrea Tirinzoni, Marcello Restelli, Alessandro Lazaric et al.ICML 2021 · 35 citations
- Stochastic Online Linear Regression: the Forward Algorithm to Replace RidgeReda Ouhamma, Odalric-Ambrym Maillard, Vianney PerchetNeurIPS 2021 · 18 citations
Related papers
- Instrumental Variable Regression with Confounder BalancingAnpeng Wu, Kun Kuang, Bo Li, Fei WuICML 2022 · 31 citations
- Transformers Handle Endogeneity in In-Context Linear RegressionHaodong Liang, Krishna Balasubramanian, Lifeng LaiICLR 2025
- Conditional Instrumental Variable Regression with Representation Learning for Causal InferenceDebo Cheng, Ziqi Xu, Jiuyong Li, Lin Liu et al.ICLR 2024 · 14 citations
- Dual Instrumental Variable RegressionKrikamol Muandet, Arash Mehrjou, Si Kai Lee, Anant RajNeurIPS 2020 · 87 citations
- Learning Decision Policies with Instrumental Variables through Double Machine LearningDaqian Shao, Ashkan Soleymani, Francesco Quinzan, Marta KwiatkowskaICML 2024 · 4 citations
