Structural Causal Bandits under Markov Equivalence
Min Woo Park, Andy Arditi, Elias Bareinboim, Sanghack Lee
Abstract
In decision-making processes, an intelligent agent with causal knowledge can optimize action spaces to avoid unnecessary exploration. A structural causal bandit framework provides guidance on how to prune actions that are unable to maximize reward by leveraging prior knowledge of the underlying causal structure among actions. A key assumption of this framework is that the agent has access to a fully-specified causal diagram representing the target system. In this paper, we extend the structural causal bandits to scenarios where the agent leverages a Markov equivalence class. In such cases, the causal structure is provided to the agent in the form of a partial ancestral graph (PAG). We propose a generalized framework for identifying potentially optimal actions within this graph structure, thereby broadening the applicability of structural causal bandits.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0a6a9702-d4ed-43c9-8a10-26f73a967c36Cited by top-tier papers5
- Counterfactual Structural Causal BanditsMin Woo Park, Sanghack LeeICLR 2026 · 1 citation
- Variance-Reduced Long-Term Rehearsal Learning with Quadratic Programming ReformulationWen-Bo Du, Tian Qin, Tian-Zuo Wang, Zhi-Hua ZhouNeurIPS 2025 · 1 citation
- Enabling Optimal Decisions in Rehearsal Learning under CARE ConditionWen-Bo Du, Hao-Yi Lei, Lue Tao, Tian-Zuo Wang et al.ICML 2025
- Polynomial-Delay MAG Listing with Novel Locally Complete Orientation RulesTian-Zuo Wang, Wen-Bo Du, Zhi-Hua ZhouICML 2025
- On Measuring Influence in Avoiding Undesired FutureLue Tao, Tian-Zuo Wang, Yuan Jiang, Zhi-Hua ZhouICLR 2026
Builds on33
- Causal Discovery in Heterogeneous Environments Under the Sparse Mechanism Shift HypothesisRonan Perry, Julius von Kügelgen, Bernhard SchölkopfNeurIPS 2022 · 84 citations
- Designing Optimal Dynamic Treatment Regimes: A Causal Reinforcement Learning ApproachJunzhe ZhangICML 2020 · 78 citations
- Partial Counterfactual Identification from Observational and Experimental DataJunzhe Zhang, Jin Tian, Elias BareinboimICML 2022 · 77 citations
- Agent Incentives: A Causal PerspectiveTom Everitt, Ryan Carey, Eric D. Langlois, Pedro A. Ortega et al.AAAI 2021 · 66 citations
- Causal Bandits with Unknown Graph StructureYangyi Lu, Amirhossein Meisami, Ambuj TewariNeurIPS 2021 · 53 citations
Related papers
- Causal Identification under Markov equivalence: Calculus, Algorithm, and CompletenessAmin Jaber, Adèle H. Ribeiro, Jiji Zhang, Elias BareinboimNeurIPS 2022 · 36 citations
- Estimating Identifiable Causal Effects on Markov Equivalence Class through Double Machine LearningYonghan Jung, Jin Tian, Elias BareinboimICML 2021 · 21 citations
- Approximate Allocation Matching for Structural Causal Bandits with Unobserved ConfoundersLai Wei, Muhammad Qasim Elahi, Mahsa Ghasemi, Murat KocaogluNeurIPS 2023 · 10 citations
- Non-Stationary Structural Causal BanditsYeahoon Kwon, Yesong Choe, Soungmin Park, Neil Dhir et al.NeurIPS 2025
- Additive Causal Bandits with Unknown GraphAlan Malek, Virginia Aglietti, Silvia ChiappaICML 2023 · 11 citations
