Training a Resilient Q-network against Observational Interference
Chao-Han Huck Yang, I-Te Danny Hung, Yi Ouyang, Pin-Yu Chen
Abstract
Deep reinforcement learning (DRL) has demonstrated impressive performance in various gaming simulators and real-world applications. In practice, however, a DRL agent may receive faulty observation by abrupt interferences such as black-out, frozen-screen, and adversarial perturbation. How to design a resilient DRL algorithm against these rare but mission-critical and safety-crucial scenarios is an essential yet challenging task. In this paper, we consider a deep q-network (DQN) framework training with an auxiliary task of observational interferences such as artificial noises. Inspired by causal inference for observational interference, we propose a causal inference based DQN algorithm called causal inference Q-network (CIQ). We evaluate the performance of CIQ in several benchmark DQN environments with different types of interferences as auxiliary labels. Our experimental results show that the proposed CIQ method could achieve higher performance and more resilience against observational interferences.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext dc39bccc-27a1-41d2-ad8a-646cd881ea0eCited by top-tier papers3
- Selector-Enhancer: Learning Dynamic Selection of Local and Non-local Attention Operation for Speech EnhancementXinmeng Xu, Weiping Tu, Yuhong YangAAAI 2023 · 8 citations
- BERRY: Bit Error Robustness for Energy-Efficient Reinforcement Learning-Based Autonomous SystemsZishen Wan, Nandhini Chandramoorthy, Karthik Swaminathan, Pin-Yu Chen et al.DAC 2023 · 7 citations
- Deconfounding Actor-Critic Network with Policy Adaptation for Dynamic Treatment RegimesChangchang Yin, Ruoqi Liu, Jeffrey M. Caterino, Ping ZhangKDD 2022 · 5 citations
Builds on9
- Explainable Reinforcement Learning through a Causal LensPrashan Madumal, Tim Miller, Liz Sonenberg, Frank VetereAAAI 2020 · 408 citations
- Behaviour Suite for Reinforcement LearningIan Osband, Yotam Doron, Matteo Hessel, John Aslanides et al.ICLR 2020 · 204 citations
- Invariant Causal Prediction for Block MDPsAmy Zhang, Clare Lyle, Shagun Sodhani, Angelos Filos et al.ICML 2020 · 153 citations
- A Causal View on Robustness of Neural NetworksCheng Zhang, Kun Zhang, Yingzhen LiNeurIPS 2020 · 103 citations
- Off-Policy Evaluation in Partially Observable EnvironmentsGuy Tennenholtz, Uri Shalit, Shie MannorAAAI 2020 · 91 citations
Related papers
- On Minimizing Adversarial Counterfactual Error in Adversarial Reinforcement LearningRoman Belaire, Arunesh Sinha, Pradeep VarakanthamICLR 2025
- Robust Reinforcement Learning on State Observations with Learned Optimal AdversaryHuan Zhang, Hongge Chen, Duane S. Boning, Cho-Jui HsiehICLR 2021 · 212 citations
- Confounding Robust Deep Reinforcement Learning: A Causal ApproachMingxuan Li, Junzhe Zhang, Elias BareinboimNeurIPS 2025 · 7 citations
- Improve Robustness of Reinforcement Learning against Observation Perturbations via l∞ Lipschitz Policy NetworksBuqing Nie, Jingtian Ji, Yangqing Fu, Yue GaoAAAI 2024 · 10 citations
- Belief-Enriched Pessimistic Q-Learning against Adversarial State PerturbationsXiaolin Sun, Zizhan ZhengICLR 2024 · 4 citations
