ACAMDA: Improving Data Efficiency in Reinforcement Learning through Guided Counterfactual Data Augmentation
Yuewen Sun, Erli Wang, Biwei Huang, Chaochao Lu, Lu Feng, Changyin Sun, Kun Zhang
Abstract
Data augmentation plays a crucial role in improving the data efficiency of reinforcement learning (RL). However, the generation of high-quality augmented data remains a significant challenge. To overcome this, we introduce ACAMDA (Adversarial Causal Modeling for Data Augmentation), a novel framework that integrates two causality-based tasks: causal structure recovery and counterfactual estimation. The unique aspect of ACAMDA lies in its ability to recover temporal causal relationships from limited non-expert datasets. The identification of the sequential cause-and-effect allows the creation of realistic yet unobserved scenarios. We utilize this characteristic to generate guided counterfactual datasets, which, in turn, substantially reduces the need for extensive data collection. By simulating various state-action pairs under hypothetical actions, ACAMDA enriches the training dataset for diverse and heterogeneous conditions. Our experimental evaluation shows that ACAMDA outperforms existing methods, particularly when applied to novel and unseen domains.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 571faf6f-1e81-4ecc-9c78-741b6980a884Cited by top-tier papers4
- Identifying Latent State-Transition Processes for Individualized Reinforcement LearningYuewen Sun, Biwei Huang, Yu Yao, Donghuo Zeng et al.NeurIPS 2024 · 2 citations
- Causal Information Prioritization for Efficient Reinforcement LearningHongye Cao, Fan Feng, Tianpei Yang, Jing Huo et al.ICLR 2025 · 1 citation
- CoBA: Counterbias Text Augmentation for Mitigating Various Spurious Correlations via Semantic TriplesKyohoon Jin, Juhwan Choi, Jungmin Yun, Junho Lee et al.EMNLP 2025
- Reward-Preserving Counterfactual State Editing for Offline Reinforcement LearningSiyu Wang, Xiaocong Chen, Mingming Gong, Yong Li et al.ICML 2026
Builds on9
- AugMix: A Simple Data Processing Method to Improve Robustness and UncertaintyDan Hendrycks, Norman Mu, Ekin Dogus Cubuk, Barret Zoph et al.ICLR 2020 · 1,572 citations
- Reinforcement Learning with Augmented DataMichael Laskin, Kimin Lee, Adam Stooke, Lerrel Pinto et al.NeurIPS 2020 · 833 citations
- Counterfactual Data Augmentation using Locally Factored DynamicsSilviu Pitis, Elliot Creager, Animesh GargNeurIPS 2020 · 126 citations
- Transient Non-stationarity and Generalisation in Deep Reinforcement LearningMaximilian Igl, Gregory Farquhar, Jelena Luketina, Wendelin Boehmer et al.ICLR 2021 · 104 citations
- Temporally Disentangled Representation LearningWeiran Yao, Guangyi Chen, Kun ZhangNeurIPS 2022 · 84 citations
Related papers
- MARLIN: Multi-Agent Reinforcement Learning for Incremental DAG DiscoveryDong Li, Zhengzhang Chen, Xujiang Zhao, Linlin Yu et al.AAAI 2026
- Learning from Counterfactual Links for Link PredictionTong Zhao, Gang Liu, Daheng Wang, Wenhao Yu et al.ICML 2022 · 127 citations
- Counterfactual Bootstrap for Robust Meta-Reinforcement LearningAi Bo, Junzhe Zhang, M. Cenk GursoyICML 2026
- Understanding when Dynamics-Invariant Data Augmentations Benefit Model-free Reinforcement Learning UpdatesNicholas Corrado, Josiah P. HannaICLR 2024 · 6 citations
- Improving Deepfake Detection with Reinforcement Learning-Based Adaptive Data AugmentationYuxuan Chou, Tao Yu, Wen Huang, Yuheng Zhang et al.AAAI 2026
