Conditional Mutual Information for Disentangled Representations in Reinforcement Learning
Mhairi Dunion, Trevor McInroe, Kevin Sebastian Luck, Josiah Hanna, Stefano V. Albrecht
Abstract
Reinforcement Learning (RL) environments can produce training data with spurious correlations between features due to the amount of training data or its limited feature coverage. This can lead to RL agents encoding these misleading correlations in their latent representation, preventing the agent from generalising if the correlation changes within the environment or when deployed in the real world. Disentangled representations can improve robustness, but existing disentanglement techniques that minimise mutual information between features require independent features, thus they cannot disentangle correlated features. We propose an auxiliary task for RL algorithms that learns a disentangled representation of high-dimensional observations with correlated features by minimising the conditional mutual information between features in the representation. We demonstrate experimentally, using continuous control tasks, that our approach improves generalisation under correlation shifts, as well as improving the training performance of RL algorithms in the presence of correlated features.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 6bea91aa-37ba-4996-8b13-c97b5946a3e7Cited by top-tier papers15
- DRED: Zero-Shot Transfer in Reinforcement Learning via Data-Regularised Environment DesignSamuel Garcin, James Doran, Shangmin Guo, Christopher G. Lucas et al.ICML 2024 · 14 citations
- Enhancing Tactile-based Reinforcement Learning for Robotic ControlElle Miller, Trevor McInroe, David Abel, Oisin Mac Aodha et al.NeurIPS 2025 · 9 citations
- Learning Disentangled Representations for Perceptual Point Cloud Quality Assessment via Mutual Information MinimizationZiyu Shan, Yujie Zhang, Yipeng Liu, Yiling XuNeurIPS 2024 · 7 citations
- Skill-aware Mutual Information Optimisation for Zero-shot Generalisation in Reinforcement LearningXuehui Yu, Mhairi Dunion, Xin Li, Stefano V. AlbrechtNeurIPS 2024 · 6 citations
- PvP: Data-Efficient Humanoid Robot Learning with Proprioceptive-Privileged Contrastive RepresentationsMingqi Yuan, Tao Yu, Haolin Song, Bo Li et al.CVPR 2026 · 3 citations
Builds on18
- CURL: Contrastive Unsupervised Representations for Reinforcement LearningMichael Laskin, Aravind Srinivas, Pieter AbbeelICML 2020 · 1,261 citations
- Image Augmentation Is All You Need: Regularizing Deep Reinforcement Learning from PixelsDenis Yarats, Ilya Kostrikov, Rob FergusICLR 2021 · 911 citations
- Reinforcement Learning with Augmented DataMichael Laskin, Kimin Lee, Adam Stooke, Lerrel Pinto et al.NeurIPS 2020 · 833 citations
- Weakly-Supervised Disentanglement Without CompromisesFrancesco Locatello, Ben Poole, Gunnar Rätsch, Bernhard Schölkopf et al.ICML 2020 · 361 citations
- Stabilizing Deep Q-Learning with ConvNets and Vision Transformers under Data AugmentationNicklas Hansen, Hao Su, Xiaolong WangNeurIPS 2021 · 189 citations
Related papers
- Temporal Disentanglement of Representations for Improved Generalisation in Reinforcement LearningMhairi Dunion, Trevor McInroe, Kevin Sebastian Luck, Josiah P. Hanna et al.ICLR 2023 · 4 citations
- Domain-Robust Visual Imitation Learning with Mutual Information ConstraintsEdoardo Cetin, Oya ÇeliktutanICLR 2021 · 4 citations
- Zero Shot Generalization of Vision-Based RL Without Data AugmentationSumeet Batra, Gaurav S. SukhatmeICML 2025
- CATAL: Causally Disentangled Task Representation Learning for Offline Meta-Reinforcement LearningShan Cong, Chao Yu, Xiangyuan LanAAAI 2026
- On Disentangled Representations Learned from Correlated DataFrederik Träuble, Elliot Creager, Niki Kilbertus, Francesco Locatello et al.ICML 2021 · 38 citations
