Conditional Mutual Information for Disentangled Representations in Reinforcement Learning
Mhairi Dunion, Trevor McInroe, Kevin Sebastian Luck, Josiah Hanna, Stefano V. Albrecht
摘要
Reinforcement Learning (RL) environments can produce training data with spurious correlations between features due to the amount of training data or its limited feature coverage. This can lead to RL agents encoding these misleading correlations in their latent representation, preventing the agent from generalising if the correlation changes within the environment or when deployed in the real world. Disentangled representations can improve robustness, but existing disentanglement techniques that minimise mutual information between features require independent features, thus they cannot disentangle correlated features. We propose an auxiliary task for RL algorithms that learns a disentangled representation of high-dimensional observations with correlated features by minimising the conditional mutual information between features in the representation. We demonstrate experimentally, using continuous control tasks, that our approach improves generalisation under correlation shifts, as well as improving the training performance of RL algorithms in the presence of correlated features.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper15
- DRED: Zero-Shot Transfer in Reinforcement Learning via Data-Regularised Environment DesignSamuel Garcin, James Doran, Shangmin Guo, Christopher G. Lucas 等ICML 2024 · 被引用 14 次
- Enhancing Tactile-based Reinforcement Learning for Robotic ControlElle Miller, Trevor McInroe, David Abel, Oisin Mac Aodha 等NeurIPS 2025 · 被引用 9 次
- Learning Disentangled Representations for Perceptual Point Cloud Quality Assessment via Mutual Information MinimizationZiyu Shan, Yujie Zhang, Yipeng Liu, Yiling XuNeurIPS 2024 · 被引用 7 次
- Skill-aware Mutual Information Optimisation for Zero-shot Generalisation in Reinforcement LearningXuehui Yu, Mhairi Dunion, Xin Li, Stefano V. AlbrechtNeurIPS 2024 · 被引用 6 次
- PvP: Data-Efficient Humanoid Robot Learning with Proprioceptive-Privileged Contrastive RepresentationsMingqi Yuan, Tao Yu, Haolin Song, Bo Li 等CVPR 2026 · 被引用 3 次
它引用的顶会 Paper18
- CURL: Contrastive Unsupervised Representations for Reinforcement LearningMichael Laskin, Aravind Srinivas, Pieter AbbeelICML 2020 · 被引用 1,261 次
- Image Augmentation Is All You Need: Regularizing Deep Reinforcement Learning from PixelsDenis Yarats, Ilya Kostrikov, Rob FergusICLR 2021 · 被引用 911 次
- Reinforcement Learning with Augmented DataMichael Laskin, Kimin Lee, Adam Stooke, Lerrel Pinto 等NeurIPS 2020 · 被引用 833 次
- Weakly-Supervised Disentanglement Without CompromisesFrancesco Locatello, Ben Poole, Gunnar Rätsch, Bernhard Schölkopf 等ICML 2020 · 被引用 361 次
- Stabilizing Deep Q-Learning with ConvNets and Vision Transformers under Data AugmentationNicklas Hansen, Hao Su, Xiaolong WangNeurIPS 2021 · 被引用 189 次
相关 Paper
- Temporal Disentanglement of Representations for Improved Generalisation in Reinforcement LearningMhairi Dunion, Trevor McInroe, Kevin Sebastian Luck, Josiah P. Hanna 等ICLR 2023 · 被引用 4 次
- Domain-Robust Visual Imitation Learning with Mutual Information ConstraintsEdoardo Cetin, Oya ÇeliktutanICLR 2021 · 被引用 4 次
- Zero Shot Generalization of Vision-Based RL Without Data AugmentationSumeet Batra, Gaurav S. SukhatmeICML 2025
- CATAL: Causally Disentangled Task Representation Learning for Offline Meta-Reinforcement LearningShan Cong, Chao Yu, Xiangyuan LanAAAI 2026
- On Disentangled Representations Learned from Correlated DataFrederik Träuble, Elliot Creager, Niki Kilbertus, Francesco Locatello 等ICML 2021 · 被引用 38 次
