EqR: Equivariant Representations for Data-Efficient Reinforcement Learning
Arnab Kumar Mondal, Vineet Jain, Kaleem Siddiqi, Siamak Ravanbakhsh
Abstract
We study a variety of notions of equivariance as an inductive bias in Reinforcement Learning (RL). In particular, we propose new mechanisms for learning representations that are equivariant to both the agent's action, as well as symmetry transformations of the state-action pairs. Whereas prior work on exploiting symmetries in deep RL can only incorporate predefined linear transformations, our approach allows non-linear symmetry transformations of state-action pairs to be learned from the data. This is achieved through 1) equivariant Lie algebraic parameterization of state and action encodings, 2) equivariant latent transition models, and 3) the incorporation of symmetrybased losses. We demonstrate the advantages of our method, which we call Equivariant representations for RL (EqR), for Atari games in a data-efficient setting limited to 100K steps of interactions with the environment.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a9c07c6a-f14d-4eed-99d2-0cba540cf8a9Cited by top-tier papers11
- EDGI: Equivariant Diffusion for Planning with Embodied AgentsJohann Brehmer, Joey Bose, Pim de Haan, Taco S. CohenNeurIPS 2023 · 52 citations
- Equivariant Adaptation of Large Pretrained ModelsArnab Kumar Mondal, Siba Smarak Panigrahi, Oumar Kaba, Sai Mudumba et al.NeurIPS 2023 · 49 citations
- Continuous MDP Homomorphisms and Homomorphic Policy GradientSahand Rezaei-Shoshtari, Rosie Zhao, Prakash Panangaden, David Meger et al.NeurIPS 2022 · 34 citations
- A General Theory of Correct, Incorrect, and Extrinsic EquivarianceDian Wang, Xupeng Zhu, Jung Yeon Park, Mingxi Jia et al.NeurIPS 2023 · 23 citations
- Structuring Representations Using Group InvariantsMehran Shakerinava, Arnab Kumar Mondal, Siamak RavanbakhshNeurIPS 2022 · 23 citations
Builds on11
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 24,064 citations
- Bootstrap Your Own Latent - A New Approach to Self-Supervised LearningJean-Bastien Grill, Florian Strub, Florent Altché, Corentin Tallec et al.NeurIPS 2020 · 9,171 citations
- Barlow Twins: Self-Supervised Learning via Redundancy ReductionJure Zbontar, Li Jing, Ishan Misra, Yann LeCun et al.ICML 2021 · 2,942 citations
- Dream to Control: Learning Behaviors by Latent ImaginationDanijar Hafner, Timothy P. Lillicrap, Jimmy Ba, Mohammad NorouziICLR 2020 · 1,852 citations
- CURL: Contrastive Unsupervised Representations for Reinforcement LearningMichael Laskin, Aravind Srinivas, Pieter AbbeelICML 2020 · 1,261 citations
Related papers
- The Surprising Effectiveness of Equivariant Models in Domains with Latent SymmetryDian Wang, Jung Yeon Park, Neel Sortur, Lawson L. S. Wong et al.ICLR 2023 · 2 citations
- Meta-learning Symmetries by ReparameterizationAllan Zhou, Tom Knowles, Chelsea FinnICLR 2021 · 105 citations
- Residual Pathway Priors for Soft Equivariance ConstraintsMarc Finzi, Greg Benton, Andrew Gordon WilsonNeurIPS 2021 · 89 citations
- MDP Homomorphic Networks: Group Symmetries in Reinforcement LearningElise van der Pol, Daniel E. Worrall, Herke van Hoof, Frans A. Oliehoek et al.NeurIPS 2020 · 203 citations
- Learning Symmetric Embeddings for Equivariant World ModelsJung Yeon Park, Ondrej Biza, Linfeng Zhao, Jan-Willem van de Meent et al.ICML 2022 · 56 citations
