Residual Pathway Priors for Soft Equivariance Constraints
Marc Finzi, Greg Benton, Andrew Gordon Wilson
Abstract
There is often a trade-off between building deep learning systems that are expressive enough to capture the nuances of the reality, and having the right inductive biases for efficient learning. We introduce Residual Pathway Priors (RPPs) as a method for converting hard architectural constraints into soft priors, guiding models towards structured solutions, while retaining the ability to capture additional complexity. Using RPPs, we construct neural network priors with inductive biases for equivariances, but without limiting flexibility. We show that RPPs are resilient to approximate or misspecified symmetries, and are as effective as fully constrained models even when symmetries are exact. We showcase the broad applicability of RPPs with dynamical systems, tabular data, and reinforcement learning. In Mujoco locomotion tasks, where contact forces and directional rewards violate strict equivariance assumptions, the RPP outperforms baseline model-free RL agents, and also improves the learned transition models for model-based RL.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers34
- Approximately Equivariant Networks for Imperfectly Symmetric DynamicsRui Wang, Robin Walters, Rose YuICML 2022 · 111 citations
- The Importance of Being Scalable: Improving the Speed and Accuracy of Neural Network Interatomic Potentials Across Chemical DomainsEric Qu, Aditi S. KrishnapriyanNeurIPS 2024 · 63 citations
- Approximation-Generalization Trade-offs under (Approximate) Group EquivarianceMircea Petrache, Shubhendu TrivediNeurIPS 2023 · 54 citations
- Learning Partial Equivariances From DataDavid W. Romero, Suhas LohitNeurIPS 2022 · 54 citations
- Relaxing Equivariance Constraints with Non-stationary Continuous FiltersTycho F. A. van der Ouderaa, David W. Romero, Mark van der WilkNeurIPS 2022 · 51 citations
Builds on17
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- CoAtNet: Marrying Convolution and Attention for All Data SizesZihang Dai, Hanxiao Liu, Quoc V. Le, Mingxing TanNeurIPS 2021 · 1,747 citations
- E(n) Equivariant Graph Neural NetworksVictor Garcia Satorras, Emiel Hoogeboom, Max WellingICML 2021 · 1,432 citations
- SE(3)-Transformers: 3D Roto-Translation Equivariant Attention NetworksFabian Fuchs, Daniel E. Worrall, Volker Fischer, Max WellingNeurIPS 2020 · 1,025 citations
- ConViT: Improving Vision Transformers with Soft Convolutional Inductive BiasesStéphane d'Ascoli, Hugo Touvron, Matthew L. Leavitt, Ari S. Morcos et al.ICML 2021 · 1,021 citations
Related papers
- EqR: Equivariant Representations for Data-Efficient Reinforcement LearningArnab Kumar Mondal, Vineet Jain, Kaleem Siddiqi, Siamak RavanbakhshICML 2022 · 32 citations
- The Surprising Effectiveness of Equivariant Models in Domains with Latent SymmetryDian Wang, Jung Yeon Park, Neel Sortur, Lawson L. S. Wong et al.ICLR 2023 · 2 citations
- Deconstructing the Inductive Biases of Hamiltonian Neural NetworksNate Gruver, Marc Anton Finzi, Samuel Don Stanton, Andrew Gordon WilsonICLR 2022 · 50 citations
- MDP Homomorphic Networks: Group Symmetries in Reinforcement LearningElise van der Pol, Daniel E. Worrall, Herke van Hoof, Frans A. Oliehoek et al.NeurIPS 2020 · 203 citations
- Latent Mixture of Symmetries for Sample-Efficient Dynamic LearningHaoran Li, Chenhan Xiao, Muhao Guo, Yang WengNeurIPS 2025 · 7 citations
