Translation Equivariant Transformer Neural Processes
Matthew Ashman, Cristiana Diaconu, Junhyuck Kim, Lakee Sivaraya, Stratis Markou, James Requeima, Wessel P. Bruinsma, Richard E. Turner
Abstract
The effectiveness of neural processes (NPs) in modelling posterior prediction maps-the mapping from data to posterior predictive distributions-has significantly improved since their inception. This improvement can be attributed to two principal factors: (1) advancements in the architecture of permutation invariant set functions, which are intrinsic to all NPs; and (2) leveraging symmetries present in the true posterior predictive map, which are problem dependent. Transformers are a notable development in permutation invariant set functions, and their utility within NPs has been demonstrated through the family of models we refer to as transformer neural processes (TNPs). Despite significant interest in TNPs, little attention has been given to incorporating symmetries. Notably, the posterior prediction maps for data that are stationary-a common assumption in spatiotemporal modelling-exhibit translation equivariance. In this paper, we introduce of a new family of translation equivariant TNPs (TE-TNPs) that incorporate translation equivariance. Through an extensive range of experiments on synthetic and real-world spatio-temporal data, we demonstrate the effectiveness of TE-TNPs relative to their nontranslation-equivariant counterparts and other NP baselines.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers8
- Estimating Interventional Distributions with Uncertain Causal Graphs through Meta-LearningAnish Dhir, Cristiana Diaconu, Valentinian Lungu, James Requeima et al.NeurIPS 2025 · 16 citations
- Approximately Equivariant Neural ProcessesMatthew Ashman, Cristiana Diaconu, Adrian Weller, Wessel P. Bruinsma et al.NeurIPS 2024 · 11 citations
- Spectral Convolutional Conditional Neural ProcessesPeiman Mohseni, Nick DuffieldNeurIPS 2025 · 10 citations
- ALINE: Joint Amortization for Bayesian Inference and Active Data AcquisitionDaolang Huang, Xinyi Wen, Ayush Bharti, Samuel Kaski et al.NeurIPS 2025 · 8 citations
- Test Time Scaling for Neural ProcessesHyungi Lee, Moonseok Choi, Hyunsu Kim, Kyunghyun Cho et al.NeurIPS 2025 · 1 citation
Builds on12
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- E(n) Equivariant Graph Neural NetworksVictor Garcia Satorras, Emiel Hoogeboom, Max WellingICML 2021 · 1,432 citations
- Transformers Can Do Bayesian InferenceSamuel Müller, Noah Hollmann, Sebastian Pineda-Arango, Josif Grabocka et al.ICLR 2022 · 287 citations
- Convolutional Conditional Neural ProcessesJonathan Gordon, Wessel P. Bruinsma, Andrew Y. K. Foong, James Requeima et al.ICLR 2020 · 200 citations
- Transformer Neural Processes: Uncertainty-Aware Meta Learning Via Sequence ModelingTung Nguyen, Aditya GroverICML 2022 · 148 citations
Related papers
- Meta-Learning Stationary Stochastic Process Prediction with Convolutional Neural ProcessesAndrew Y. K. Foong, Wessel P. Bruinsma, Jonathan Gordon, Yann Dubois et al.NeurIPS 2020 · 96 citations
- Gridded Transformer Neural Processes for Spatio-Temporal DataMatthew Ashman, Cristiana Diaconu, Eric Langezaal, Adrian Weller et al.ICML 2025
- Revisiting Neural Processes via Fourier Transform and Volterra SeriesPeiman Mohseni, Nick Duffield, Raymond K WongICML 2026
- Practical Equivariances via Relational Conditional Neural ProcessesDaolang Huang, Manuel Haussmann, Ulpu Remes, S. T. John et al.NeurIPS 2023 · 14 citations
- Group Equivariant Conditional Neural ProcessesMakoto Kawano, Wataru Kumagai, Akiyoshi Sannai, Yusuke Iwasawa et al.ICLR 2021 · 22 citations
