Infusing Lattice Symmetry Priors in Attention Mechanisms for Sample-Efficient Abstract Geometric Reasoning
Mattia Atzeni, Mrinmaya Sachan, Andreas Loukas
Abstract
The Abstraction and Reasoning Corpus (ARC) (Chollet, 2019) and its most recent language-complete instantiation (LARC) has been postulated as an important step towards general AI. Yet, even state-of-the-art machine learning models struggle to achieve meaningful performance on these problems, falling behind non-learning based approaches. We argue that solving these tasks requires extreme generalization that can only be achieved by proper accounting for core knowledge priors. As a step towards this goal, we focus on geometry priors and introduce LatFormer, a model that incorporates lattice symmetry priors in attention masks. We show that, for any transformation of the hypercubic lattice, there exists a binary attention mask that implements that group action. Hence, our study motivates a modification to the standard attention mechanism, where attention weights are scaled using soft masks generated by a convolutional network. Experiments on synthetic geometric reasoning show that LatFormer requires 2 orders of magnitude fewer data than standard attention and transformers. Moreover, our results on ARC and LARC tasks that incorporate geometric priors provide preliminary evidence that these complex datasets do not lie out of the reach of deep learning models.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 8ca64d59-509c-42a4-b862-37a78b584b2aCited by top-tier papers1
Ask how each one uses itBuilds on8
- On the Relationship between Self-Attention and Convolutional LayersJean-Baptiste Cordonnier, Andreas Loukas, Martin JaggiICLR 2020 · 629 citations
- COTR: Correspondence Transformer for Matching Across ImagesWei Jiang, Eduard Trulls, Jan Hosang, Andrea Tagliasacchi et al.ICCV 2021 · 318 citations
- CoMIR: Contrastive Multimodal Image Representation for RegistrationNicolas Pielawski, Elisabeth Wetzer, Johan Öfverstedt, Jiahao Lu et al.NeurIPS 2020 · 110 citations
- SPECTRE: Spectral Conditioning Helps to Overcome the Expressivity Limits of One-shot Graph GeneratorsKarolis Martinkus, Andreas Loukas, Nathanaël Perraudin, Roger WattenhoferICML 2022 · 109 citations
- Text-based RL Agents with Commonsense Knowledge: New Challenges, Environments and BaselinesKeerthiram Murugesan, Mattia Atzeni, Pavan Kapanipathi, Pushkar Shukla et al.AAAI 2021 · 60 citations
Related papers
- ARC Is a Vision Problem!Keya Hu, Ali Cy, Linlu Qiu, Xiaoman Delores Ding et al.CVPR 2026 · 23 citations
- The Lattice Representation Hypothesis of Large Language ModelsBo XiongICLR 2026 · 3 citations
- Think Visually, Reason Textually: Vision-Language Synergy in Abstract ReasoningBeichen Zhang, Yuhang Zang, Xiaoyi Dong, Yuhang Cao et al.CVPR 2026
- Platonic Transformers: A Solid Choice For EquivarianceMohammad Mohaiminul Islam, Rishabh Anand, David Wessels, Friso de Kruiff et al.ICML 2026 · 7 citations
- Wyckoff Transformer: Generation of Symmetric CrystalsNikita Kazeev, Wei Nong, Ignat Romanov, Ruiming Zhu et al.ICML 2025
