Geometric Hyena Networks for Large-scale Equivariant Learning
Artem Moskalev, Mangal Prakash, Junjie Xu, Tianyu Cui, Rui Liao, Tommaso Mansi
Abstract
Processing global geometric context while preserving equivariance is crucial when modeling biological, chemical, and physical systems. Yet, this is challenging due to the computational demands of equivariance and global context at scale. Standard methods such as equivariant self-attention suffer from quadratic complexity, while local methods such as distance-based message passing sacrifice global information. Inspired by the recent success of state-space and long-convolutional models, we introduce Geometric Hyena, the first equivariant long-convolutional model for geometric systems. Geometric Hyena captures global geometric context at sub-quadratic complexity while maintaining equivariance to rotations and translations. Evaluated on all-atom property prediction of large RNA molecules and full protein molecular dynamics, Geometric Hyena outperforms existing equivariant models while requiring significantly less memory and compute that equivariant self-attention. Notably, our model processes the geometric context of 30k tokens 20× faster than the equivariant transformer and allows 72× longer context within the same budget.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 9420a581-e7b4-4a98-a376-fc279aef9bf3Cited by top-tier papers4
- Platonic Transformers: A Solid Choice For EquivarianceMohammad Mohaiminul Islam, Rishabh Anand, David Wessels, Friso de Kruiff et al.ICML 2026 · 7 citations
- BioBO: Biology-informed Bayesian Optimization for Perturbation DesignYanke Li, Tianyu Cui, Tommaso Mansi, Mangal Prakash et al.ICLR 2026 · 2 citations
- DualEqui: A Dual-Space Hierarchical Equivariant Network for Large BiomoleculesJunjie Xu, Jiahao Zhang, Mangal Prakash, Xiang Zhang et al.NeurIPS 2025 · 2 citations
- Erwin: A Tree-based Hierarchical Transformer for Large-scale Physical SystemsMaksim Zhdanov, Max Welling, Jan-Willem van de MeentICML 2025
Builds on35
- FlashAttention: Fast and Memory-Efficient Exact Attention with IO-AwarenessTri Dao, Daniel Y. Fu, Stefano Ermon, Atri Rudra et al.NeurIPS 2022 · 5,493 citations
- Implicit Neural Representations with Periodic Activation FunctionsVincent Sitzmann, Julien N. P. Martel, Alexander W. Bergman, David B. Lindell et al.NeurIPS 2020 · 4,008 citations
- Efficiently Modeling Long Sequences with Structured State SpacesAlbert Gu, Karan Goel, Christopher RéICLR 2022 · 3,482 citations
- Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space ModelLianghui Zhu, Bencheng Liao, Qian Zhang, Xinlong Wang et al.ICML 2024 · 1,725 citations
- E(n) Equivariant Graph Neural NetworksVictor Garcia Satorras, Emiel Hoogeboom, Max WellingICML 2021 · 1,432 citations
Related papers
- Hyena Hierarchy: Towards Larger Convolutional Language ModelsMichael Poli, Stefano Massaroli, Eric Nguyen, Daniel Y. Fu et al.ICML 2023 · 481 citations
- Flash Inference: Near Linear Time Inference for Long Convolution Sequence Models and BeyondCostin-Andrei Oncescu, Sanket Purandare, Stratos Idreos, Sham M. KakadeICLR 2025
- Spatial Attention Kinetic Networks with E(n)-EquivarianceYuanqing Wang, John D. ChoderaICLR 2023 · 11 citations
- Equivariant Spatio-Temporal Attentive Graph Networks to Simulate Physical DynamicsLiming Wu, Zhichao Hou, Jirui Yuan, Yu Rong et al.NeurIPS 2023 · 34 citations
- Laughing Hyena Distillery: Extracting Compact Recurrences From ConvolutionsStefano Massaroli, Michael Poli, Daniel Y. Fu, Hermann Kumbong et al.NeurIPS 2023 · 31 citations
