Learning Polynomial Problems with SL(2, R)-Equivariance
Hannah Lawrence, Mitchell Tong Harris
摘要
Optimizing and certifying the positivity of polynomials are fundamental primitives across mathematics and engineering applications, from dynamical systems to operations research. However, solving these problems in practice requires large semidefinite programs, with poor scaling in dimension and degree. In this work, we demonstrate for the first time that neural networks can effectively solve such problems in a data-driven fashion, achieving tenfold speedups while retaining high accuracy. Moreover, we observe that these polynomial learning problems are equivariant to the non-compact group , which consists of area-preserving linear transformations. We therefore adapt our learning pipelines to accommodate this structure, including data augmentation, a new -equivariant architecture, and an architecture equivariant with respect to its maximal compact subgroup, . Surprisingly, the most successful approaches in practice do not enforce equivariance to the entire group, which we prove arises from an unusual lack of architecture universality for in particular. A consequence of this result, which is of independent interest, is that there exists an equivariant function for which there is no sequence of equivariant polynomials multiplied by arbitrary invariants that approximates the original function. This is a rare example of a symmetric problem where data augmentation outperforms a fully equivariant architecture, and provides interesting lessons in both theory and practice for other problems with non-compact symmetries.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper11
- Generalizing Convolutional Neural Networks for Equivariance to Lie Groups on Arbitrary Continuous DataMarc Finzi, Samuel Stanton, Pavel Izmailov, Andrew Gordon WilsonICML 2020 · 被引用 372 次
- A Practical Method for Constructing Equivariant Multilayer Perceptrons for Arbitrary Matrix GroupsMarc Finzi, Max Welling, Andrew Gordon WilsonICML 2021 · 被引用 226 次
- Frame Averaging for Invariant and Equivariant Network DesignOmri Puny, Matan Atzmon, Edward J. Smith, Ishan Misra 等ICLR 2022 · 被引用 177 次
- Lorentz Group Equivariant Neural Network for Particle PhysicsAlexander Bogatskiy, Brandon M. Anderson, Jan T. Offermann, Marwah Roussi 等ICML 2020 · 被引用 164 次
- Provably Strict Generalisation Benefit for Equivariant ModelsBryn Elesedy, Sheheryar ZaidiICML 2021 · 被引用 100 次
相关 Paper
- Scalars are universal: Equivariant machine learning, structured like classical physicsSoledad Villar, David W. Hogg, Kate Storey-Fisher, Weichi Yao 等NeurIPS 2021 · 被引用 185 次
- Neural Sum-of-Squares: Certifying the Nonnegativity of Polynomials with TransformersNico Pelleriti, Christoph Spiegel, Shiwei Liu, David Martínez-Rubio 等ICLR 2026 · 被引用 2 次
- Equivariance with Learned Canonicalization FunctionsSékou-Oumar Kaba, Arnab Kumar Mondal, Yan Zhang, Yoshua Bengio 等ICML 2023 · 被引用 109 次
- On the hardness of learning under symmetriesBobak T. Kiani, Thien Le, Hannah Lawrence, Stefanie Jegelka 等ICLR 2024 · 被引用 14 次
- Interpretable Discovery of One-parameter Subgroups: A Modular Framework for Elliptical, Hyperbolic, and Parabolic SymmetriesPavan Karjol, Vivek Kashyap, Rohan Venkatesh Kashyap, Prathosh APICML 2026
