Symmetry Teleportation for Accelerated Optimization
Bo Zhao, Nima Dehmamy, Robin Walters, Rose Yu
Abstract
Existing gradient-based optimization methods update parameters locally, in a direction that minimizes the loss function. We study a different approach, symmetry teleportation, that allows parameters to travel a large distance on the loss level set, in order to improve the convergence speed in subsequent steps. Teleportation exploits symmetries in the loss landscape of optimization problems. We derive loss-invariant group actions for test functions in optimization and multi-layer neural networks, and prove a necessary condition for teleportation to improve convergence rate. We also show that our algorithm is closely related to second order methods. Experimentally, we show that teleportation improves the convergence speed of gradient descent and AdaGrad for several optimization problems including test functions, multi-layer regressions, and MNIST classification. Our code is available at https://github.com/Rose-
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7f30c2d1-20f2-4d91-a329-ad72784871d3Cited by top-tier papers16
- Hidden Symmetries of ReLU NetworksJ. Elisenda Grigsby, Kathryn Lindsey, David RolnickICML 2023 · 35 citations
- Improving Convergence and Generalization Using Parameter SymmetriesBo Zhao, Robert M. Gower, Robin Walters, Rose YuICLR 2024 · 24 citations
- Scale Equivariant Graph MetanetworksIoannis Kalogeropoulos, Giorgos Bouritsas, Yannis PanagakisNeurIPS 2024 · 24 citations
- Parameter Symmetry and Noise Equilibrium of Stochastic Gradient DescentLiu Ziyin, Mingze Wang, Hongchao Li, Lei WuNeurIPS 2024 · 23 citations
- Neural Thermodynamics: Entropic Forces in Deep and Universal Representation LearningLiu Ziyin, Yizhou Xu, Isaac L. ChuangNeurIPS 2025 · 11 citations
Builds on4
- Geometry of the Loss Landscape in Overparameterized Neural Networks: Symmetries and InvariancesBerfin Simsek, François Ged, Arthur Jacot, Francesco Spadaro et al.ICML 2021 · 136 citations
- Neural Mechanics: Symmetry and Broken Conservation Laws in Deep Learning DynamicsDaniel Kunin, Javier Sagastuy-Breña, Surya Ganguli, Daniel L. K. Yamins et al.ICLR 2021 · 100 citations
- Understanding the Dynamics of Gradient Flow in Overparameterized Linear modelsSalma Tarmoun, Guilherme França, Benjamin D. Haeffele, René VidalICML 2021 · 76 citations
- On the Explicit Role of Initialization on the Convergence and Implicit Bias of Overparametrized Linear NetworksHancheng Min, Salma Tarmoun, René Vidal, Enrique MalladaICML 2021 · 53 citations
Related papers
- Global curvature for second-order optimization of neural networksAlberto BernacchiaICML 2025
- How Does Adaptive Optimization Impact Local Neural Network Geometry?Kaiqi Jiang, Dhruv Malik, Yuanzhi LiNeurIPS 2023 · 26 citations
- Continual Optimization with Symmetry Teleportation for Multi-Task LearningZhipeng Zhou, Ziqiao Meng, Pengcheng Wu, Peilin Zhao et al.NeurIPS 2025 · 4 citations
- Scalable Decentralized Learning with TeleportationYuki Takezawa, Sebastian U. StichICLR 2025
- A Tale of Two Symmetries: Exploring the Loss Landscape of Equivariant ModelsYuqing Xie, Tess E. SmidtNeurIPS 2025 · 9 citations
