Lune

NeurIPS2025Top-tier venue

Nonlinearly Preconditioned Gradient Methods: Momentum and Stochastic Analysis

Konstantinos A. Oikonomidis, Jan Quan, Panagiotis Patrinos

2025Year
6Citations

Abstract

We study nonlinearly preconditioned gradient methods for smooth nonconvex optimization problems, focusing on sigmoid preconditioners that inherently perform a form of gradient clipping akin to the widely used gradient clipping technique. Building upon this idea, we introduce a novel heavy ball-type algorithm and provide convergence guarantees under a generalized smoothness condition that is less restrictive than traditional Lipschitz smoothness, thus covering a broader class of functions. Additionally, we develop a stochastic variant of the base method and study its convergence properties under different noise assumptions. We compare the proposed algorithms with baseline methods on diverse tasks from machine learning including neural network training.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext eab35a2a-b151-44e2-ad62-1d5efd92e884

Builds on19

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines