Lune

ICML2023Top-tier venue

Two Losses Are Better Than One: Faster Optimization Using a Cheaper Proxy

Blake E. Woodworth, Konstantin Mishchenko, Francis R. Bach

2023Year
9Citations
2Top-tier citations

Abstract

We present an algorithm for minimizing an objective with hard-to-compute gradients by using a related, easier-to-access function as a proxy. Our algorithm is based on approximate proximal point iterations on the proxy combined with relatively few stochastic gradients from the objective. When the difference between the objective and the proxy is δ\delta-smooth, our algorithm guarantees convergence at a rate matching stochastic gradient descent on a δ\delta-smooth objective, which can lead to substantially better sample efficiency. Our algorithm has many potential applications in machine learning, and provides a principled means of leveraging synthetic data, physics simulators, mixed public and private data, and more.

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext a4c696a0-8483-4109-846a-c5facd02cbb0

Cited by top-tier papers2

Ask how each one uses it

Builds on4

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines