Sinkhorn Barycenter via Functional Gradient Descent
Zebang Shen, Zhenfu Wang, Alejandro Ribeiro, Hamed Hassani
Abstract
In this paper, we consider the problem of computing the barycenter of a set of probability distributions under the Sinkhorn divergence. This problem has recently found applications across various domains, including graphics, learning, and vision, as it provides a meaningful mechanism to aggregate knowledge. Unlike previous approaches which directly operate in the space of probability measures, we recast the Sinkhorn barycenter problem as an instance of unconstrained functional optimization and develop a novel functional gradient descent method named Sinkhorn Descent (SD). We prove that SD converges to a stationary point at a sublinear rate, and under reasonable assumptions, we further show that it asymptotically finds a global minimizer of the Sinkhorn barycenter problem. Moreover, by providing a mean-field analysis, we show that SD preserves the weak convergence of empirical measures. Importantly, the computational complexity of SD scales linearly in the dimension d and we demonstrate its scalability by solving a 100-dimensional Sinkhorn barycenter problem. Analysis In this section, we analyze the finite time convergence and the mean field limit of SD under the following assumptions on the ground cost function c and the kernel function k of the RKHS H d . Assumption 4.1. The ground cost function c(x, y) is bounded, i.e. ∀x, y ∈ X , c(x, y) ≤ M c ; G c -Lipschitz continuous, i.e. ∀x, x , y ∈ X , |c(x, y) -c(x , y)| ≤ G c x -x ; and L c -Lipschitz smooth, i.e. ∀x, x , y ∈ X , ∇ 1 c(x, y) -∇ 1 c(x , y) ≤ L c x -x . Assumption 4.2. The kernel function k(x, y) is bounded, i.e. ∀x, y ∈ X , k(x, y) ≤ D k ; G k -Lipschitz continuous, i.e. ∀x, x , y ∈ X , |k(x, y) -k(x , y)| ≤ G c x -x .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 61d637c2-385e-410f-920f-79e3036b22e8Cited by top-tier papers4
- Nonparametric Iterative Machine TeachingChen Zhang, Xiaofeng Cao, Weiyang Liu, Ivor W. Tsang et al.ICML 2023 · 13 citations
- Nonparametric Teaching of Implicit Neural RepresentationsChen Zhang, Steven Tin Sui Luo, Jason Chun Lok Li, Yik-Chung Wu et al.ICML 2024 · 12 citations
- Nonparametric Teaching for Multiple LearnersChen Zhang, Xiaofeng Cao, Weiyang Liu, Ivor W. Tsang et al.NeurIPS 2023 · 8 citations
- Nonparametric Teaching for Graph Property LearnersChen Zhang, Weixin Bu, Zeyi Ren, Zhengwu Liu et al.ICML 2025
Related papers
- Optimal Transport Barycenter via Nonconvex-Concave Minimax OptimizationKaheon Kim, Rentian Yao, Changbo Zhu, Xiaohui ChenICML 2025
- Mirror Descent with Relative Smoothness in Measure Spaces, with application to Sinkhorn and EMPierre-Cyril Aubin-Frankowski, Anna Korba, Flavien LégerNeurIPS 2022 · 61 citations
- Sobolev Gradient Ascent for Optimal Transport: Barycenter Optimization and Convergence AnalysisKaheon Kim, Bohan Zhou, Changbo Zhu, Xiaohui ChenICLR 2026 · 6 citations
- Hilbert Sinkhorn Divergence for Optimal TransportQian Li, Zhichao Wang, Gang Li, Jun Pang et al.CVPR 2021
- Efficient Approximation Algorithm for Computing Wasserstein Barycenter under Euclidean MetricPankaj K. Agarwal, Sharath Raghvendra, Pouyan Shirzadian, Keegan YaoSODA 2025
