Sinkhorn Barycenter via Functional Gradient Descent
Zebang Shen, Zhenfu Wang, Alejandro Ribeiro, Hamed Hassani
摘要
In this paper, we consider the problem of computing the barycenter of a set of probability distributions under the Sinkhorn divergence. This problem has recently found applications across various domains, including graphics, learning, and vision, as it provides a meaningful mechanism to aggregate knowledge. Unlike previous approaches which directly operate in the space of probability measures, we recast the Sinkhorn barycenter problem as an instance of unconstrained functional optimization and develop a novel functional gradient descent method named Sinkhorn Descent (SD). We prove that SD converges to a stationary point at a sublinear rate, and under reasonable assumptions, we further show that it asymptotically finds a global minimizer of the Sinkhorn barycenter problem. Moreover, by providing a mean-field analysis, we show that SD preserves the weak convergence of empirical measures. Importantly, the computational complexity of SD scales linearly in the dimension d and we demonstrate its scalability by solving a 100-dimensional Sinkhorn barycenter problem. Analysis In this section, we analyze the finite time convergence and the mean field limit of SD under the following assumptions on the ground cost function c and the kernel function k of the RKHS H d . Assumption 4.1. The ground cost function c(x, y) is bounded, i.e. ∀x, y ∈ X , c(x, y) ≤ M c ; G c -Lipschitz continuous, i.e. ∀x, x , y ∈ X , |c(x, y) -c(x , y)| ≤ G c x -x ; and L c -Lipschitz smooth, i.e. ∀x, x , y ∈ X , ∇ 1 c(x, y) -∇ 1 c(x , y) ≤ L c x -x . Assumption 4.2. The kernel function k(x, y) is bounded, i.e. ∀x, y ∈ X , k(x, y) ≤ D k ; G k -Lipschitz continuous, i.e. ∀x, x , y ∈ X , |k(x, y) -k(x , y)| ≤ G c x -x .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- Nonparametric Iterative Machine TeachingChen Zhang, Xiaofeng Cao, Weiyang Liu, Ivor W. Tsang 等ICML 2023 · 被引用 13 次
- Nonparametric Teaching of Implicit Neural RepresentationsChen Zhang, Steven Tin Sui Luo, Jason Chun Lok Li, Yik-Chung Wu 等ICML 2024 · 被引用 12 次
- Nonparametric Teaching for Multiple LearnersChen Zhang, Xiaofeng Cao, Weiyang Liu, Ivor W. Tsang 等NeurIPS 2023 · 被引用 8 次
- Nonparametric Teaching for Graph Property LearnersChen Zhang, Weixin Bu, Zeyi Ren, Zhengwu Liu 等ICML 2025
相关 Paper
- Optimal Transport Barycenter via Nonconvex-Concave Minimax OptimizationKaheon Kim, Rentian Yao, Changbo Zhu, Xiaohui ChenICML 2025
- Mirror Descent with Relative Smoothness in Measure Spaces, with application to Sinkhorn and EMPierre-Cyril Aubin-Frankowski, Anna Korba, Flavien LégerNeurIPS 2022 · 被引用 61 次
- Sobolev Gradient Ascent for Optimal Transport: Barycenter Optimization and Convergence AnalysisKaheon Kim, Bohan Zhou, Changbo Zhu, Xiaohui ChenICLR 2026 · 被引用 6 次
- Hilbert Sinkhorn Divergence for Optimal TransportQian Li, Zhichao Wang, Gang Li, Jun Pang 等CVPR 2021
- Efficient Approximation Algorithm for Computing Wasserstein Barycenter under Euclidean MetricPankaj K. Agarwal, Sharath Raghvendra, Pouyan Shirzadian, Keegan YaoSODA 2025
