A Wasserstein Minimax Framework for Mixed Linear Regression
Theo Diamandis, Yonina C. Eldar, Alireza Fallah, Farzan Farnia, Asuman E. Ozdaglar
Abstract
Multi-modal distributions are commonly used to model clustered data in statistical learning tasks. In this paper, we consider the Mixed Linear Regression (MLR) problem. We propose an optimal transport-based framework for MLR problems, Wasserstein Mixed Linear Regression (WMLR), which minimizes the Wasserstein distance between the learned and target mixture regression models. Through a model-based duality analysis, WMLR reduces the underlying MLR task to a nonconvex-concave minimax optimization problem, which can be provably solved to find a minimax stationary point by the Gradient Descent Ascent (GDA) algorithm. In the special case of mixtures of two linear regression models, we show that WMLR enjoys global convergence and generalization guarantees. We prove that WMLR's sample complexity grows linearly with the dimension of data. Finally, we discuss the application of WMLR to the federated learning task where the training samples are collected by multiple agents in a network. Unlike the Expectation Maximization algorithm, WMLR directly extends to the distributed, federated learning setting. We support our theoretical results through several numerical experiments, which highlight our framework's ability to handle the federated learning setting with mixture models.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 35884ed7-727e-41dc-8d77-0894ea31caf9Cited by top-tier papers5
- Learning Mixtures of Linear Dynamical SystemsYanxi Chen, H. Vincent PoorICML 2022 · 22 citations
- Collaboration Equilibrium in Federated LearningSen Cui, Jian Liang, Weishen Pan, Kun Chen et al.KDD 2022 · 17 citations
- Imbalanced Mixed Linear RegressionPini Zilber, Boaz NadlerNeurIPS 2023 · 6 citations
- Convergence of Online Learning Algorithm for a Mixture of Multiple Linear RegressionsYujing Liu, Zhixin Liu, Lei GuoICML 2024 · 2 citations
- Covariate-Guided Clusterwise Linear Regression for Generalization to Unseen DataDohyun Bu, Hyunho Kim, Jong-Seok LeeICLR 2026
Builds on6
- SCAFFOLD: Stochastic Controlled Averaging for Federated LearningSai Praneeth Karimireddy, Satyen Kale, Mehryar Mohri, Sashank J. Reddi et al.ICML 2020 · 3,875 citations
- Personalized Federated Learning with Theoretical Guarantees: A Model-Agnostic Meta-Learning ApproachAlireza Fallah, Aryan Mokhtari, Asuman E. OzdaglarNeurIPS 2020 · 1,354 citations
- An Efficient Framework for Clustered Federated LearningAvishek Ghosh, Jichan Chung, Dong Yin, Kannan RamchandranNeurIPS 2020 · 1,329 citations
- On Gradient Descent Ascent for Nonconvex-Concave Minimax ProblemsTianyi Lin, Chi Jin, Michael I. JordanICML 2020 · 587 citations
- Robust Federated Learning: The Case of Affine Distribution ShiftsAmirhossein Reisizadeh, Farzan Farnia, Ramtin Pedarsani, Ali JadbabaieNeurIPS 2020 · 196 citations
Related papers
- An Optimal Transport-based Latent Mixer for Robust Multi-modal LearningFengjiao Gong, Angxiao Yue, Hongteng XuAAAI 2025
- Global Convergence of Federated Learning for Mixed RegressionLili Su, Jiaming Xu, Pengkun YangNeurIPS 2022 · 9 citations
- A Communication-efficient Algorithm with Linear Convergence for Federated Minimax LearningZhenyu Sun, Ermin WeiNeurIPS 2022 · 20 citations
- Accelerated Dual Method for Distributed Optimization: An Inexact-Gradient View of Local UpdatesJunchi Yang, Ziyang Zeng, Linxuan Pan, Murat Yildirim et al.ICML 2026
- Federated Multi-Objective LearningHaibo Yang, Zhuqing Liu, Jia Liu, Chaosheng Dong et al.NeurIPS 2023 · 28 citations
