Joker: Joint Optimization Framework for Lightweight Kernel Machines
Junhong Zhang, Zhihui Lai
摘要
Kernel methods are powerful tools for nonlinear learning with well-established theory. The scalability issue has been their long-standing challenge. Despite the existing success, there are two limitations in large-scale kernel methods: (i) The memory overhead is too high for users to afford; (ii) existing efforts mainly focus on kernel ridge regression (KRR), while other models lack study. In this paper, we propose Joker, a joint optimization framework for diverse kernel models, including KRR, logistic regression, and support vector machines. We design a dual block coordinate descent method with trust region (DBCD-TR) and adopt kernel approximation with randomized features, leading to low memory costs and high efficiency in large-scale learning. Experiments show that Joker saves up to 90% memory but achieves comparable training time and performance (or even better) than the state-of-the-art methods.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper7
- Kernel Methods Through the Roof: Handling Billions of Points EfficientlyGiacomo Meanti, Luigi Carratino, Lorenzo Rosasco, Alessandro RudiNeurIPS 2020 · 被引用 138 次
- On the Similarity between the Laplace and Neural Tangent KernelsAmnon Geifman, Abhay Kumar Yadav, Yoni Kasten, Meirav Galun 等NeurIPS 2020 · 被引用 118 次
- The Deep Bootstrap Framework: Good Online Learners are Good Offline GeneralizersPreetum Nakkiran, Behnam Neyshabur, Hanie SedghiICLR 2021 · 被引用 75 次
- Error Bounds for Learning with Vector-Valued Random FeaturesSamuel Lanthaler, Nicholas H. NelsenNeurIPS 2023 · 被引用 26 次
- Toward Large Kernel ModelsAmirhesam Abedsoltan, Mikhail Belkin, Parthe PanditICML 2023 · 被引用 23 次
相关 Paper
- ParK: Sound and Efficient Kernel Ridge Regression by Feature Space PartitionsLuigi Carratino, Stefano Vigogna, Daniele Calandriello, Lorenzo RosascoNeurIPS 2021 · 被引用 7 次
- Fast Training of Large Kernel Models with Delayed ProjectionsAmirhesam Abedsoltan, Siyuan Ma, Parthe Pandit, Misha BelkinNeurIPS 2025 · 被引用 2 次
- Distributed Randomized Sketching Kernel LearningRong Yin, Yong Liu, Dan MengAAAI 2022 · 被引用 4 次
- Large-Scale Learning with Fourier Features and Tensor DecompositionsFrederiek Wesel, Kim BatselierNeurIPS 2021 · 被引用 20 次
- Effective Distributed Learning with Random Features: Improved Bounds and AlgorithmsYong Liu, Jiankun Liu, Shuqiang WangICLR 2021 · 被引用 21 次
