Automatic Optimization of Matrix Implementations for Distributed Machine Learning and Linear Algebra
Shangyu Luo, Dimitrije Jankov, Binhang Yuan, Chris Jermaine
Abstract
Machine learning (ML) computations are often expressed using vectors, matrices, or higher-dimensional tensors. Such data structures can have many different implementations, especially in a distributed environment: a matrix could be stored as row or column vectors, tiles of different sizes, or relationally, as a set of (rowIndex, colIndex, value) triples. Many other storage formats are possible. The choice of format can have a profound impact on the performance of a ML computation. In this paper, we propose a framework for automatic optimization of the physical implementation of a complex ML or linear algebra (LA) computation in a distributed environment, develop algorithms for solving this problem, and show, through a prototype on top of a distributed relational database system, that our ideas can radically speed up common ML and LA computations.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 03c661a6-3d4d-4141-bb79-54053bf6eaebCited by top-tier papers5
- Autoscheduling for sparse tensor algebra with an asymptotic cost modelWillow Ahrens, Fredrik Kjolstad, Saman P. AmarasinghePLDI 2022 · 30 citations
- In-Database Machine Learning with CorgiPile: Stochastic Gradient Descent without Full Data ShuffleLijie Xu, Shuang Qiu, Binhang Yuan, Jiawei Jiang et al.SIGMOD 2022 · 11 citations
- AWARE: Workload-aware, Redundancy-exploiting Linear AlgebraSebastian Baunsgaard, Matthias BoehmSIGMOD 2023 · 4 citations
- Givens QR Decomposition over Relational DatabasesDan Olteanu, Nils Vortmeier, Dorde ZivanovicSIGMOD 2022 · 3 citations
- TranSQL + : Serving Large Language Models with SQL on Low-Resource HardwareWenbo Sun, Qiming Guo, Wenlu Wang, Rihan HaiSIGMOD 2026 · 1 citation
Builds on1
Related papers
- Tensor Relational Algebra for Distributed Machine Learning System DesignBinhang Yuan, Dimitrije Jankov, Jia Zou, Yuxin Tang et al.VLDB 2021 · 33 citations
- Optimizing Tensor Programs on Flexible StorageMaximilian Schleich, Amir Shaikhha, Dan SuciuSIGMOD 2023 · 21 citations
- Auto-Differentiation of Relational Computations for Very Large Scale Machine LearningYuxin Tang, Zhimin Ding, Dimitrije Jankov, Binhang Yuan et al.ICML 2023 · 7 citations
- Distributed Numerical and Machine Learning Computations via Two-Phase Execution of Aggregated Join TreesDimitrije Jankov, Binhang Yuan, Shangyu Luo, Chris JermaineVLDB 2021 · 9 citations
- Deinsum: Practically I/O Optimal Multi-Linear AlgebraAlexandros Nikolaos Ziogas, Grzegorz Kwasniewski, Tal Ben-Nun, Timo Schneider et al.SC 2022 · 2 citations
