CoLA: Exploiting Compositional Structure for Automatic and Efficient Numerical Linear Algebra
Andres Potapczynski, Marc Finzi, Geoff Pleiss, Andrew Gordon Wilson
Abstract
Many areas of machine learning and science involve large linear algebra problems, such as eigendecompositions, solving linear systems, computing matrix exponentials, and trace estimation. The matrices involved often have Kronecker, convolutional, block diagonal, sum, or product structure. In this paper, we propose a simple but general framework for large-scale linear algebra problems in machine learning, named CoLA (Compositional Linear Algebra). By combining a linear operator abstraction with compositional dispatch rules, CoLA automatically constructs memory and runtime efficient numerical algorithms. Moreover, CoLA provides memory efficient automatic differentiation, low precision computation, and GPU acceleration in both JAX and PyTorch, while also accommodating new objects, operations, and rules in downstream packages via multiple dispatch. CoLA can accelerate many algebraic operations, while making it easy to prototype matrix structures and algorithms, providing an appealing drop-in tool for virtually any computational effort that requires linear algebra. We showcase its efficacy across a broad range of applications, including partial differential equations, Gaussian processes, equivariant model construction, and unsupervised learning.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 974da47d-5710-486f-beb2-b52253e7f2acCited by top-tier papers7
- Hidden Breakthroughs in Language Model TrainingSara Kangaslahti, Elan Rosenfeld, Naomi SaphraICLR 2026 · 17 citations
- Searching for Efficient Linear Layers over a Continuous Space of Structured MatricesAndres Potapczynski, Shikai Qiu, Marc Finzi, Christopher Ferri et al.NeurIPS 2024 · 11 citations
- Hamiltonian Monte Carlo Inference of Marginalized Linear Mixed-Effects ModelsJinlin Lai, Justin Domke, Daniel R. SheldonNeurIPS 2024 · 2 citations
- Diffusing Differentiable RepresentationsYash Savani, Marc Finzi, J. Zico KolterNeurIPS 2024 · 1 citation
- Compute-Optimal LLMs Provably Generalize Better with ScaleMarc Anton Finzi, Sanyam Kapoor, Diego Granziol, Anming Gu et al.ICLR 2025
Builds on3
- A Practical Method for Constructing Equivariant Multilayer Perceptrons for Arbitrary Matrix GroupsMarc Finzi, Max Welling, Andrew Gordon WilsonICML 2021 · 226 citations
- Hungry Hungry Hippos: Towards Language Modeling with State Space ModelsDaniel Y. Fu, Tri Dao, Khaled Kamal Saab, Armin W. Thomas et al.ICLR 2023 · 117 citations
- A Stable and Scalable Method for Solving Initial Value PDEs with Neural NetworksMarc Anton Finzi, Andres Potapczynski, Matthew Choptuik, Andrew Gordon WilsonICLR 2023 · 1 citation
Related papers
- Gradients of Functions of Large MatricesNicholas Krämer, Pablo Moreno-Muñoz, Hrittik Roy, Søren HaubergNeurIPS 2024 · 6 citations
- CoLA: Compute-Efficient Pre-Training of LLMs via Low-Rank ActivationZiyue Liu, Ruijie Zhang, Zhengyang Wang, Mingsong Yan et al.EMNLP 2025
- Memory safe computations with XLA compilerArtem Artemev, Yuze An, Tilman Roeder, Mark van der WilkNeurIPS 2022 · 10 citations
- A Simple and Efficient Tensor CalculusSören Laue, Matthias Mitterreiter, Joachim GiesenAAAI 2020 · 40 citations
- Learning, Solving and Optimizing PDEs with TensorGalerkin: an efficient high-performance Galerkin assembly algorithmShizheng Wen, Mingyuan chi, Tianwei Yu, Ben Moseley et al.ICML 2026 · 1 citation
