Composing Linear Layers from Irreducibles
Travis Pence, Daisuke Yamada, Vikas Singh
Abstract
Contemporary large models often exhibit behaviors suggesting the presence of low-level primitives that compose into modules with richer functionality, but these fundamental building blocks remain poorly understood. We investigate this compositional structure in linear layers by asking: can we identify/synthesize linear transformations from a minimal set of geometric primitives? Using Clifford algebra, we show that linear layers can be expressed as compositions of bivectors -- geometric objects encoding oriented planes -- and introduce a differentiable algorithm that decomposes them into products of rotors. This construction uses only O(log^2 d) parameters, versus O(d^2) required by dense matrices. Applied to the key, query, and value projections in LLM attention layers, our rotor-based layers match the performance of strong baselines such as block-Hadamard and low-rank approximations. Our findings provide an algebraic perspective on how these geometric primitives can compose into higher-level functions within deep models.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 0c3e1b1c-cfc1-4d1b-8af0-5f51ec413b03Cited by top-tier papers1
Ask how each one uses itBuilds on15
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- Scaling Vision with Sparse Mixture of ExpertsCarlos Riquelme, Joan Puigcerver, Basil Mustafa, Maxim Neumann et al.NeurIPS 2021 · 1,213 citations
- Thinking Like TransformersGail Weiss, Yoav Goldberg, Eran YahavICML 2021 · 183 citations
- Compositional Generalization via Neural-Symbolic Stack MachinesXinyun Chen, Chen Liang, Adams Wei Yu, Dawn Song et al.NeurIPS 2020 · 112 citations
- Clifford Group Equivariant Neural NetworksDavid Ruhe, Johannes Brandstetter, Patrick ForréNeurIPS 2023 · 85 citations
Related papers
- Geometric Algebra TransformerJohann Brehmer, Pim de Haan, Sönke Behrends, Taco S. CohenNeurIPS 2023 · 81 citations
- Geometric Clifford Algebra NetworksDavid Ruhe, Jayesh K. Gupta, Steven De Keninck, Max Welling et al.ICML 2023 · 58 citations
- GLGENN: A Novel Parameter-Light Equivariant Neural Networks Architecture Based on Clifford Geometric AlgebrasEkaterina Filimoshina, Dmitry ShirokovICML 2025
- Clifford Group Equivariant Simplicial Message Passing NetworksCong Liu, David Ruhe, Floor Eijkelboom, Patrick ForréICLR 2024 · 20 citations
- Clifford Neural Layers for PDE ModelingJohannes Brandstetter, Rianne van den Berg, Max Welling, Jayesh K. GuptaICLR 2023 · 21 citations
