QuadEnhancer: Leveraging Quadratic Transformations to Enhance Deep Neural Networks
Qian Chen, Linxin Yang, Akang Wang, Xiaodong Luo, Yin Zhang
Abstract
The combination of linear transformations and non-linear activation functions forms the foundation of most modern deep neural networks, enabling them to approximate highly complex functions. This paper explores the introduction of quadratic transformations to further increase nonlinearity in neural networks, with the aim of enhancing the performance of existing architectures. To reduce parameter complexity and computational complexity, we propose a lightweight quadratic enhancer that uses low-rankness, weight sharing, and sparsification techniques. For a fixed architecture, the proposed approach introduces quadratic interactions between features at every layer, while only adding negligible amounts of additional model parameters and forward computations. We conduct a set of proof-of-concept experiments for the proposed method across three tasks: image classification, text classification, and fine-tuning large-language models. In all tasks, the proposed approach demonstrates clear and substantial performance gains.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 3745c118-6237-4eaf-bb66-b577bfef998cCited by top-tier papers1
Ask how each one uses itBuilds on3
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn et al.ICLR 2021 · 21,477 citations
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu et al.ICLR 2022 · 18,833 citations
- LLM-Adapters: An Adapter Family for Parameter-Efficient Fine-Tuning of Large Language ModelsZhiqiang Hu, Lei Wang, Yihuai Lan, Wanyu Xu et al.EMNLP 2023 · 200 citations
Related papers
- LQF: Linear Quadratic Fine-TuningAlessandro Achille, Aditya Golatkar, Avinash Ravichandran, Marzia Polito et al.CVPR 2021
- DeLighT: Deep and Light-weight TransformerSachin Mehta, Marjan Ghazvininejad, Srinivasan Iyer, Luke Zettlemoyer et al.ICLR 2021 · 96 citations
- Deformable Butterfly: A Highly Structured and Sparse Linear TransformRui Lin, Jie Ran, King Hung Chiu, Graziano Chesi et al.NeurIPS 2021 · 17 citations
- Deep Learning with Learnable Product-Structured ActivationsSaanjali Maharaj, Prasanth B. NairICLR 2026
- Revisiting Feature Interactions from the Perspective of Quadratic Neural Networks for Click-through Rate PredictionHonghao Li, Yiwen Zhang, Yi Zhang, Lei Sang et al.KDD 2025 · 1 citation
