PowerMLP: An Efficient Version of KAN
Ruichen Qiu, Yibo Miao, Shiwen Wang, Yifan Zhu, Lijia Yu, Xiao-Shan Gao
摘要
The Kolmogorov-Arnold Network (KAN) is a new network architecture known for its high accuracy in several tasks such as function fitting and PDE solving. The superior expressive capability of KAN arises from the Kolmogorov-Arnold representation theorem and learnable spline functions. However, the computation of spline functions involves multiple iterations, which renders KAN significantly slower than MLP, thereby increasing the cost associated with model training and deployment. The authors of KAN have also noted that "the biggest bottleneck of KANs lies in its slow training. KANs are usually 10x slower than MLPs, given the same number of parameters." To address this issue, we propose a novel MLP-type neural network PowerMLP that employs simpler non-iterative spline function representation, offering approximately the same training time as MLP while theoretically demonstrating stronger expressive power than KAN. Furthermore, we compare the FLOPs of KAN and Pow-erMLP, quantifying the faster computation speed of Pow-erMLP. Our comprehensive experiments demonstrate that PowerMLP generally achieves higher accuracy and a training speed about 40 times faster than KAN in various tasks.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Generalization Bounds for Kolmogorov-Arnold Networks (KANs) and Enhanced KANs with Lower Lipschitz ComplexityPengqi Li, Lizhong Ding, Jiarun Fu, Chunhui Zhang 等NeurIPS 2025 · 被引用 8 次
- A Geometric Algebra-informed NeRF Framework for Generalizable Wireless Channel PredictionJingzhou Shen, Luis Lago Enamorado, Shiwen Mao, Xuyu WangINFOCOM 2026 · 被引用 1 次
相关 Paper
- KAN: Kolmogorov-Arnold NetworksZiming Liu, Yixuan Wang, Sachin Vaidya, Fabian Ruehle 等ICLR 2025
- On the expressiveness and spectral bias of KANsYixuan Wang, Jonathan W. Siegel, Ziming Liu, Thomas Y. HouICLR 2025
- Initialization Schemes for Kolmogorov–Arnold Networks: An Empirical StudySpyros Rigas, Dhruv Verma, Georgios Alexandridis, Yixuan WangICLR 2026 · 被引用 13 次
- Incorporating Arbitrary Matrix Group Equivariance into KANsLexiang Hu, Yisen Wang, Zhouchen LinICML 2025
- FeKAN: Efficient Kolmogorov-Arnold Networks Accelerator Using FeFET-based CAM and LUTXuliang Yu, Yu Qian, Xunzhao Yin, Cheng Zhuo 等DAC 2025 · 被引用 1 次
