Kolmogorov-Arnold Networks Still Catastrophically Forget but Differently from MLP
Anton Lee, Heitor Murilo Gomes, Yaqian Zhang, W. Bastiaan Kleijn
摘要
Catastrophic forgetting is when a neural network loses previously learnt information after learning a new task sequentially. Avoiding catastrophic forgetting could reduce the resources necessary to update neural networks. Recently, Kolmogorov-Arnold Networks (KAN) gained the community's attention as preliminary experiments suggest KAN avoid catastrophic forgetting. KAN replace neural network edges with learnable B-splines and sum incoming edges in nodes. Proponents of KAN argue they avoid forgetting, are more accurate, are interpretable, and use fewer parameters. Our work investigates the claims that KAN avoid catastrophic forgetting, finding that they fail to do so on more complex datasets containing features that overlap between tasks. We give a simple explanation as to why and how KAN catastrophically forget. Motivated by evidence suggesting KAN are superior for symbolic regression, we augment KAN in the same ways as multilayer perceptron (MLP) to perform continual learning tasks, making special accommodations to support KAN. Our experiments found that unmodified KAN often forget more than MLP, but KAN can be better than MLP when combined with continual learning strategies. We aim to highlight some of the current shortcomings and strengths associated with KAN for continual learning.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Unifying Locality of KANs and Feature Drift Compensation Projection for Data-Free Replay Based Continual Face Forgery DetectionTianshuo Zhang, Siran Peng, Li Gao, Haoyuan Zhang 等AAAI 2026 · 被引用 1 次
- Catastrophic Forgetting in Kolmogorov-Arnold NetworksMohammad Marufur Rahman, Guanchu Wang, Kaixiong Zhou, Minghan Chen 等AAAI 2026 · 被引用 1 次
它引用的顶会 Paper3
- Dark Experience for General Continual Learning: a Strong, Simple BaselinePietro Buzzega, Matteo Boschini, Angelo Porrello, Davide Abati 等NeurIPS 2020 · 被引用 1,494 次
- Supermasks in SuperpositionMitchell Wortsman, Vivek Ramanujan, Rosanne Liu, Aniruddha Kembhavi 等NeurIPS 2020 · 被引用 364 次
- Wide Neural Networks Forget Less CatastrophicallySeyed-Iman Mirzadeh, Arslan Chaudhry, Dong Yin, Huiyi Hu 等ICML 2022 · 被引用 84 次
相关 Paper
- KAC: Kolmogorov-Arnold Classifier for Continual LearningYusong Hu, Zichen Liang, Fei Yang, Qibin Hou 等CVPR 2025
- Neuro-Symbolic Continual Learning: Knowledge, Reasoning Shortcuts and Concept RehearsalEmanuele Marconato, Gianpaolo Bontempo, Elisa Ficarra, Simone Calderara 等ICML 2023 · 被引用 34 次
- KAN: Kolmogorov-Arnold NetworksZiming Liu, Yixuan Wang, Sachin Vaidya, Fabian Ruehle 等ICLR 2025
- PowerMLP: An Efficient Version of KANRuichen Qiu, Yibo Miao, Shiwen Wang, Yifan Zhu 等AAAI 2025 · 被引用 13 次
- Learning curves for continual learning in neural networks: Self-knowledge transfer and forgettingRyo Karakida, Shotaro AkahoICLR 2022 · 被引用 16 次
