SEDONA: Search for Decoupled Neural Networks toward Greedy Block-wise Learning
Myeongjang Pyeon, Jihwan Moon, Taeyoung Hahn, Gunhee Kim
摘要
Backward locking and update locking are well-known sources of inefficiency in backpropagation that prevent from concurrently updating layers. Several works have recently suggested using local error signals to train network blocks asynchronously to overcome these limitations. However, they often require numerous iterations of trial-and-error to find the best configuration for local training, including how to decouple network blocks and which auxiliary networks to use for each block. In this work, we propose a differentiable search algorithm named SEDONA to automate this process. Experimental results show that our algorithm can consistently discover transferable decoupled architectures for VGG and ResNet variants, and significantly outperforms the ones trained with end-to-end backpropagation and other state-of-the-art greedy-leaning methods in CIFAR-10, Tiny-ImageNet and ImageNet.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Scaling Supervised Local Learning with Augmented Auxiliary NetworksChenxiang Ma, Jibin Wu, Chenyang Si, Kay Chen TanICLR 2024 · 被引用 9 次
- Masked Image Modeling with Local Multi-Scale ReconstructionHaoqing Wang, Yehui Tang, Yunhe Wang, Jianyuan Guo 等CVPR 2023
- DeeperForward: Enhanced Forward-Forward Training for Deeper and Better PerformanceLiang Sun, Yang Zhang, Weizhao He, Jiajun Wen 等ICLR 2025
它引用的顶会 Paper3
- Understanding and Robustifying Differentiable Architecture SearchArber Zela, Thomas Elsken, Tonmoy Saikia, Yassine Marrakchi 等ICLR 2020 · 被引用 408 次
- Stabilizing Differentiable Architecture Search via Perturbation-based RegularizationXiangning Chen, Cho-Jui HsiehICML 2020 · 被引用 235 次
- Decoupled Greedy Learning of CNNsEugene Belilovsky, Michael Eickenberg, Edouard OyallonICML 2020 · 被引用 134 次
相关 Paper
- IDARTS: Interactive Differentiable Architecture SearchSong Xue, Runqi Wang, Baochang Zhang, Tian Wang 等ICCV 2021 · 被引用 10 次
- On the Acceleration of Deep Learning Model Parallelism With StalenessAn Xu, Zhouyuan Huo, Heng HuangCVPR 2020
- LoCo: Local Contrastive Representation LearningYuwen Xiong, Mengye Ren, Raquel UrtasunNeurIPS 2020 · 被引用 87 次
- Can Forward Gradient Match Backpropagation?Louis Fournier, Stéphane Rivaud, Eugene Belilovsky, Michael Eickenberg 等ICML 2023 · 被引用 33 次
- Neural Architecture Search for Lightweight Non-Local NetworksYingwei Li, Xiaojie Jin, Jieru Mei, Xiaochen Lian 等CVPR 2020
