DyRep: Bootstrapping Training with Dynamic Re-parameterization
Tao Huang, Shan You, Bohan Zhang, Yuxuan Du, Fei Wang, Chen Qian, Chang Xu
摘要
Structural re-parameterization (Rep) methods achieve noticeable improvements on simple VGG-style networks. Despite the prevalence, current Rep methods simply re-parameterize all operations into an augmented network, including those that rarely contribute to the model's performance. As such, the price to pay is an expensive computational overhead to manipulate these unnecessary behaviors. To eliminate the above caveats, we aim to boot-strap the training with minimal cost by devising a dynamic re-parameterization (DyRep) method, which encodes Rep technique into the training process that dynamically evolves the network structures. Concretely, our proposal adaptively finds the operations which contribute most to the loss in the network, and applies Rep to enhance their representational capacity. Besides, to suppress the noisy and redundant operations introduced by Rep, we devise a de-parameterization technique for a more compact re-parameterization. With this regard, DyRep is more efficient than Rep since it smoothly evolves the given network instead of constructing an over-parameterized network. Experimental results demonstrate our effectiveness, e.g., DyRep improves the accuracy of ResNet-18 by 2.04% on ImageNet and reduces 22% runtime over the baseline. Code is avail-able at: https://github.com/hunto/DyRep.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Knowledge Distillation from A Stronger TeacherTao Huang, Shan You, Fei Wang, Chen Qian 等NeurIPS 2022 · 被引用 477 次
- Depth-Supervised Fusion Network for Seamless-Free Image StitchingZhiying Jiang, Ruhao Yan, Zengxi Zhang, Bowei Zhang 等NeurIPS 2025 · 被引用 3 次
- RepAn: Enhanced Annealing through Re-parameterizationXiang Fei, Xiawu Zheng, Yan Wang, Fei Chao 等CVPR 2024
- EMFormer: Efficient Multi-Scale Transformer for Accumulative Context Weather Forecastinghao chen, Tao Han, Jie ZHANG, Song Guo 等ICML 2026
- Reparameterization through Spatial Gradient ScalingAlexander Detkov, Mohammad Salameh, Muhammad Fetrat Qharabagh, Jialin Zhang 等ICLR 2023
它引用的顶会 Paper10
- Pruning neural networks without any data by iteratively conserving synaptic flowHidenori Tanaka, Daniel Kunin, Daniel L. K. Yamins, Surya GanguliNeurIPS 2020 · 被引用 884 次
- Picking Winning Tickets Before Training by Preserving Gradient FlowChaoqi Wang, Guodong Zhang, Roger B. GrosseICLR 2020 · 被引用 743 次
- Instances as QueriesYuxin Fang, Shusheng Yang, Xinggang Wang, Yu Li 等ICCV 2021 · 被引用 331 次
- Weakly Supervised Contrastive LearningMingkai Zheng, Fei Wang, Shan You, Chen Qian 等ICCV 2021 · 被引用 153 次
- Locally Free Weight Sharing for Network Width SearchXiu Su, Shan You, Tao Huang, Fei Wang 等ICLR 2021 · 被引用 45 次
相关 Paper
- RepVGG: Making VGG-Style ConvNets Great AgainXiaohan Ding, Xiangyu Zhang, Ningning Ma, Jungong Han 等CVPR 2021
- Online Convolutional ReparameterizationMu Hu, Junyi Feng, Jiashen Hua, Baisheng Lai 等CVPR 2022 · 被引用 90 次
- ResRep: Lossless CNN Pruning via Decoupling Remembering and ForgettingXiaohan Ding, Tianxiang Hao, Jianchao Tan, Ji Liu 等ICCV 2021 · 被引用 202 次
- RepSR: Training Efficient VGG-style Super-Resolution Networks with Structural Re-Parameterization and Batch NormalizationXintao Wang, Chao Dong, Ying ShanACM MM 2022 · 被引用 44 次
- Make RepVGG Greater Again: A Quantization-Aware ApproachXiangxiang Chu, Liang Li, Bo ZhangAAAI 2024 · 被引用 70 次
