Bayesian Optimization Meets Self-Distillation
HyunJae Lee, Heon Song, Hyeonsoo Lee, Gihyeon Lee, Suyeong Park, Donggeun Yoo
摘要
Bayesian optimization (BO) has contributed greatly to improving model performance by suggesting promising hyperparameter configurations iteratively based on observations from multiple training trials. However, only partial knowledge (i.e., the measured performances of trained models and their hyperparameter configurations) from previous trials is transferred. On the other hand, Self-Distillation (SD) only transfers partial knowledge learned by the task model itself. To fully leverage the various knowledge gained from all training trials, we propose the BOSS framework, which combines BO and SD. BOSS suggests promising hyperparameter configurations through BO and carefully selects pre-trained models from previous trials for SD, which are otherwise abandoned in the conventional BO process. BOSS achieves significantly better performance than both BO and SD in a wide range of tasks including general image classification, learning with noisy labels, semi-supervised learning, and medical image analysis tasks. Our code is available at https://github.com/sooperset/boss .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper6
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- SRM: A Style-Based Recalibration Module for Convolutional Neural NetworksHyunJae Lee, Hyo-Eun Kim, Hyeonseob NamICCV 2019 · 被引用 286 次
- Towards Understanding Ensemble, Knowledge Distillation and Self-Distillation in Deep LearningZeyuan Allen-Zhu, Yuanzhi LiICLR 2023 · 被引用 151 次
- Knowledge Distillation: Bad Models Can Be Good Role ModelsGal Kaplun, Eran Malach, Preetum Nakkiran, Shai Shalev-ShwartzNeurIPS 2022 · 被引用 19 次
- Regularizing Class-Wise Predictions via Self-Knowledge DistillationSukmin Yun, Jongjin Park, Kimin Lee, Jinwoo ShinCVPR 2020
相关 Paper
- : Unlocking the Performance Ceiling for Pretrained OptimizersMuqi Han, Ruoqi Xing, KAI WU, Xiaoyu Zhang 等ICML 2026
- Self-supervised Label Augmentation via Input TransformationsHankook Lee, Sung Ju Hwang, Jinwoo ShinICML 2020 · 被引用 218 次
- Systematic comparison of semi-supervised and self-supervised learning for medical image classificationZhe Huang, Ruijie Jiang, Shuchin Aeron, Michael C. HughesCVPR 2024
- Refine Myself by Teaching Myself: Feature Refinement via Self-Knowledge DistillationMingi Ji, Seungjae Shin, Seunghyun Hwang, Gibeom Park 等CVPR 2021
- A Quantile-based Approach for Hyperparameter Transfer LearningDavid Salinas, Huibin Shen, Valerio PerroneICML 2020 · 被引用 50 次
