Transfer Learning with Affine Model Transformation
Shunya Minami, Kenji Fukumizu, Yoshihiro Hayashi, Ryo Yoshida
摘要
Supervised transfer learning has received considerable attention due to its potential to boost the predictive power of machine learning in scenarios where data are scarce. Generally, a given set of source models and a dataset from a target domain are used to adapt the pre-trained models to a target domain by statistically learning domain shift and domain-specific factors. While such procedurally and intuitively plausible methods have achieved great success in a wide range of real-world applications, the lack of a theoretical basis hinders further methodological development. This paper presents a general class of transfer learning regression called affine model transfer, following the principle of expected-square loss minimization. It is shown that the affine model transfer broadly encompasses various existing methods, including the most common procedure based on neural feature extractors. Furthermore, the current paper clarifies theoretical properties of the affine model transfer such as generalization error and excess risk. Through several case studies, we demonstrate the practical benefits of modeling and estimating inter-domain commonality and domain-specific factors separately with the affine-type transfer models. *1.09 ± 0.232 *0.969 ± 0.144 *0.927 ± 0.170 Augmented 2.47 ± 0.406 1.90 ± 0.515 1.67 ± 0.552 *1.31 ± 0.214 1.16 ± 0.225 *0.984 ± 0.149 *0.897 ± 0.138 HTL-offset 2.29 ± 0.621 *1.69 ± 0.507 *1.49 ± 0.513 *1.22 ± 0.269 *1.09 ± 0.233 *0.969 ± 0.144 *0.925 ± 0.171 HTL-scale 2.32 ± 0.599 *1.71 ± 0.516 1.51 ± 0.513 *1.24 ± 0.271 *1.12 ± 0.234 *0.999 ± 0.175 0.948 ± 0.172 AffineTL-full *2.23 ± 0.554 *1.71 ± 0.501 *1.45 ± 0.458 *1.21 ± 0.256 *1.06 ± 0.219 *0.974 ± 0.164 *0.870 ± 0.121 AffineTL-const *2.30 ± 0.565 *1.73 ± 0.420 *1.48 ± 0.527 *1.20 ± 0.243 *1.04 ± 0.217 *0.963 ± 0.161 *0.884 ± 0.136
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper6
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- On the Theory of Transfer Learning: The Importance of Task DiversityNilesh Tripuraneni, Michael I. Jordan, Chi JinNeurIPS 2020 · 被引用 263 次
- Coupling-based Invertible Neural Networks Are Universal Diffeomorphism ApproximatorsTakeshi Teshima, Isao Ishikawa, Koichi Tojo, Kenta Oono 等NeurIPS 2020 · 被引用 129 次
- SciRepEval: A Multi-Format Benchmark for Scientific Document RepresentationsAmanpreet Singh, Mike D'Arcy, Arman Cohan, Doug Downey 等EMNLP 2023 · 被引用 45 次
- PAC-Net: A Model Pruning Approach to Inductive Transfer LearningSanghoon Myung, In Huh, Wonik Jang, Jae Myung Choe 等ICML 2022 · 被引用 18 次
相关 Paper
- Minimax Lower Bounds for Transfer Learning with Linear and One-hidden Layer Neural NetworksSeyed Mohammadreza Mousavi Kalan, Zalan Fabian, Salman Avestimehr, Mahdi SoltanolkotabiNeurIPS 2020 · 被引用 37 次
- A General Class of Transfer Learning Regression without Implementation CostShunya Minami, Song Liu, Stephen Wu, Kenji Fukumizu 等AAAI 2021 · 被引用 8 次
- Near-Optimal Linear Regression under Distribution ShiftQi Lei, Wei Hu, Jason D. LeeICML 2021 · 被引用 45 次
- Adversarial Training Helps Transfer Learning via Better RepresentationsZhun Deng, Linjun Zhang, Kailas Vodrahalli, Kenji Kawaguchi 等NeurIPS 2021 · 被引用 60 次
- Test-time Adaptation for Regression by Subspace AlignmentKazuki Adachi, Shin'ya Yamaguchi, Atsutoshi Kumagai, Tomoki HamagamiICLR 2025
