EMT-NAS: Transferring architectural knowledge between tasks from different datasets
Peng Liao, Yaochu Jin, Wenli Du
摘要
The success of multi-task learning (MTL) can largely be attributed to the shared representation of related tasks, allowing the models to better generalise. In deep learning, this is usually achieved by sharing a common neural network architecture and jointly training the weights. However, the joint training of weighting parameters on multiple related tasks may lead to performance degradation, known as negative transfer. To address this issue, this work proposes an evolutionary multi-tasking neural architecture search (EMT-NAS) algorithm to accelerate the search process by transferring architectural knowledge across multiple related tasks. In EMT-NAS, unlike the traditional MTL, the model for each task has a personalised network architecture and its own weights, thus offering the capability of effectively alleviating negative transfer. A fitness re-evaluation method is suggested to alleviate fluctuations in performance evaluations resulting from parameter sharing and the mini-batch gradient descent training method, thereby avoiding losing promising solutions during the search process. To rigorously verify the performance of EMT-NAS, the classification tasks used in the empirical assessments are derived from different datasets, including the CIFAR-10 and CIFAR-100, and four MedMNIST datasets. Extensive comparative experiments on different numbers of tasks demonstrate that EMT-NAS takes 8% and up to 40% on CIFAR and MedMNIST, respectively, less time to find competitive neural architectures than its single-task counterparts.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper12
- Searching for MobileNetV3Andrew Howard, Ruoming Pang, Hartwig Adam, Quoc V. Le 等ICCV 2019 · 被引用 9,163 次
- PC-DARTS: Partial Channel Connections for Memory-Efficient Architecture SearchYuhui Xu, Lingxi Xie, Xiaopeng Zhang, Xin Chen 等ICLR 2020 · 被引用 691 次
- AdaShare: Learning What To Share For Efficient Deep Multi-Task LearningXimeng Sun, Rameswar Panda, Rogério Feris, Kate SaenkoNeurIPS 2020 · 被引用 337 次
- Multi-Objective Meta LearningFeiyang Ye, Baijiong Lin, Zhixiong Yue, Pengxin Guo 等NeurIPS 2021 · 被引用 71 次
- TransTailor: Pruning the Pre-trained Model for Improved Transfer LearningBingyan Liu, Yifeng Cai, Yao Guo, Xiangqun ChenAAAI 2021 · 被引用 69 次
相关 Paper
- Automatic Multi-Task Learning Framework with Neural Architecture Search in RecommendationsShen Jiang, Guanghui Zhu, Yue Wang, Chunfeng Yuan 等KDD 2024 · 被引用 5 次
- MTL-NAS: Task-Agnostic Neural Architecture Search Towards General-Purpose Multi-Task LearningYuan Gao, Haoping Bai, Zequn Jie, Jiayi Ma 等CVPR 2020
- Towards Fast Adaptation of Neural Architectures with Meta LearningDongze Lian, Yin Zheng, Yintao Xu, Yanxiong Lu 等ICLR 2020 · 被引用 95 次
- M-NAS: Meta Neural Architecture SearchJiaxing Wang, Jiaxiang Wu, Haoli Bai, Jian ChengAAAI 2020 · 被引用 34 次
- Breaking Multi-Task Curse: Reward-Weighted Evolution for Black-Box Many-Task OptimizationYanchi Li, Jiao Liu, Wenyin Gong, Qiong Gu 等ICML 2026
