DEPARA: Deep Attribution Graph for Deep Knowledge Transferability
Jie Song, Yixin Chen, Jingwen Ye, Xinchao Wang, Chengchao Shen, Feng Mao, Mingli Song
Abstract
Exploring the intrinsic interconnections between the knowledge encoded in PRe-trained Deep Neural Networks (PR-DNNs) of heterogeneous tasks sheds light on their mutual transferability, and consequently enables knowledge transfer from one task to another so as to reduce the training effort of the latter. In this paper, we propose the DEeP Attribution gRAph (DEPARA) to investigate the transferability of knowledge learned from PR-DNNs. In DEPARA, nodes correspond to the inputs and are represented by their vectorized attribution maps with regards to the outputs of the PR-DNN. Edges denote the relatedness between inputs and are measured by the similarity of their features extracted from the PR-DNN. The knowledge transferability of two PR-DNNs is measured by the similarity of their corresponding DEPARAs. We apply DEPARA to two important yet understudied problems in transfer learning: pre-trained model selection and layer selection. Extensive experiments are conducted to demonstrate the effectiveness and superiority of the proposed method in solving both these problems. Code, data and models reproducing the results in this paper are available at https://github.com/zju-vipa/ DEPARA .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 2e4f2ed6-b03b-40de-bd82-732277ae6098Cited by top-tier papers13
- Factorizable Graph Convolutional NetworksYiding Yang, Zunlei Feng, Mingli Song, Xinchao WangNeurIPS 2020 · 175 citations
- Task Switching Network for Multi-task LearningGuolei Sun, Thomas Probst, Danda Pani Paudel, Nikola Popovic et al.ICCV 2021 · 57 citations
- Model Spider: Learning to Rank Pre-Trained Models EfficientlyYi-Kai Zhang, Ting-Ji Huang, Yao-Xiang Ding, De-Chuan Zhan et al.NeurIPS 2023 · 57 citations
- Progressive Network Grafting for Few-Shot Knowledge DistillationChengchao Shen, Xinchao Wang, Youtan Yin, Jie Song et al.AAAI 2021 · 55 citations
- Meta Discovery: Learning to Discover Novel Classes given Very Limited DataHaoang Chi, Feng Liu, Wenjing Yang, Long Lan et al.ICLR 2022 · 52 citations
Builds on3
- Similarity-Preserving Knowledge DistillationFrederick Tung, Greg MoriICCV 2019 · 1,214 citations
- Rethinking ImageNet Pre-TrainingKaiming He, Ross B. Girshick, Piotr DollárICCV 2019 · 1,188 citations
- Correlation Congruence for Knowledge DistillationBaoyun Peng, Xiao Jin, Dongsheng Li, Shunfeng Zhou et al.ICCV 2019 · 625 citations
Related papers
- GraphControl: Adding Conditional Control to Universal Graph Pre-trained Models for Graph Domain Transfer LearningYun Zhu, Yaoke Wang, Haizhou Shi, Zhenshuo Zhang et al.WWW 2024 · 43 citations
- Understanding the Transferability of Representations via Task-RelatednessAkshay Mehra, Yunbei Zhang, Jihun HammNeurIPS 2024 · 13 citations
- Ranking Neural CheckpointsYandong Li, Xuhui Jia, Ruoxin Sang, Yukun Zhu et al.CVPR 2021
- Knowledge Consistency between Neural Networks and BeyondRuofan Liang, Tianlin Li, Longfei Li, Jing Wang et al.ICLR 2020 · 30 citations
- Model Selection with Model Zoo via Graph LearningZiyu Li, Hilco van der Wilk, Danning Zhan, Megha Khosla et al.ICDE 2024 · 6 citations
