Neural Data Server: A Large-Scale Search Engine for Transfer Learning Data
Xi Yan, David Acuna, Sanja Fidler
摘要
Transfer learning has proven to be a successful technique to train deep learning models in the domains where little training data is available. The dominant approach is to pretrain a model on a large generic dataset such as ImageNet and finetune its weights on the target domain. However, in the new era of an ever increasing number of massive datasets, selecting the relevant data for pretraining is a critical issue. We introduce Neural Data Server (NDS), a large-scale search engine for finding the most useful transfer learning data to the target domain. NDS consists of a dataserver which indexes several large popular image datasets, and aims to recommend data to a client, an end-user with a target application with its own small labeled dataset. The dataserver represents large datasets with a much more compact mixture-of-experts model, and employs it to perform data search in a series of dataserverclient transactions at a low computational cost. We show the effectiveness of NDS in various transfer learning scenarios, demonstrating state-of-the-art performance on several target datasets and tasks such as image classification, object detection and instance segmentation. Neural Data Server is available as a web-service at http://aidemo s.cs.toronto.edu/nds/.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper22
- Divergence-aware Federated Self-Supervised LearningWeiming Zhuang, Yonggang Wen, Shuai ZhangICLR 2022 · 被引用 123 次
- Exploring and Predicting Transferability across NLP TasksTu Vu, Tong Wang, Tsendsuren Munkhdalai, Alessandro Sordoni 等EMNLP 2020 · 被引用 104 次
- Scalable Transfer Learning with Expert ModelsJoan Puigcerver, Carlos Riquelme Ruiz, Basil Mustafa, Cédric Renggli 等ICLR 2021 · 被引用 70 次
- Scalable Diverse Model Selection for Accessible Transfer LearningDaniel Bolya, Rohit Mittapalli, Judy HoffmanNeurIPS 2021 · 被引用 61 次
- Data Acquisition for Improving Machine Learning ModelsYifan Li, Xiaohui Yu, Nick KoudasVLDB 2021 · 被引用 57 次
它引用的顶会 Paper6
- Practical Secure Aggregation for Privacy-Preserving Machine LearningKallista A. Bonawitz, Vladimir Ivanov, Ben Kreuter, Antonio Marcedone 等CCS 2017 · 被引用 3,936 次
- Rethinking ImageNet Pre-TrainingKaiming He, Ross B. Girshick, Piotr DollárICCV 2019 · 被引用 1,188 次
- Gated-SCNN: Gated Shape CNNs for Semantic SegmentationTowaki Takikawa, David Acuna, Varun Jampani, Sanja FidlerICCV 2019 · 被引用 710 次
- Task2Vec: Task Embedding for Meta-LearningAlessandro Achille, Michael Lam, Rahul Tewari, Avinash Ravichandran 等ICCV 2019 · 被引用 359 次
- Meta-Sim: Learning to Generate Synthetic DatasetsAmlan Kar, Aayush Prakash, Ming-Yu Liu, Eric Cameracci 等ICCV 2019 · 被引用 272 次
相关 Paper
- Scalable Neural Data Server: A Data Recommender for Transfer LearningTianshi Cao, Sasha Doubov, David Acuna, Sanja FidlerNeurIPS 2021 · 被引用 7 次
- Which Model to Transfer? Finding the Needle in the Growing HaystackCédric Renggli, André Susano Pinto, Luka Rimanic, Joan Puigcerver 等CVPR 2022 · 被引用 13 次
- Task-Adaptive Neural Network Search with Meta-Contrastive LearningWonyong Jeong, Hayeon Lee, Geon Park, Eunyoung Hyung 等NeurIPS 2021 · 被引用 17 次
- NASTransfer: Analyzing Architecture Transferability in Large Scale Neural Architecture SearchRameswar Panda, Michele Merler, Mayoore S. Jaiswal, Hui Wu 等AAAI 2021 · 被引用 10 次
- GAIA: A Transfer Learning System of Object Detection That Fits Your NeedsXingyuan Bu, Junran Peng, Junjie Yan, Tieniu Tan 等CVPR 2021
