Nautilus: An Optimized System for Deep Transfer Learning over Evolving Training Datasets
Supun Nakandala, Arun Kumar
Abstract
Deep learning (DL) has revolutionized unstructured data analytics. But in most cases, DL needs massive labeled datasets and large compute clusters, which hinders its adoption. These limitations can be overcome using a popular paradigm called deep transfer learning (DTL). With DTL, one adapts a pre-trained DL model instead of training a model from scratch. Thus, DTL reduces the massive training data and compute requirements to train a model. During adaptation, a common practice is to freeze most pre-trained model parts and adapt only the remaining. Since no single adaptation scheme is universally the best, one often evaluates several schemes, which is also called model selection.
We also observed that data labeling for DTL is seldom a oneoff process. One often updates their labeled data intermittently by adding new labeled data and performs model selection to evaluate the accuracy of the trained models. Today, one executes this workload by performing computations for the entire pre-trained model and repeats it for every model selection cycle. This approach results in redundant computations in frozen model parts and causes usability and system inefficiency issues.
In this work, we reimagine DTL model selection in the presence of frozen layers as an instance of multi-query optimization and propose two optimizations that reduce redundant computations and training overheads. We implement our optimizations into a data system called Nautilus. Experiments with end-to-end workloads on benchmark datasets show that Nautilus reduces DTL model selection runtimes by up to 5X compared to the current practice.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers3
- Lotan: Bridging the Gap between GNNs and Scalable Graph Analytics EnginesYuhao Zhang, Arun KumarVLDB 2023 · 11 citations
- StarfishDB: A Query Execution Engine for Relational Probabilistic ProgrammingOuael Ben Amara, Sami Hadouaj, Niccolò MeneghettiSIGMOD 2024 · 2 citations
- Alsatian: Optimizing Model Search for Deep Transfer LearningNils Strassenburg, Boris Glavic, Tilmann RablSIGMOD 2025 · 2 citations
Builds on8
- ZeRO-Offload: Democratizing Billion-Scale Model TrainingJie Ren, Samyam Rajbhandari, Reza Yazdani Aminabadi, Olatunji Ruwase et al.USENIX ATC 2021 · 657 citations
- Analyzing and Mitigating Data Stalls in DNN TrainingJayashree Mohan, Amar Phanishayee, Ashish Raniwala, Vijay ChidambaramVLDB 2021 · 142 citations
- Cerebro: A Data System for Optimized Deep Learning Model SelectionSupun Nakandala, Yuhao Zhang, Arun KumarVLDB 2020 · 61 citations
- Distributed Deep Learning on Data Systems: A Comparative Analysis of ApproachesYuhao Zhang, Frank Mcquillan, Nandish Jayaram, Nikhil Kak et al.VLDB 2021 · 35 citations
- LIMA: Fine-grained Lineage Tracing and Reuse in Machine Learning SystemsArnab Phani, Benjamin Rath, Matthias BoehmSIGMOD 2021 · 30 citations
Related papers
- Vista: Optimized System for Declarative Feature Transfer from Deep CNNs at ScaleSupun Nakandala, Arun KumarSIGMOD 2020 · 12 citations
- AdaFilter: Adaptive Filter Fine-Tuning for Deep Transfer LearningYunhui Guo, Yandong Li, Liqiang Wang, Tajana RosingAAAI 2020 · 44 citations
- A Visual Analytics Framework for Explaining and Diagnosing Transfer Learning ProcessesYuxin Ma, Arlen Fan, Jingrui He, Arun Reddy Nelakurthi et al.IEEE VIS 2020 · 36 citations
- Automated Deep Learning Optimization via DSL-Based Source Code TransformationRuixin Wang, Minghai Lu, Cody Hao Yu, Yi-Hsiang Lai et al.ISSTA 2024 · 1 citation
- Selectivity Drives Productivity: Efficient Dataset Pruning for Enhanced Transfer LearningYihua Zhang, Yimeng Zhang, Aochuan Chen, Jinghan Jia et al.NeurIPS 2023 · 18 citations
