MorphingDB: A Task-Centric AI-Native DBMS for Model Management and Inference
Sai Wu, Ruichen Xia, Dingyu Yang, Rui Wang, Huihang Lai, Jiarui Guan, Jiameng Bai, Dongxiang Zhang, Xiu Tang, Zhongle Xie, Peng Lu, Gang Chen
摘要
The increasing demand for deep neural inference within database environments has driven the emergence of AI?native DBMSs. However, existing solutions either rely on model-centric designs requiring developers to manually select, configure, and maintain models, resulting in high development overhead, or adopt task-centric AutoML approaches with high computational costs and poor DBMS integration. We present MorphingDB, a task-centric AI-native DBMS that automates model storage, selection, and inference within PostgreSQL. To enable flexible, I/O-efficient storage of deep learning models, we first introduce specialized schemas and multi-dimensional tensor data types to support BLOB-based all-in-one and decoupled model storage. Then we design a transfer learning framework for model selection in two phases, which builds a transferability subspace via offline embedding of historical tasks and employs online projection through feature-aware mapping for real-time tasks. To further optimize inference throughput, we propose pre-embedding with vectoring sharing to eliminate redundant computations and DAG-based batch pipelines with cost-aware scheduling to minimize the inference time. Implemented as a PostgreSQL extension with LibTorch, MorphingDB outperforms AI-native DBMSs (EvaDB, Madlib, GaussML) and AutoML platforms (AutoGluon, AutoKeras, AutoSklearn) across nine public datasets, encompassing series, NLP, and image tasks. Our evaluation demonstrates a robust balance among accuracy, resource consumption, and time cost in model selection and significant gains in throughput and resource efficiency.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper11
- ALBERT: A Lite BERT for Self-supervised Learning of Language RepresentationsZhenzhong Lan, Mingda Chen, Sebastian Goodman, Kevin Gimpel 等ICLR 2020 · 被引用 7,418 次
- TSMixer: Lightweight MLP-Mixer Model for Multivariate Time Series ForecastingVijay Ekambaram, Arindam Jati, Nam Nguyen, Phanwadee Sinthong 等KDD 2023 · 被引用 221 次
- Scalable Diverse Model Selection for Accessible Transfer LearningDaniel Bolya, Rohit Mittapalli, Judy HoffmanNeurIPS 2021 · 被引用 61 次
- End-to-end Optimization of Machine Learning Prediction QueriesKwanghyun Park, Karla Saur, Dalitso Banda, Rathijit Sen 等SIGMOD 2022 · 被引用 50 次
- Vexless: A Serverless Vector Data Management System Using Cloud FunctionsYongye Su, Yinqi Sun, Minjia Zhang, Jianguo WangSIGMOD 2024 · 被引用 27 次
相关 Paper
- Distributed Deep Learning on Data Systems: A Comparative Analysis of ApproachesYuhao Zhang, Frank Mcquillan, Nandish Jayaram, Nikhil Kak 等VLDB 2021 · 被引用 35 次
- Powering In-Database Dynamic Model Slicing for Structured Data AnalyticsLingze Zeng, Naili Xing, Shaofeng Cai, Gang Chen 等VLDB 2024 · 被引用 7 次
- ArrayMorph: Optimizing Hyperslab Queries on the Cloud for Machine Learning PipelinesRuochen Jiang, Spyros BlanasVLDB 2025
- CactusDB: Unlock Co-Optimization Opportunities for SQL Queries and AI/ML Model InferencesLixi Zhou, Kanchan Chowdhury, Lulu Xie, Jaykumar Tandel 等ICDE 2026
- EVA: A Symbolic Approach to Accelerating Exploratory Video Analytics with Materialized ViewsZhuangdi Xu, Gaurav Tarlok Kakkar, Joy Arulraj, Umakishore RamachandranSIGMOD 2022 · 被引用 26 次
