Optimizing An In-memory Database System For AI-powered On-line Decision Augmentation Using Persistent Memory
Cheng Chen, Jun Yang, Mian Lu, Taize Wang, Zhao Zheng, Yuqiang Chen, Wenyuan Dai, Bingsheng He, Weng-Fai Wong, Guoan Wu, Yuping Zhao, Andy Rudoff
摘要
On-line decision augmentation (OLDA) has been considered as a promising paradigm for real-time decision making powered by Artificial Intelligence (AI). OLDA has been widely used in many applications such as real-time fraud detection, personalized recommendation, etc. On-line inference puts real-time features extracted from multiple time windows through a pre-trained model to evaluate new data to support decision making. Feature extraction is usually the most time-consuming operation in many OLDA data pipelines. In this work, we started by studying how existing in-memory databases can be leveraged to efficiently support such real-time feature extractions. However, we found that existing in-memory databases cost hundreds or even thousands of milliseconds. This is unacceptable for OLDA applications with strict real-time constraints. We therefore propose FEDB ( F eature E ngineering D ata b ase), a distributed in-memory database system designed to efficiently support on-line feature extraction. Our experimental results show that FEDB can be one to two orders of magnitude faster than the state-of-the-art in-memory databases on real-time feature extraction. Furthermore, we explore the use of the Intel Optane DC Persistent Memory Module (PMEM) to make FEDB more cost-effective. When comparing the proposed PMEM-optimized persistent skiplist to the FEDB using DRAM+SSD, PMEM-based FEDB can shorten the tail latency up to 19.7%, reduce the recovery time up to 99.7%, and save up to 58.4% total cost of a real OLDA pipeline.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Optimizing Data Pipelines for Machine Learning in Feature StoresRui Liu, Kwanghyun Park, Fotis Psallidas, Xiaoyong Zhu 等VLDB 2023 · 被引用 10 次
- Biathlon: Harnessing Model Resilience for Accelerating ML Inference PipelinesChaokun Chang, Eric Lo, Chunxiao YeVLDB 2024 · 被引用 5 次
它引用的顶会 Paper3
- FlatStore: An Efficient Log-Structured Key-Value Storage Engine for Persistent MemoryYoumin Chen, Youyou Lu, Fan Yang, Qing Wang 等ASPLOS 2020 · 被引用 166 次
- Deep Learning Models for Selectivity Estimation of Multi-Attribute QueriesShohedul Hasan, Saravanan Thirumuruganathan, Jees Augustine, Nick Koudas 等SIGMOD 2020 · 被引用 101 次
- Dash: Scalable Hashing on Persistent MemoryBaotong Lu, Xiangpeng Hao, Tianzheng Wang, Eric LoVLDB 2020 · 被引用 8 次
相关 Paper
- PerMA-Bench: Benchmarking Persistent Memory AccessLawrence Benson, Leon Papke, Tilmann RablVLDB 2022 · 被引用 19 次
- APEX: A High-Performance Learned Index on Persistent MemoryBaotong Lu, Jialin Ding, Eric Lo, Umar Farooq Minhas 等VLDB 2022 · 被引用 73 次
- Single Machine Graph Analytics on Massive Datasets Using Intel Optane DC Persistent MemoryGurbinder Gill, Roshan Dathathri, Loc Hoang, Ramesh Peri 等VLDB 2020 · 被引用 82 次
- Maximizing Persistent Memory Bandwidth Utilization for OLAP WorkloadsBjörn Daase, Lars Jonas Bollmeier, Lawrence Benson, Tilmann RablSIGMOD 2021 · 被引用 38 次
- PetPS: Supporting Huge Embedding Models with Persistent MemoryMinhui Xie, Youyou Lu, Qing Wang, Yangyang Feng 等VLDB 2023 · 被引用 10 次
