On inter-operator data transfers in query processing
Harshad Deshmukh, Bruhathi Sundarmurthy, Jignesh M. Patel
摘要
In designing query processing primitives, a crucial design choice is the method for data transfer between two operators in a query plan. As we were considering this critical design mechanism for an in-memory database system that we are building, we quickly realized that (surprisingly) there isn't a clear definition of this concept. Papers are full of ad hoc use of terms like pipelining and blocking, but these terms are not crisply defined, making it hard to fully understand the results attributed to these concepts. To address this limitation, we introduce a clear terminology for how to think about data transfer between operators in a query pipeline. We argue that there isn't a clear definition of pipelining and blocking, and that there is a full spectrum of techniques based on a simple concept called unit-of-transfer. Next, we develop an analytical model for inter-operator communication, and highlight the key parameters that impact performance (for in-memory database settings). Armed with this model, we then apply it to the system we are designing and highlight the insights that we gathered from this exercise. We find that the gap between the traditional “pipelining” and “non-pipelining” methods of query processing, w.r.t. key factors such as performance and memory footprint is quite narrow, and thus system designers should likely rethink the notion of “pipelining” vs. “blocking” for in-memory database systems.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- Taking Analytic Databases to the BankAlexandar Devic, Martin Prammer, Kevin P. Gaffney, Siddhartha Balakrishna Rai 等ISCA 2026
- FPGA for Aggregate Processing: The Good, The Bad, and The UglyZubeyr F. Eryilmaz, Aarati Kakaraparthy, Jignesh M. Patel, Rathijit Sen 等ICDE 2021 · 被引用 12 次
- No Cap, This Memory Slaps: Breaking Through the Memory Wall of Transactional Database Systems with Processing-in-MemoryHyoungjoo Kim, Yiwei Zhao, Andrew Pavlo, Phillip B. GibbonsVLDB 2025 · 被引用 7 次
- Database Processing-in-Memory: An Experimental StudyTiago Rodrigo Kepe, Eduardo C. de Almeida, Marco A. Z. AlvesVLDB 2020 · 被引用 22 次
- Scaling GPU-Accelerated Databases beyond GPU Memory SizeYinan Li, Bailu Ding, Ziyun Wei, Lukas M. Maas 等VLDB 2025 · 被引用 7 次
