Parallelism-Optimizing Data Placement for Faster Data-Parallel Computations
Nirvik Baruah, Peter Kraft, Fiodar Kazhamiaka, Peter Bailis, Matei Zaharia
摘要
Systems performing large data-parallel computations, including online analytical processing (OLAP) systems like Druid and search engines like Elasticsearch, are increasingly being used for business-critical real-time applications where providing low query latency is paramount. In this paper, we investigate an underexplored factor in the performance of data-parallel queries: their parallelism. We find that to minimize the tail latency of data-parallel queries, it is critical to place data such that the data items accessed by each individual query are spread across as many machines as possible so that each query can leverage the computational resources of as many machines as possible. To optimize parallelism and minimize tail latency in real systems, we develop a novel parallelism-optimizing data placement algorithm that defines a linearly-computable measure of query parallelism, uses it to frame data placement as an optimization problem, and leverages a new optimization problem partitioning technique to scale to large cluster sizes. We apply this algorithm to popular systems such as Solr and MongoDB and show that it reduces p99 latency by 7-64% on data-parallel workloads.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper4
- AlpaServe: Statistical Multiplexing with Model Parallelism for Deep Learning ServingZhuohan Li, Lianmin Zheng, Yinmin Zhong, Vincent Liu 等OSDI 2023 · 被引用 211 次
- Decouple and Decompose: Scaling Resource Allocation with DeDeZhiying Xu, Minlan Yu, Francis Y. YanOSDI 2025 · 被引用 5 次
- SkyPIE: A Fast & Accurate Oracle for Object PlacementTiemo Bang, Chris Douglas, Natacha Crooks, Joseph M. HellersteinSIGMOD 2024
- A Resource-centric Analysis and Optimization of NoSQL Workloads using Distressed Resource Volume MetricGunika Verma, Aashutosh A V, Pooja Srinivas, Yogesh Simmhan 等VLDB 2026
它引用的顶会 Paper3
- Solving Large-Scale Granular Resource Allocation Problems Efficiently with POPDeepak Narayanan, Fiodar Kazhamiaka, Firas Abuzaid, Peter Kraft 等SOSP 2021 · 被引用 56 次
- Shard Manager: A Generic Shard Management Framework for Geo-distributed ApplicationsSangmin Lee, Zhenhua Guo, Omer Sunercan, Jun Ying 等SOSP 2021 · 被引用 18 次
- Data-Parallel Actors: A Programming Model for Scalable Query Serving SystemsPeter Kraft, Fiodar Kazhamiaka, Peter Bailis, Matei ZahariaNSDI 2022
相关 Paper
- Taking Analytic Databases to the BankAlexandar Devic, Martin Prammer, Kevin P. Gaffney, Siddhartha Balakrishna Rai 等ISCA 2026
- Cool, a COhort OnLine analytical processing systemZhongle Xie, Hongbin Ying, Cong Yue, Meihui Zhang 等ICDE 2020 · 被引用 4 次
- Scalable top-k retrieval with SpartaGali Sheffi, Dmitry Basin, Edward Bortnikov, David Carmel 等PPoPP 2020
- Airphant: Cloud-oriented Document IndexingSupawit Chockchowwat, Chaitanya Sood, Yongjoo ParkICDE 2022 · 被引用 5 次
- Terabyte-Scale Analytics in the Blink of an EyeBowen Wu, Wei Cui, Carlo Curino, Matteo Interlandi 等VLDB 2026 · 被引用 10 次
