Trident: Task Scheduling over Tiered Storage Systems in Big Data Platforms
Herodotos Herodotou, Elena Kakoulli
摘要
The recent advancements in storage technologies have popularized the use of tiered storage systems in data-intensive compute clusters. The Hadoop Distributed File System (HDFS), for example, now supports storing data in memory, SSDs, and HDDs, while OctopusFS and hatS offer fine-grained storage tiering solutions. However, the task schedulers of big data platforms (such as Hadoop and Spark) will assign tasks to available resources only based on data locality information, and completely ignore the fact that local data is now stored on a variety of storage media with different performance characteristics. This paper presents Trident, a principled task scheduling approach that is designed to make optimal task assignment decisions based on both locality and storage tier information. Trident formulates task scheduling as a minimum cost maximum matching problem in a bipartite graph and uses a standard solver for finding the optimal solution. In addition, Trident utilizes two novel pruning algorithms for bounding the size of the graph, while still guaranteeing optimality. Trident is implemented in both Spark and Hadoop, and evaluated extensively using a realistic workload derived from Facebook traces as well as an industry-validated benchmark, demonstrating significant benefits in terms of application performance and cluster efficiency.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- Fine-Grained Modeling and Optimization for Intelligent Resource Management in Big Data ProcessingChenghao Lyu, Qi Fan, Fei Song, Arnab Sinha 等VLDB 2022 · 被引用 14 次
- A Spark Optimizer for Adaptive, Fine-Grained Parameter TuningChenghao Lyu, Qi Fan, Philippe Guyard, Yanlei DiaoVLDB 2024 · 被引用 9 次
- S/C: Speeding up Data Materialization with Bounded MemoryZhaoheng Li, Xinyu Pi, Yongjoo ParkICDE 2023 · 被引用 7 次
它引用的顶会 Paper1
相关 Paper
- Adaptive Low-level Storage of Very Large Knowledge GraphsJacopo Urbani, Ceriel J. H. JacobsWWW 2020 · 被引用 10 次
- TCO-driven Storage Provisioning for Exascale Data CentersTimothy Kim, Saurabh Kadekodi, Arif Merchant, Prashant Nema 等EuroSys 2026
- Spark-based Cloud Data Analytics using Multi-Objective OptimizationFei Song, Khaled Zaouk, Chenghao Lyu, Arnab Sinha 等ICDE 2021 · 被引用 15 次
- Concealing Compression-accelerated I/O for HPC Applications through In Situ Task SchedulingSian Jin, Sheng Di, Frédéric Vivien, Daoce Wang 等EuroSys 2024 · 被引用 13 次
- Boosting Task Scheduling Data Locality with Low-latency, HW-accelerated Label PropagationLucas Morais, Juan Miguel De Haro Ruiz, Alfredo Goldman, Guido Araujo 等MICRO 2025 · 被引用 1 次
