HILP: Accounting for Workload-Level Parallelism in System-on-Chip Design Space Exploration
Joseph Rogers, Lieven Eeckhout, Magnus Jahre
摘要
High-performance System-on-Chip (SoC) architectures are becoming increasingly complex and heterogeneous, and the days when a single application could utilize all of an SoC's hardware resources are all but over. The SoC's workload, i.e., the set of independent applications that the SoC typically executes, therefore has a significant impact on its efficiency. Accounting for Workload-Level Parallelism (WLP) in early-stage design space exploration is thus critical as later-stage analysis steps must focus on favorable design points to yield optimal results. Unfortunately, state-of-the-art MultiAmdahl and Gables fall short because they only model the extremes of minimal and maximal WLP.
We hence propose HILP, the first early-stage design space exploration approach for heterogeneous SoCs that fully accounts for WLP. Our key observation is that scheduling a workload of independent multi-phase applications on a heterogeneous SoC is an instance of the classic job-shop scheduling optimization problem and thus can be solved using integer linear programming. HILP therefore uses a high-performance integer linear programming solver to find a near-optimal schedule that minimizes the overall execution time of the workload, i.e., it schedules the dependent phases of all applications in the workload on the cores and accelerators of the target SoC to maximize performance while respecting power consumption and memory bandwidth constraints. We validate HILP by demonstrating that it captures the performance effects of Amdahl's law, the memory wall, and dark silicon, and then use it to explore the impact of WLP across a large SoC design space, yielding multiple insights. The key takeaway is that modeling WLP is necessary to ensure that more detailed, later-stage design tasks focus on the most favorable parts of the vast design space of heterogeneous SoCs.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper3
- AutoTM: Automatic Tensor Movement in Heterogeneous Memory Systems using Integer Linear ProgrammingMark Hildebrand, Jawad Khan, Sanjeev Trika, Jason Lowe-Power 等ASPLOS 2020 · 被引用 70 次
- HSM: A Hybrid Slowdown Model for Multitasking GPUsXia Zhao, Magnus Jahre, Lieven EeckhoutASPLOS 2020 · 被引用 36 次
- AIO: An Abstraction for Performance Analysis Across Diverse Accelerator ArchitecturesJoseph Rogers, Taha Soliman, Magnus JahreISCA 2024 · 被引用 5 次
相关 Paper
- ARTEMIS: Agile Discovery of Efficient Real-Time Systems-on-Chips in the Heterogeneous EraSubhankar Pal, Aporva Amarnath, Behzad Boroujerdian, Augusto Vega 等HPCA 2025 · 被引用 2 次
- PCCS: Processor-Centric Contention-aware Slowdown Model for Heterogeneous System-on-ChipsYuanchao Xu, Mehmet Esat Belviranli, Xipeng Shen, Jeffrey S. VetterMICRO 2021 · 被引用 15 次
- Automated Task Scheduling for Cloth and Deformable Body Simulations in Heterogeneous Computing EnvironmentsChengzhu He, Zhendong Wang, Zhaorui Meng, Junfeng Yao 等SIGGRAPH 2025 · 被引用 2 次
- Heterogeneity-Aware Cluster Scheduling Policies for Deep Learning WorkloadsDeepak Narayanan, Keshav Santhanam, Fiodar Kazhamiaka, Amar Phanishayee 等OSDI 2020 · 被引用 286 次
- Shared Memory-contention-aware Concurrent DNN Execution for Diversely Heterogeneous System-on-ChipsIsmet Dagli, Mehmet E. BelviranliPPoPP 2024 · 被引用 18 次
