Co-Located Parallel Scheduling of Threads to Optimize Cache Sharing
Corey Tessler, Venkata Prashant Modekurthy, Nathan Fisher, Abusayeed Saifullah, Alleyn Murphy
Abstract
For hard-real time systems, cache memory increases execution time variability, increasing the complexity of timing analysis. As such, cache memory is often treated exclusively as a detractor to schedulability. Cache-aware co-located scheduling aims to improve schedulability by carefully scheduling threads to share cached values. Cache sharing between threads potentially reduces task execution times and increases schedulability with fewer resources. Antithetically, co-located scheduling may reduce parallelism, decreasing efficiency. Thus, identifying the optimal set of threads to co-locate that minimizes the resources required while ensuring timing constraints is a complex challenge. This work establishes optimal co-location as NP-Hard in the strong sense. It offers an approximation method for the co-located scheduling of Fork-Join tasks named 3-PARM-HD. The approximation has a 3-factor guarantee and a resource augmentation bound of 3. The simulated evaluation shows 3-PARM-HD increases schedulability compared to an optimal intractable algorithm (without co-location) scheduling 28% more tasks with 30% fewer cores. Simulated results show 3-PARM-HD outperforms a 2-factor approximation for traditional makespan, scheduling 39% more tasks with 44% fewer cores. An experimental RISC-V evaluation running on a QEMU platform confirms the benefits of 3-PARM-HD, scheduling and executing tasks deemed unschedulable by a 2-factor makespan approximation without co-location.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get 445fc56f-73f5-4ddf-8b7a-dd20a9c8e8e0Related papers
- Co-Optimizing Cache Partitioning and Multi-Core Task Scheduling: Exploit Cache Sensitivity or Not?Binqi Sun, Debayan Roy, Tomasz Kloda, Andrea Bastoni et al.RTSS 2023 · 4 citations
- Precise and scalable shared cache contention analysis for WCET estimationWei Zhang, Mingsong Lv, Wanli Chang, Lei JuDAC 2022 · 12 citations
- A Cache/Algorithm Co-design for Parallel Real-Time Systems with Data Dependency on Multi/Many-core System-on-ChipsZhe Jiang, Shuai Zhao, Ran Wei, Yiyang Gao et al.DAC 2024 · 3 citations
- Per-Bank Bandwidth Regulation of Shared Last-Level Cache for Real-Time SystemsConnor Sullivan, Alex Manley, Mohammad Alian, Heechul YunRTSS 2024 · 4 citations
- Tight Cache Contention Analysis for WCET Estimation on Multicore SystemsShuai Zhao, Jieyu Jiang, Shenlin Cai, Yaowei Liang et al.RTSS 2025 · 1 citation
