Per-Bank Bandwidth Regulation of Shared Last-Level Cache for Real-Time Systems
Connor Sullivan, Alex Manley, Mohammad Alian, Heechul Yun
Abstract
Modern commercial-off-the-shelf (COTS) multicore processors have advanced memory hierarchies that enhance memory-level parallelism (MLP), which is crucial for high performance. To support high MLP, shared last-level caches (LLCs) are divided into multiple banks, allowing parallel access. However, uneven distribution of cache requests from the cores, especially when requests from multiple cores are concentrated on a single bank, can result in significant contention affecting all cores that access the cache. Such cache bank contention can even be maliciously induced-known as cache bank-aware denial-of-service (DoS) attacks-in order to jeopardize the system’s timing predictability. In this paper, we propose a per-bank bandwidth regulation approach for multi-banked shared LLC based multicore realtime systems. By regulating bandwidth on a per-bank basis, the approach aims to prevent unnecessary throttling of cache accesses to non-contended banks, thus improving overall performance (throughput) without compromising isolation benefits of throttling. We implement our approach on a RISC-V system-on-chip (SoC) platform using FireSim and evaluate extensively using both synthetic and real-world workloads. Our evaluation results show that the proposed per-bank regulation approach effectively protects real-time tasks from co-running cache bank-aware DoS attacks, and offers up to a performance improvement for the throttled benign best-effort tasks compared to prior bank-oblivious bandwidth throttling approaches.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext d8dd3ad4-a4c7-49e7-a72e-54f211777efdBuilds on3
- E-WarP: A System-wide Framework for Memory Bandwidth Profiling and ManagementParul Sohal, Rohan Tabish, Ulrich Drepper, Renato MancusoRTSS 2020 · 34 citations
- Making Powerful Enemies on NVIDIA GPUsTyler Yandrofski, Jingyuan Chen, Nathan Otterness, James H. Anderson et al.RTSS 2022 · 15 citations
- PolyRhythm: Adaptive Tuning of a Multi-Channel Attack Template for Timing InterferenceAo Li, Marion Sudvarg, Han Liu, Zhiyuan Yu et al.RTSS 2022 · 12 citations
Related papers
- FaSe: fast selective flushing to mitigate contention-based cache timing attacksTuo Li, Sri ParameswaranDAC 2022 · 3 citations
- Predictable sharing of last-level cache partitions for multi-core safety-critical systemsZhuanhao Wu, Hiren D. PatelDAC 2022 · 6 citations
- HybCache: Hybrid Side-Channel-Resilient Caches for Trusted Execution EnvironmentsGhada Dessouky, Tommaso Frassetto, Ahmad-Reza SadeghiUSENIX Security 2020
- Co-Located Parallel Scheduling of Threads to Optimize Cache SharingCorey Tessler, Venkata Prashant Modekurthy, Nathan Fisher, Abusayeed Saifullah et al.RTSS 2023 · 5 citations
- Precise and scalable shared cache contention analysis for WCET estimationWei Zhang, Mingsong Lv, Wanli Chang, Lei JuDAC 2022 · 12 citations
