Delay and Bypass: Ready and Criticality Aware Instruction Scheduling in Out-of-Order Processors
Mehdi Alipour, Stefanos Kaxiras, David Black-Schaffer, Rakesh Kumar
Abstract
Flexible instruction scheduling is essential for performance in out-of-order processors. This is typically achieved by using CAM-based Instruction Queues (IQs) that provide complete flexibility in choosing ready instructions for execution, but at the cost of significant scheduling energy.
In this work we seek to reduce the instruction scheduling energy by reducing the depth and width of the IQ. We do so by classifying instructions based on their readiness and criticality, and using this information to bypass the IQ for instructions that will not benefit from its expensive scheduling structures and delay instructions that will not harm performance. Combined, these approaches allow us to offload a significant portion of the instructions from the IQ to much cheaper FIFO-based scheduling structures without hurting performance. As a result we can reduce the IQ depth and width by half, thereby saving energy.
Our design, Delay and Bypass (DNB), is the first design to explicitly address both readiness and criticality to reduce scheduling energy. By handling both classes we are able to achieve 95% of the baseline out-of-order performance while only using 33% of the scheduling energy. This represents a significant improvement over previous designs which addressed only criticality or readiness (91%/89% performance at 74%/53% energy).
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f050802f-8b0a-42fc-8be0-c486ea8994d3Cited by top-tier papers4
- CRISP: critical slice prefetchingHeiner Litz, Grant Ayers, Parthasarathy RanganathanASPLOS 2022 · 33 citations
- Harvesting Memory-bound CPU Stall Cycles in Software with MSHZhihong Luo, Sam Son, Sylvia Ratnasamy, Scott ShenkerOSDI 2024 · 5 citations
- Criticality-Aware Instruction-Centric Bandwidth Partitioning for Data Center ApplicationsLiren Zhu, Liujia Li, Jianyu Wu, Yiming Yao et al.HPCA 2025 · 4 citations
- Clockhands: Rename-free Instruction Set Architecture for Out-of-order ProcessorsToru Koizumi, Ryota Shioya, Shu Sugita, Taichi Amano et al.MICRO 2023 · 4 citations
Related papers
- Reconstructing Out-of-Order Issue QueueIpoom Jeong, Jiwon Lee, Myung Kuk Yoon, Won Woo RoMICRO 2022 · 9 citations
- Orinoco: Ordered Issue and Unordered Commit with Non-Collapsible QueuesDibei Chen, Tairan Zhang, Yi Huang, Jianfeng Zhu et al.ISCA 2023 · 1 citation
- CASINO Core Microarchitecture: Generating Out-of-Order Schedules Using Cascaded In-Order Scheduling WindowsIpoom Jeong, Seihoon Park, Changmin Lee, Won Woo RoHPCA 2020 · 13 citations
- NOREBA: a compiler-informed non-speculative out-of-order commit processorAli Hajiabadi, Andreas Diavastos, Trevor E. CarlsonASPLOS 2021 · 7 citations
- Localizing the Tag Comparisons in the Wakeup Logic to Reduce Energy Consumption of the Issue QueueKenichiro Mori, Sota Kosugi, Hiroto Yoshida, Hajime Shimada et al.MICRO 2024 · 1 citation
