Exploring Instruction Fusion Opportunities in General Purpose Processors
Sawan Singh, Arthur Perais, Alexandra Jimborean, Alberto Ros
摘要
The Complex Instruction Set Computer (CISC) paradigm has led to the introduction of instruction cracking in which an architectural instruction is divided into multiple microarchitectural instructions (-ops). However, the dual concept, instruction fusion is also prevalent in modern microarchitectures to maximize resource utilization. In essence, some architectural instructions are too complex to be executed as a unit, so they should be cracked, while others are too simple to waste resources on executing them as a unit, so they should be fused with others. In this paper, we focus on instruction fusion and explore opportunities for fusing additional instructions in a high-performance general purpose pipeline. We show that enabling fusion for common RISC-V idioms improves performance by 7%. Then, we determine experimentally that enabling fusion only for memory instructions achieves 86% of the potential of fusion in this particular case. Finally, we propose the Helios microarchitecture, able to fuse non-consecutive and noncontiguous memory instructions, and discuss microarchitectural changes required to do so efficiently while preserving correctness. Helios allows to fuse an additional 5.5% of dynamic instructions, yielding a 14.2% performance uplift over no fusion (8.2% over baseline fusion).
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper1
相关 Paper
- Improving the Utilization of Micro-operation Caches in x86 ProcessorsJagadish B. Kotra, John KalamatianosMICRO 2020 · 被引用 8 次
- Dual-Issue Execution of Mixed Integer and Floating-Point Workloads on Energy-Efficient In-Order RISC-V CoresLuca Colagrande, Luca BeniniDAC 2025 · 被引用 2 次
- ARCANE: Adaptive RISC-V Cache Architecture for Near-memory ExtensionsVincenzo Petrolo, Flavia Guella, Michele Caon, Pasquale Davide Schiavone 等DAC 2025 · 被引用 1 次
- GoPTX: Fine-grained GPU Kernel Fusion by PTX-level Instruction Flow WeavingKan Wu, Zejia Lin, Mengyue Xi, Zhongchun Zheng 等DAC 2025 · 被引用 1 次
- Co-Utilizing SIMD and Scalar to Accelerate the Data Analytics WorkloadsZewen Sun, Zhifang Li, Chuliang WengICDE 2023
