Clockhands: Rename-free Instruction Set Architecture for Out-of-order Processors
Toru Koizumi, Ryota Shioya, Shu Sugita, Taichi Amano, Yuya Degawa, Junichiro Kadomoto, Hidetsugu Irie, Shuichi Sakai
Abstract
Out-of-order superscalar processors are currently the only architecture that speeds up irregular programs, but they suffer from poor power efficiency. To tackle this issue, we focused on how to specify register operands. Specifying operands by register names, as conventional RISC does, requires register renaming, resulting in poor power efficiency and preventing an increase in the front-end width. In contrast, a recently proposed architecture called STRAIGHT specifies operands by inter-instruction distance, thereby eliminating register renaming. However, STRAIGHT has strong constraints on instruction placement, which generally results in a large increase in the number of instructions.
We propose Clockhands, a novel instruction set architecture that has multiple register groups and specifies a value as "the value written in this register group 𝑘 times before." Clockhands does not require register renaming as in STRAIGHT. In contrast, Clockhands has much looser constraints on instruction placement than STRAIGHT, allowing programs to be written with almost the same number of instructions as Conventional RISC. We implemented a cycle-accurate simulator, FPGA implementation, and first-step compiler for Clockhands and evaluated benchmarks including SPEC CPU. On a machine with an eight-fetch width, the evaluation results showed that Clockhands consumes 7.4% less energy than RISC while having performance comparable to RISC. This energy reduction increases significantly to 24.4% when simulating a futuristic up-scaled processor with a 16-fetch width, which shows that Clockhands enables a wider front-end.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 092966a2-3705-4982-b911-a45422892be4Builds on3
- Delay and Bypass: Ready and Criticality Aware Instruction Scheduling in Out-of-Order ProcessorsMehdi Alipour, Stefanos Kaxiras, David Black-Schaffer, Rakesh KumarHPCA 2020 · 15 citations
- CASINO Core Microarchitecture: Generating Out-of-Order Schedules Using Cascaded In-Order Scheduling WindowsIpoom Jeong, Seihoon Park, Changmin Lee, Won Woo RoHPCA 2020 · 13 citations
- Reconstructing Out-of-Order Issue QueueIpoom Jeong, Jiwon Lee, Myung Kuk Yoon, Won Woo RoMICRO 2022 · 9 citations
Related papers
- Leveraging Targeted Value Prediction to Unlock New Hardware Strength Reduction PotentialArthur PeraisMICRO 2021 · 11 citations
- Precise Runahead ExecutionAjeya Naithani, Josué Feliu, Almutaz Adileh, Lieven EeckhoutHPCA 2020 · 32 citations
- Architecting Value Prediction around In-Order ExecutionPierre Ravenel, Arthur Perais, Benoît Dupont de Dinechin, Frédéric PétrotHPCA 2025 · 2 citations
- DiAG: a dataflow-inspired architecture for general-purpose processorsDong Kai Wang, Nam Sung KimASPLOS 2021 · 8 citations
- ATR: Out-of-Order Register Release Exploiting Atomic RegionsYinyuan Zhao, Surim Oh, Mingsheng Xu, Heiner LitzMICRO 2025 · 2 citations
