Enhancing Thread-Level Parallelism in Asymmetric Multicores using Transparent Instruction Offloading
Jeckson Dellagostin Souza, Madhavan Manivannan, Miquel Pericàs, Antonio Carlos Schneider Beck
摘要
Asymmetric multicore architectures (AMC) with single-ISA can accelerate multi-threaded applications by running the serial region on the big core and the parallel region on multiple small cores. In such architectures, all cores implement resource-expensive and application-specific instruction extensions (e.g., SIMD and FP). We argue that instead of implementing such extensions in the big core, the resources must be traded-off to increase the number of small cores. Furthermore, when the big core requires such instruction extensions, we offload execution to the small cores. This design mainly leverages the observation that SIMD/FP operations are more frequently executed inside parallel regions. The proposed AMC provides an additional 1.76x speedup and 12.4% energy savings compared to a traditional AMC of the same area due to enhanced thread-level parallelism (TLP) exploitation.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- big.VLITTLE: On-Demand Data-Parallel Acceleration for Mobile Systems on ChipTuan Ta, Khalid Al-Hawaj, Nick Cebry, Yanghui Ou 等MICRO 2022 · 被引用 10 次
- Occamy: Elastically Sharing a SIMD Co-processor across Multiple CPU CoresZhongcheng Zhang, Yan Ou, Ying Liu, Chenxi Wang 等ASPLOS 2023 · 被引用 4 次
- Dual-Issue Execution of Mixed Integer and Floating-Point Workloads on Energy-Efficient In-Order RISC-V CoresLuca Colagrande, Luca BeniniDAC 2025 · 被引用 2 次
- Asymmetry-aware scalable lockingNian Liu, Jinyu Gu, Dahai Tang, Kenli Li 等PPoPP 2022 · 被引用 7 次
- SHADOW: Simultaneous Multi-Threading Architecture with Asymmetric ThreadsIshita Chaturvedi, Bhargav Reddy Godala, Abiram Gangavaram, Daniel Flyer 等MICRO 2025 · 被引用 1 次
