Software-Defined Vector Processing on Manycore Fabrics
Philip Bedoukian, Neil Adit, Edwin Peguero, Adrian Sampson
摘要
We describe a tiled architecture that can fluidly transition between manycore (MIMD) and vector (SIMD) execution. The hardware provides a software-defined vector programming model that lets applications aggregate groups of manycore tiles into logical vector engines. In manycore mode, the machine behaves as a standard parallel processor. In vector mode, groups of tiles repurpose their functional units as vector execution lanes and scratchpads as vector memory banks. The key mechanism is an instruction forwarding network: a single tile fetches instructions and sends them to other trailing cores. Most cores disable their frontends and instruction caches, so vector groups amortize the intrinsic hardware costs of von Neumann control. Vector groups also use a decoupled access/execute scheme to centralize their memory requests and issue coalesced, wide loads.
We augment an existing RISC-V manycore design with a minimal hardware extension to implement software-defined vectors. Cyclelevel simulation results show that software-defined vectors improve performance by an average of 1.7× over standard MIMD execution while saving 22% of the energy. Compared to a similarly configured GPU, the architecture improves performance by 1.9×.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper1
相关 Paper
- Scalable, Programmable and Dense: The HammerBlade Open-Source RISC-V ManycoreDai Cheol Jung, Max Ruttenberg, Paul Gao, Scott Davidson 等ISCA 2024 · 被引用 11 次
- Efficient and scalable core multiplexing with M³vNils Asmussen, Sebastian Haas, Carsten Weinhold, Till Miemietz 等ASPLOS 2022 · 被引用 11 次
- To PIM or not for emerging general purpose processing in DDR memory systemsAlexandar Devic, Siddhartha Balakrishna Rai, Anand Sivasubramaniam, Ameen Akel 等ISCA 2022 · 被引用 51 次
- big.VLITTLE: On-Demand Data-Parallel Acceleration for Mobile Systems on ChipTuan Ta, Khalid Al-Hawaj, Nick Cebry, Yanghui Ou 等MICRO 2022 · 被引用 10 次
- SARIS: Accelerating Stencil Computations on Energy-Efficient RISC-V Compute Clusters with Indirect Stream RegistersPaul Scheffler, Luca Colagrande, Luca BeniniDAC 2024 · 被引用 3 次
