StreamPIM: Streaming Matrix Computation in Racetrack Memory
Yuda An, Yunxiao Tang, Shushu Yi, Li Peng, Xiurui Pan, Guangyu Sun, Zhaochu Luo, Qiao Li, Jie Zhang
摘要
Racetrack memory (RM) techniques have become promising solutions to resolve the memory wall issue as they increase memory density, reduce energy consumption and are capable of building processing-in-memory (PIM) architectures. RM can place arithmetic logic units in or near its memory arrays to process tasks offloaded by the host. While there already exist multiple studies of processing in RM, these solutions, unfortunately, suffer from data transfer overheads imposed by the loose coupling of the memory core and the computation units. To address this issue, we propose StreamPIM, a new processing-in-RM architecture, which tightly couples the memory core and the computation units. Specifically, StreamPIM directly constructs a matrix processor from domain-wall nanowires without the usage of CMOS-based computation units. It also designs a domainwall nanowire-based bus, which can eliminate electromagnetic conversion. StreamPIM further optimizes the performance by leveraging RM internal parallelism. Our evaluation results show that StreamPIM achieves 39.1 × higher performance and saves 58.4 × energy consumption, compared with the traditional computing platform.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- CORUSCANT: Fast Efficient Processing-in-Racetrack MemoriesSébastien Ollivier, Stephen Longofono, Prayash Dutta, Jingtong Hu 等MICRO 2022 · 被引用 15 次
- UM-PIM: DRAM-based PIM with Uniform & Shared Memory SpaceYilong Zhao, Mingyu Gao, Fangxin Liu, Yiwei Hu 等ISCA 2024 · 被引用 26 次
- AmgR: Algebraic Multigrid Accelerated on ReRAMMingjia Fan, Xiaotian Tian, Yintao He, Junxian Li 等DAC 2023 · 被引用 8 次
- Max-PIM: Fast and Efficient Max/Min Searching in DRAMFan Zhang, Shaahin Angizi, Deliang FanDAC 2021 · 被引用 10 次
- PIPF-DRAM: processing in precharge-free DRAMNezam Rohbani, Mohammad Arman Soleimani, Hamid Sarbazi-AzadDAC 2022 · 被引用 7 次
