ReCG: ReRAM-Accelerated Sparse Conjugate Gradient
Mingjia Fan, Xiaoming Chen, Dechuang Yang, Zhou Jin, Weifeng Liu
摘要
Solving sparse linear systems is crucial in scientific computing. Sparse Conjugate Gradient (CG) is one of the most well-known iterative solvers with high efficiency and low storage requirements. However, the performance of sparse CG solvers implemented on storage-compute separated architectures is greatly limited by the irregular memory access and the large amount of data transmission.
In this paper, we propose a processing-in-memory (PIM) architecture, ReCG, based on the resistive random-access memory (ReRAM) to accelerate sparse CG solvers. The design of ReCG faces three major challenges: (1) how to make complex sparse CG more suitable for acceleration with ReRAM-based architecture, (2) how to map sparse and irregular operations to regular crossbars that are more suitable for dense operations, and (3) how to coordinate the dataflow among hardware units to minimize the impact of the poor write endurance of ReRAMs on CG acceleration. To address these challenges, we (1) classify the kernels of sparse CG by exploring the commonality of operations and design a flexible and dedicated architecture, (2) efficiently implement the sparse and irregular operations by utilizing both content-addressable memory (CAM) and multiply-and-accumulate (MAC) crossbars, and (3) develop a novel scheduling strategy for the dataflow. The experimental results show that ReCG improves the performance by up to three, one and one order of magnitude compared with PETSc on CPU and GPU and CALLIPEPLA on FPGA, respectively, and the energy consumption is reduced by up to two, two and one order of magnitude.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- AmgT: Algebraic Multigrid Solver on Tensor CoresYuechen Lu, Lijie Zeng, Tengcheng Wang, Xu Fu 等SC 2024 · 被引用 17 次
- Mille-feuille: A Tile-Grained Mixed Precision Single-Kernel Conjugate Gradient Solver on GPUsDechuang Yang, Yuxuan Zhao, Yiduo Niu, Weile Jia 等SC 2024 · 被引用 8 次
它引用的顶会 Paper4
- GaaS-X: Graph Analytics Accelerator Supporting Sparse Data Representation using Crossbar ArchitecturesNagadastagiri Challapalle, Sahithi Rampalli, Linghao Song, Nandhini Chandramoorthy 等ISCA 2020 · 被引用 67 次
- PIM-DH: ReRAM-based processing-in-memory architecture for deep hashing accelerationFangxin Liu, Wenbo Zhao, Yongbiao Chen, Zongwu Wang 等DAC 2022 · 被引用 11 次
- FSPA: An FeFET-based Sparse Matrix-Dense Vector Multiplication AcceleratorXiaoyu Zhang, Zerun Li, Rui Liu, Xiaoming Chen 等DAC 2023 · 被引用 10 次
- AmgR: Algebraic Multigrid Accelerated on ReRAMMingjia Fan, Xiaotian Tian, Yintao He, Junxian Li 等DAC 2023 · 被引用 8 次
相关 Paper
- ReSMiPS: A ReRAM-based Sparse Mixed-precision Solver with Fast Matrix Reordering AlgorithmYuyang Fu, Jiancong Li, Jia Chen, Zhiwei Zhou 等DAC 2025
- PIMGCN: A ReRAM-Based PIM Design for Graph Convolutional Network AccelerationTao Yang, Dongyue Li, Yibo Han, Yilong Zhao 等DAC 2021 · 被引用 39 次
- ReSiPE: ReRAM-based Single-Spiking Processing-In-Memory EngineZiru Li, Bonan Yan, Hai Helen LiDAC 2020 · 被引用 18 次
- ReFloat: Low-Cost Floating-Point Processing in ReRAM for Accelerating Iterative Linear SolversLinghao Song, Fan Chen, Hai Li, Yiran ChenSC 2023 · 被引用 8 次
- ReSMA: accelerating approximate string matching using ReRAM-based content addressable memoryHuize Li, Hai Jin, Long Zheng, Yu Huang 等DAC 2022 · 被引用 7 次
