An Analog Preconditioner for Solving Linear Systems
Ben Feinberg, Ryan Wong, T. Patrick Xiao, Christopher H. Bennett, Jacob N. Rohan, Erik G. Boman, Matthew J. Marinella, Sapan Agarwal, Engin Ipek
摘要
Over the past decade as Moore's Law has slowed, the need for new forms of computation that can provide sustainable performance improvements has risen. A new method, called in situ computing, has shown great potential to accelerate matrix vector multiplication (MVM), an important kernel for a diverse range of applications from neural networks to scientific computing. Existing in situ accelerators for scientific computing, however, have a significant limitation: these accelerators provide no acceleration for preconditioning-a key bottleneck in linear solvers and in scientific computing workflows.
This paper enables in situ acceleration for state-of-the-art linear solvers by demonstrating how to use a new in situ matrix inversion accelerator for analog preconditioning. As existing techniques that enable high precision and scalability for in situ MVM are inapplicable to in situ matrix inversion, new techniques to compensate for circuit non-idealities are proposed. Additionally, a new approach to bit slicing that enables splitting operands across multiple devices without external digital logic is proposed. For scalability, this paper demonstrates how in situ matrix inversion kernels can work in tandem with existing domain decomposition techniques to accelerate the solutions of arbitrarily large linear systems. The analog kernel can be directly integrated into existing preconditioning workflows, leveraging several well-optimized numerical linear algebra tools to improve the behavior of the circuit. The result is an analog preconditioner that is more effective (up to 50% fewer iterations) than the widely used incomplete LU factorization preconditioner, ILU(0), while also reducing the energy and execution time of each approximate solve operation by 1025x and 105x respectively.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- AmgR: Algebraic Multigrid Accelerated on ReRAMMingjia Fan, Xiaotian Tian, Yintao He, Junxian Li 等DAC 2023 · 被引用 8 次
- ReFloat: Low-Cost Floating-Point Processing in ReRAM for Accelerating Iterative Linear SolversLinghao Song, Fan Chen, Hai Li, Yiran ChenSC 2023 · 被引用 8 次
- DARTH-PUM: A Hybrid Processing-Using-Memory ArchitectureRyan Wong, Ben Feinberg, Saugata GhoseASPLOS 2026 · 被引用 1 次
相关 Paper
- PHANES: ReRAM-based photonic accelerator for deep neural networksYinyi Liu, Jiaqi Liu, Yuxiang Fu, Shixi Chen 等DAC 2022 · 被引用 4 次
- A GPU-based multilevel additive schwarz preconditioner for cloth and deformable body simulationBotao Wu, Zhendong Wang, Huamin WangSIGGRAPH 2022 · 被引用 48 次
- On the Intrinsic Robustness of NVM Crossbars Against Adversarial AttacksDeboleena Roy, Indranil Chakraborty, Timur Ibrayev, Kaushik RoyDAC 2021 · 被引用 16 次
- A Diagonal Block Memory-Aware Polynomial Preconditioner for Linear and Eigenvalue SolversXiaojian Yang, Yuhui Ni, Fan Yuan, Shengguo Li 等PPoPP 2026 · 被引用 1 次
- Algorithm/Hardware Co-Design for In-Memory Neural Network Computing with Minimal Peripheral Circuit OverheadHyungjun Kim, Yulhwa Kim, Sungju Ryu, Jae-Joon KimDAC 2020 · 被引用 17 次
