CIM-MLC: A Multi-level Compilation Stack for Computing-In-Memory Accelerators
Songyun Qu, Shixin Zhao, Bing Li, Yintao He, Xuyi Cai, Lei Zhang, Ying Wang
摘要
In recent years, various computing-in-memory (CIM) processors have been presented, showing superior performance over traditional architectures. To unleash the potential of various CIM architectures, such as device precision, crossbar size, and crossbar number, it is necessary to develop compilation tools that are fully aware of the CIM architectural details and implementation diversity. However, due to the lack of architectural support in current popular open-source compiling stacks such as TVM, existing CIM designs either manually deploy networks or build their own compilers, which is time-consuming and labor-intensive. Although some works expose the specific CIM device programming interfaces to compilers, they are often bound to a fixed CIM architecture, lacking the flexibility to support the CIM architectures with different computing granularity. On the other hand, existing compilation works usually consider the scheduling of limited operation types (such as crossbar-bound matrix-vector multiplication). Unlike conventional processors, CIM accelerators are featured by their diverse architecture, circuit, and device, which cannot be simply abstracted by a single level if we seek to fully explore the advantages brought by CIM.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- CINM (Cinnamon): A Compilation Infrastructure for Heterogeneous Compute In-Memory and Compute Near-Memory ParadigmsAsif Ali Khan, Hamid Farzaneh, Karl Friedrich Alexander Friebel, Clément Fournier 等ASPLOS 2024 · 被引用 7 次
- CIMFlow: An Integrated Framework for Systematic Design and Evaluation of Digital CIM ArchitecturesYingjie Qi, Jianlei Yang, Yiou Wang, Yikun Wang 等DAC 2025 · 被引用 2 次
- Be CIM or Be Memory: A Dual-mode-aware DNN Compiler for CIM AcceleratorsShixin Zhao, Yuming Li, Bing Li, Yintao He 等ASPLOS 2025 · 被引用 2 次
它引用的顶会 Paper4
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- FORMS: Fine-grained Polarized ReRAM-based In-situ Computation for Mixed-signal DNN AcceleratorGeng Yuan, Payman Behnam, Zhengang Li, Ali Shafiee 等ISCA 2021 · 被引用 73 次
- TARe: Task-Adaptive in-situ ReRAM Computing for Graph LearningYintao He, Ying Wang, Cheng Liu, Huawei Li 等DAC 2021 · 被引用 12 次
- Processing-in-SRAM acceleration for ultra-low power visual 3D perceptionYuquan He, Songyun Qu, Gangliang Lin, Cheng Liu 等DAC 2022 · 被引用 4 次
相关 Paper
- PIMCOMP: A Universal Compilation Framework for Crossbar-based PIM DNN AcceleratorsXiaotian Sun, Xinyu Wang, Wanqian Li, Lei Wang 等DAC 2023 · 被引用 19 次
- DCC: Data-Centric Compilation of Machine Learning Kernels for Processing-In-Memory ArchitecturesPeiming Yang, Sankeerth Durvasula, Ivan Fernandez, Mohammad Sadrosadati 等ISCA 2026 · 被引用 3 次
- To PIM or not for emerging general purpose processing in DDR memory systemsAlexandar Devic, Siddhartha Balakrishna Rai, Anand Sivasubramaniam, Ameen Akel 等ISCA 2022 · 被引用 51 次
- AutoDCIM: An Automated Digital CIM CompilerJia Chen, Fengbin Tu, Kunming Shao, Fengshi Tian 等DAC 2023 · 被引用 24 次
- ComPASS: A Compatible PIM Protocol Architecture and Scheduling Solution for Processor-PIM CollaborationSeunghyuk Yu, Hyeonu Kim, Kyoungho Jeun, Sunyoung Hwang 等MICRO 2025 · 被引用 4 次
