C4CAM: A Compiler for CAM-based In-memory Accelerators
Hamid Farzaneh, João Paulo Cardoso de Lima, Mengyuan Li, Asif Ali Khan, Xiaobo Sharon Hu, Jerónimo Castrillón
摘要
Machine learning and data analytics applications increasingly suffer from the high latency and energy consumption of conventional von Neumann architectures. Recently, several in-memory and near-memory systems have been proposed to remove this von Neumann bottleneck. Platforms based on contentaddressable memories (CAMs) are particularly interesting due to their efficient support for the search-based operations that form the foundation for many applications, including K-nearest neighbors (KNN), high-dimensional computing (HDC), recommender systems, and one-shot learning among others. Today, these platforms are designed by hand and can only be programmed with low-level code, accessible only to hardware experts. In this paper, we introduce C4CAM, the first compiler framework to quickly explore CAM configurations and to seamlessly generate code from high-level TorchScript code. C4CAM employs a hierarchy of abstractions that progressively lowers programs, allowing code transformations at the most suitable abstraction level. Depending on the type and technology, CAM arrays exhibit varying latencies and power profiles. Our framework allows analyzing the impact of such differences in terms of system-level performance and energy consumption, and thus supports designers in selecting appropriate designs for a given application.
Index Terms-Content addressable memories (CAM), compute in memory (CIM), TCAM, MLIR
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- CINM (Cinnamon): A Compilation Infrastructure for Heterogeneous Compute In-Memory and Compute Near-Memory ParadigmsAsif Ali Khan, Hamid Farzaneh, Karl Friedrich Alexander Friebel, Clément Fournier 等ASPLOS 2024 · 被引用 7 次
- Be CIM or Be Memory: A Dual-mode-aware DNN Compiler for CIM AcceleratorsShixin Zhao, Yuming Li, Bing Li, Yintao He 等ASPLOS 2025 · 被引用 2 次
它引用的顶会 Paper3
- EDAM: edit distance tolerant approximate matching content addressable memoryRobert Hanhan, Esteban Garzón, Zuher Jahshan, Adam Teman 等ISCA 2022 · 被引用 37 次
- iMARS: an in-memory-computing architecture for recommendation systemsMengyuan Li, Ann Franchesca Laguna, Dayane Reis, Xunzhao Yin 等DAC 2022 · 被引用 14 次
- CINM (Cinnamon): A Compilation Infrastructure for Heterogeneous Compute In-Memory and Compute Near-Memory ParadigmsAsif Ali Khan, Hamid Farzaneh, Karl Friedrich Alexander Friebel, Clément Fournier 等ASPLOS 2024 · 被引用 7 次
相关 Paper
- CAP: A General Purpose Computation-in-memory with Content Addressable Processing ParadigmZhiheng Yue, Shaojun Wei, Yang Hu, Shouyi YinDAC 2024 · 被引用 1 次
- Efficient Weight Mapping and Resource Scheduling on Crossbar-based Multi-core CIM SystemsHanjie Liu, Sifan Sun, Aifei Zhang, Haiyan Qin 等DAC 2025 · 被引用 1 次
- CAMPER: Exploring the Potential of Content Addressable Memory for 3D Point Cloud Efficient Range SearchJiapei Zheng, Lizhou Wu, Yutong Su, Jingyi Wang 等DAC 2024 · 被引用 1 次
- CIM-MLC: A Multi-level Compilation Stack for Computing-In-Memory AcceleratorsSongyun Qu, Shixin Zhao, Bing Li, Yintao He 等ASPLOS 2024 · 被引用 11 次
- OptiPIM: Optimizing Processing-in-Memory Acceleration Using Integer Linear ProgrammingJiantao Liu, Minxuan Zhou, Yue Pan, Chien-Yi Yang 等ISCA 2025 · 被引用 6 次
