paper-with-me

홈 › Papers

COMPASS: A Compiler Framework for Resource-Constrained Crossbar-Array Based In-Memory Deep Learning Accelerators

2025-01-12 · Jihoon Park, Jeongin Choe, Dohyun Kim, Jae-Joon Kim

Recently, crossbar array based in-memory accelerators have been gaining interest due to their high throughput and energy efficiency. While software and compiler support for the in-memory accelerators has also been introduced, they are currently limited to the case where all weights are assumed to be on-chip. This limitation becomes apparent with the significantly increasing network sizes compared to the in-memory footprint. Weight replacement schemes are essential to address this issue. We propose COMPASS, a compiler framework for resource-constrained crossbar-based processing-in-memory (PIM) deep neural network (DNN) accelerators. COMPASS is specially targeted for networks that exceed the capacity of PIM crossbar arrays, necessitating access to external memories. We propose an algorithm to determine the optimal partitioning that divides the layers so that each partition can be accelerated on chip. Our scheme takes into account the data dependence between layers, core utilization, and the number of write instructions to minimize latency, memory accesses, and improve energy efficiency. Simulation results demonstrate that COMPASS can accommodate much more networks using a minimal memory footprint, while improving throughput by 1.78X and providing 1.28X savings in energy-delay product (EDP) over baseline partitioning methods.

📄 PDF Abstract BibTeX arXiv:2501.06780

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Crossbar-aware neural network pruning

2018-07-25 · Ling Liang, Lei Deng, Yueling Zeng, Xing Hu 외

Crossbar architecture based devices have been widely adopted in neural network accelerators by taking advantage of the high efficiency on vector-matrix multiplication (VMM) operations. However, in the case of convolution…

Network Pruning

Optimizing Binary and Ternary Neural Network Inference on RRAM Crossbars using CIM-Explorer

2025-05-20 · Rebecca Pelke, José Cubero-Cascante, Nils Bosbach, Niklas Degener 외

Using Resistive Random Access Memory (RRAM) crossbars in Computing-in-Memory (CIM) architectures offers a promising solution to overcome the von Neumann bottleneck. Due to non-idealities like cell variability, RRAM cross…

Quantization

Mixed-Precision Training and Compilation for RRAM-based Computing-in-Memory Accelerators

2026-01-29 · Rebecca Pelke, Joel Klein, Jose Cubero-Cascante, Nils Bosbach 외 arxiv

Computing-in-Memory (CIM) accelerators are a promising solution for accelerating Machine Learning (ML) workloads, as they perform Matrix-Vector Multiplications (MVMs) on crossbar arrays directly in memory. Although the b…

Reinforcement Learning

eIQ Neutron: Redefining Edge-AI Inference with Integrated NPU and Compiler Innovations

2025-09-17 · Lennart Bamberg, Filippo Minnella, Roberto Bosio, Fabrizio Ottati 외 arxiv

Neural Processing Units (NPUs) are key to enabling efficient AI inference in resource-constrained edge environments. While peak tera operations per second (TOPS) is often used to gauge performance, it poorly reflects rea…

CIM-MLC: A Multi-level Compilation Stack for Computing-In-Memory Accelerators

2024-01-23 · Songyun Qu, Shixin Zhao, Bing Li, Yintao He 외

In recent years, various computing-in-memory (CIM) processors have been presented, showing superior performance over traditional architectures. To unleash the potential of various CIM architectures, such as device precis…

Scheduling