paper-with-me

Papers

Neural-PIM: Efficient Processing-In-Memory with Neural Approximation of Peripherals

2022-01-30 · Weidong Cao, Yilong Zhao, Adith Boloor, Yinhe Han, Xuan Zhang, Li Jiang

Processing-in-memory (PIM) architectures have demonstrated great potential in accelerating numerous deep learning tasks. Particularly, resistive random-access memory (RRAM) devices provide a promising hardware substrate to build PIM accelerators due to their abilities to realize efficient in-situ vector-matrix multiplications (VMMs). However, existing PIM accelerators suffer from frequent and energy-intensive analog-to-digital (A/D) conversions, severely limiting their performance. This paper presents a new PIM architecture to efficiently accelerate deep learning tasks by minimizing the required A/D conversions with analog accumulation and neural approximated peripheral circuits. We first characterize the different dataflows employed by existing PIM accelerators, based on which a new dataflow is proposed to remarkably reduce the required A/D conversions for VMMs by extending shift and add (S+A) operations into the analog domain before the final quantizations. We then leverage a neural approximation method to design both analog accumulation circuits (S+A) and quantization circuits (ADCs) with RRAM crossbar arrays in a highly-efficient manner. Finally, we apply them to build an RRAM-based PIM accelerator (i.e., \textbf{Neural-PIM}) upon the proposed analog dataflow and evaluate its system-level performance. Evaluations on different benchmarks demonstrate that Neural-PIM can improve energy efficiency by 5.36x (1.73x) and speed up throughput by 3.43x (1.59x) without losing accuracy, compared to the state-of-the-art RRAM-based PIM accelerators, i.e., ISAAC (CASCADE).

📄 PDF Abstract BibTeX arXiv:2201.12861

Code (0)

등록된 구현이 없습니다.

Tasks

Quantization

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

PIM-DRAM: Accelerating Machine Learning Workloads using Processing in Commodity DRAM

2021-05-08 · Sourjya Roy, Mustafa Ali, Anand Raghunathan

Deep Neural Networks (DNNs) have transformed the field of machine learning and are widely deployed in many applications involving image, video, speech and natural language processing. The increasing compute demands of DN…

BIG-bench Machine LearningGPUMedical Diagnosis

Containing Analog Data Deluge at Edge through Frequency-Domain Compression in Collaborative Compute-in-Memory Networks

2023-09-20 · Nastaran Darabi, Amit R. Trivedi

Edge computing is a promising solution for handling high-dimensional, multispectral analog data from sensors and IoT devices for applications such as autonomous drones. However, edge devices' limited storage and computin…

Edge-computing

Openmv: A Python powered, extensible machine vision camera

2017-11-01 · Ibrahim Abdelkader, Yasser El-Sonbaty, Mohamed El-Habrouk

Advances in semiconductor manufacturing processes and large scale integration keep pushing demanding applications further away from centralized processing, and closer to the edges of the network (i.e. Edge Computing). It…

Edge-computing

Project Tracyn: Generative Artificial Intelligence based Peripherals Trace Synthesizer

2024-11-10 · Zhibai Huang, Yihan Shen, Yongchen Xie, Zhixiang Wei 외

Peripheral Component Interconnect Express (PCIe) is the de facto interconnect standard for high-speed peripherals and CPUs. Prototyping and optimizing PCIe devices for emerging scenarios is an ongoing challenge. Since Tr…

CPU

PiEEG kit -- bioscience Lab in home for your Brain and Body

2025-03-05 · Ildar Rakhmatulin

PiEEG kit is a multifunctional, compact, and mobile device that allows measure EEG, EMG, EOG, and EKG signals. The PiEEG Box incorporates the Raspberry Pi-based PiEEG shield, an EEG electrode cap, a display screen, addit…

EEG