paper-with-me

Papers

Bias-Scalable Near-Memory CMOS Analog Processor for Machine Learning

2022-02-10 · Pratik Kumar, Ankita Nandi, Shantanu Chakrabartty, Chetan Singh Thakur

Bias-scalable analog computing is attractive for implementing machine learning (ML) processors with distinct power-performance specifications. For instance, ML implementations for server workloads are focused on higher computational throughput for faster training, whereas ML implementations for edge devices are focused on energy-efficient inference. In this paper, we demonstrate the implementation of bias-scalable approximate analog computing circuits using the generalization of the margin-propagation principle called shape-based analog computing (S-AC). The resulting S-AC core integrates several near-memory compute elements, which include: (a) non-linear activation functions; (b) inner-product compute circuits; and (c) a mixed-signal compressive memory, all of which can be scaled for performance or power while preserving its functionality. Using measured results from prototypes fabricated in a 180nm CMOS process, we demonstrate that the performance of computing modules remains robust to transistor biasing and variations in temperature. In this paper, we also demonstrate the effect of bias-scalability and computational accuracy on a simple ML regression task.

📄 PDF Abstract BibTeX arXiv:2202.05022

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

High-Throughput In-Memory Computing for Binary Deep Neural Networks with Monolithically Integrated RRAM and 90nm CMOS

2019-09-16 · Shihui Yin, Xiaoyu Sun, Shimeng Yu, Jae-sun Seo

Deep learning hardware designs have been bottlenecked by conventional memories such as SRAM due to density, leakage and parallel computing challenges. Resistive devices can address the density and volatility issues, but …

Edge-computing

Process, Bias and Temperature Scalable CMOS Analog Computing Circuits for Machine Learning

2022-05-11 · Pratik Kumar, Ankita Nandi, Shantanu Chakrabartty, Chetan Singh Thakur

Analog computing is attractive compared to digital computing due to its potential for achieving higher computational density and higher energy efficiency. However, unlike digital circuits, conventional analog computing c…

BIG-bench Machine Learning

A Microprocessor implemented in 65nm CMOS with Configurable and Bit-scalable Accelerator for Programmable In-memory Computing

2018-11-09 · Hongyang Jia, Yinqi Tang, Hossein Valavi, Jintao Zhang 외

This paper presents a programmable in-memory-computing processor, demonstrated in a 65nm CMOS technology. For data-centric workloads, such as deep neural networks, data movement often dominates when implemented with toda…

CPU

Analog CMOS-based Resistive Processing Unit for Deep Neural Network Training

2017-06-20 · Seyoung Kim, Tayfun Gokmen, Hyung-Min Lee, Wilfried E. Haensch

Recently we have shown that an architecture based on resistive processing unit (RPU) devices has potential to achieve significant acceleration in deep neural network (DNN) training compared to today's software-based DNN …

CPUGPU

An Analog Neural Network Computing Engine using CMOS-Compatible Charge-Trap-Transistor (CTT)

2017-09-19 · Yuan Du, Li Du, Xuefeng Gu, Jieqiong Du 외

An analog neural network computing engine based on CMOS-compatible charge-trap transistor (CTT) is proposed in this paper. CTT devices are used as analog multipliers. Compared to digital multipliers, CTT-based analog mul…