paper-with-me

홈 › Papers

Heterogeneous Integration of In-Memory Analog Computing Architectures with Tensor Processing Units

2023-04-18 · Mohammed E. Elbtity, Brendan Reidy, Md Hasibul Amin, Ramtin Zand

Tensor processing units (TPUs), specialized hardware accelerators for machine learning tasks, have shown significant performance improvements when executing convolutional layers in convolutional neural networks (CNNs). However, they struggle to maintain the same efficiency in fully connected (FC) layers, leading to suboptimal hardware utilization. In-memory analog computing (IMAC) architectures, on the other hand, have demonstrated notable speedup in executing FC layers. This paper introduces a novel, heterogeneous, mixed-signal, and mixed-precision architecture that integrates an IMAC unit with an edge TPU to enhance mobile CNN performance. To leverage the strengths of TPUs for convolutional layers and IMAC circuits for dense layers, we propose a unified learning algorithm that incorporates mixed-precision training techniques to mitigate potential accuracy drops when deploying models on the TPU-IMAC architecture. The simulations demonstrate that the TPU-IMAC configuration achieves up to $2.59\times$ performance improvements, and $88\%$ memory reductions compared to conventional TPU architectures for various CNN models while maintaining comparable accuracy. The TPU-IMAC architecture shows potential for various applications where energy efficiency and high performance are essential, such as edge computing and real-time processing in mobile devices. The unified training algorithm and the integration of IMAC and TPU architectures contribute to the potential impact of this research on the broader machine learning landscape.

📄 PDF Abstract BibTeX arXiv:2304.09258

Code (0)

등록된 구현이 없습니다.

Tasks

Edge-computing

Similar Papers 제목 키워드 기반

A Heterogeneous In-Memory Computing Cluster For Flexible End-to-End Inference of Real-World Deep Neural Networks

2022-01-04 · Angelo Garofalo, Gianmarco Ottavi, Francesco Conti, Geethan Karunaratne 외

Deployment of modern TinyML tasks on small battery-constrained IoT devices requires high computational energy efficiency. Analog In-Memory Computing (IMC) using non-volatile memory (NVM) promises major efficiency improve…

Kernel Approximation using Analog In-Memory Computing

2024-11-05 · Julian Büchel, Giacomo Camposampiero, Athanasios Vasilopoulos, Corey Lammie 외

Kernel functions are vital ingredients of several machine learning algorithms, but often incur significant memory and computational costs. We introduce an approach to kernel approximation in machine learning algorithms s…

Analog, In-memory Compute Architectures for Artificial Intelligence

2023-01-13 · Patrick Bowen, Guy Regev, Nir Regev, Bruno Pedroni 외

This paper presents an analysis of the fundamental limits on energy efficiency in both digital and analog in-memory computing architectures, and compares their performance to single instruction, single data (scalar) mach…

Reliability-Aware Deployment of DNNs on In-Memory Analog Computing Architectures

2022-10-02 · Md Hasibul Amin, Mohammed Elbtity, Ramtin Zand

Conventional in-memory computing (IMC) architectures consist of analog memristive crossbars to accelerate matrix-vector multiplication (MVM), and digital functional units to realize nonlinear vector (NLV) operations in d…

A Microprocessor implemented in 65nm CMOS with Configurable and Bit-scalable Accelerator for Programmable In-memory Computing

2018-11-09 · Hongyang Jia, Yinqi Tang, Hossein Valavi, Jintao Zhang 외

This paper presents a programmable in-memory-computing processor, demonstrated in a 65nm CMOS technology. For data-centric workloads, such as deep neural networks, data movement often dominates when implemented with toda…

CPU