paper-with-me

Papers

Energy-efficiency Limits on Training AI Systems using Learning-in-Memory

2024-02-21 · Zihao Chen, Johannes Leugering, Gert Cauwenberghs, Shantanu Chakrabartty

Learning-in-memory (LIM) is a recently proposed paradigm to overcome fundamental memory bottlenecks in training machine learning systems. While compute-in-memory (CIM) approaches can address the so-called memory-wall (i.e. energy dissipated due to repeated memory read access) they are agnostic to the energy dissipated due to repeated memory writes at the precision required for training (the update-wall), and they don't account for the energy dissipated when transferring information between short-term and long-term memories (the consolidation-wall). The LIM paradigm proposes that these bottlenecks, too, can be overcome if the energy barrier of physical memories is adaptively modulated such that the dynamics of memory updates and consolidation match the Lyapunov dynamics of gradient-descent training of an AI model. In this paper, we derive new theoretical lower bounds on energy dissipation when training AI systems using different LIM approaches. The analysis presented here is model-agnostic and highlights the trade-off between energy efficiency and the speed of training. The resulting non-equilibrium energy-efficiency bounds have a similar flavor as that of Landauer's energy-dissipation bounds. We also extend these limits by taking into account the number of floating-point operations (FLOPs) used for training, the size of the AI model, and the precision of the training parameters. Our projections suggest that the energy-dissipation lower-bound to train a brain scale AI system (comprising of $10^{15}$ parameters) using LIM is $10^8 \sim 10^9$ Joules, which is on the same magnitude the Landauer's adiabatic lower-bound and $6$ to $7$ orders of magnitude lower than the projections obtained using state-of-the-art AI accelerator hardware lower-bounds.

📄 PDF Abstract BibTeX arXiv:2402.14878

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

Analog, In-memory Compute Architectures for Artificial Intelligence

2023-01-13 · Patrick Bowen, Guy Regev, Nir Regev, Bruno Pedroni 외

This paper presents an analysis of the fundamental limits on energy efficiency in both digital and analog in-memory computing architectures, and compares their performance to single instruction, single data (scalar) mach…

OPIMA: Optical Processing-In-Memory for Convolutional Neural Network Acceleration

2024-07-11 · Febin Sunny, Amin Shafiee, Abhishek Balasubramaniam, Mahdi Nikdast 외

Recent advances in machine learning (ML) have spotlighted the pressing need for computing architectures that bridge the gap between memory bandwidth and processing power. The advent of deep neural networks has pushed tra…

Solving Boltzmann Optimization Problems with Deep Learning

2024-01-30 · Fiona Knoll, John T. Daly, Jess J. Meyer

Decades of exponential scaling in high performance computing (HPC) efficiency is coming to an end. Transistor based logic in complementary metal-oxide semiconductor (CMOS) technology is approaching physical limits beyond…

Deep Learning

An Adaptive Synaptic Array using Fowler-Nordheim Dynamic Analog Memory

2021-04-13 · Darshit Mehta, Kenji Aono, Shantanu Chakrabartty

In this paper we present a synaptic array that uses dynamical states to implement an analog memory for energy-efficient training of machine learning (ML) systems. Each of the analog memory elements is a micro-dynamical s…

Training DNN IoT Applications for Deployment On Analog NVM Crossbars

2019-10-30 · Fernando García-Redondo, Shidhartha Das, Glen Rosendale

A trend towards energy-efficiency, security and privacy has led to a recent focus on deploying DNNs on microcontrollers. However, limits on compute and memory resources restrict the size and the complexity of the ML mode…

Quantization