paper-with-me

홈 › Papers

Accelerating TinyML Inference on Microcontrollers through Approximate Kernels

2024-09-25 · Giorgos Armeniakos, Georgios Mentzos, Dimitrios Soudris

The rapid growth of microcontroller-based IoT devices has opened up numerous applications, from smart manufacturing to personalized healthcare. Despite the widespread adoption of energy-efficient microcontroller units (MCUs) in the Tiny Machine Learning (TinyML) domain, they still face significant limitations in terms of performance and memory (RAM, Flash). In this work, we combine approximate computing and software kernel design to accelerate the inference of approximate CNN models on MCUs. Our kernel-based approximation framework firstly unpacks the operands of each convolution layer and then conducts an offline calculation to determine the significance of each operand. Subsequently, through a design space exploration, it employs a computation skipping approximation strategy based on the calculated significance. Our evaluation on an STM32-Nucleo board and 2 popular CNNs trained on the CIFAR-10 dataset shows that, compared to state-of-the-art exact inference, our Pareto optimal solutions can feature on average 21% latency reduction with no degradation in Top-1 classification accuracy, while for lower accuracy requirements, the corresponding reduction becomes even more pronounced.

📄 PDF Abstract BibTeX arXiv:2409.16815

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Convolution A convolution is a type of matrix operation, consisting of a kernel, a small matrix of weights, that slides over input data performing element-wise multiplication with the…

Similar Papers 제목 키워드 기반

MinUn: Accurate ML Inference on Microcontrollers

2022-10-29 · Shikhar Jaiswal, Rahul Kiran Kranti Goli, Aayan Kumar, Vivek Seshadri 외

Running machine learning inference on tiny devices, known as TinyML, is an emerging research area. This task requires generating inference code that uses memory frugally, a task that standard ML frameworks are ill-suited…

TinyOL: TinyML with Online-Learning on Microcontrollers

2021-03-15 · Haoyu Ren, Darko Anicic, Thomas Runkler

Tiny machine learning (TinyML) is a fast-growing research area committed to democratizing deep learning for all-pervasive microcontrollers (MCUs). Challenged by the constraints on power, memory, and computation, TinyML h…

Optimizing TinyML: The Impact of Reduced Data Acquisition Rates for Time Series Classification on Microcontrollers

2024-09-17 · Riya Samanta, Bidyut Saha, Soumya K. Ghosh, Ram Babu Roy

Tiny Machine Learning (TinyML) enables efficient, lowcost, and privacy preserving machine learning inference directly on microcontroller units (MCUs) connected to sensors. Optimizing models for these constrained environm…

Privacy PreservingTime SeriesTime Series Classification

MEMA Runtime Framework: Minimizing External Memory Accesses for TinyML on Microcontrollers

2023-04-12 · Andrew Sabot, Vikas Natesh, H. T. Kung, Wei-Te Ting

We present the MEMA framework for the easy and quick derivation of efficient inference runtimes that minimize external memory accesses for matrix multiplication on TinyML systems. The framework accounts for hardware reso…

Heuristic SearchScheduling

MicroNets: Neural Network Architectures for Deploying TinyML Applications on Commodity Microcontrollers

2020-10-21 · Colby Banbury, Chuteng Zhou, Igor Fedorov, Ramon Matas Navarro 외

Executing machine learning workloads locally on resource constrained microcontrollers (MCUs) promises to drastically expand the application space of IoT. However, so-called TinyML presents severe technical challenges, as…

Anomaly DetectionKeyword SpottingNeural Architecture Search