paper-with-me

홈 › Papers

Deep Learning Inference on Embedded Devices: Fixed-Point vs Posit

2018-05-22 · Seyed H. F. Langroudi, Tej Pandit, Dhireesha Kudithipudi

Performing the inference step of deep learning in resource constrained environments, such as embedded devices, is challenging. Success requires optimization at both software and hardware levels. Low precision arithmetic and specifically low precision fixed-point number systems have become the standard for performing deep learning inference. However, representing non-uniform data and distributed parameters (e.g. weights) by using uniformly distributed fixed-point values is still a major drawback when using this number system. Recently, the posit number system was proposed, which represents numbers in a non-uniform manner. Therefore, in this paper we are motivated to explore using the posit number system to represent the weights of Deep Convolutional Neural Networks. However, we do not apply any quantization techniques and hence the network weights do not require re-training. The results of this exploration show that using the posit number system outperformed the fixed point number system in terms of accuracy and memory utilization.

📄 PDF Abstract BibTeX arXiv:1805.08624

Code (0)

등록된 구현이 없습니다.

Tasks

Deep LearningQuantization

Similar Papers 제목 키워드 기반

IFQ-Net: Integrated Fixed-point Quantization Networks for Embedded Vision

2019-11-19 · Hongxing Gao, Wei Tao, Dongchao Wen, Tse-Wei Chen 외

Deploying deep models on embedded devices has been a challenging problem since the great success of deep learning based networks. Fixed-point networks, which represent their data with low bits fixed-point and thus give r…

Face DetectionImage ClassificationQuantization

Deep Convolutional Neural Network Inference with Floating-point Weights and Fixed-point Activations

2017-03-08 · Liangzhen Lai, Naveen Suda, Vikas Chandra

Deep convolutional neural network (CNN) inference requires significant amount of memory and computation, which limits its deployment on embedded devices. To alleviate these problems to some extent, prior research utilize…

Trainable Fixed-Point Quantization for Deep Learning Acceleration on FPGAs

2024-01-31 · Dingyi Dai, Yichi Zhang, Jiahao Zhang, Zhanqiu Hu 외

Quantization is a crucial technique for deploying deep learning models on resource-constrained devices, such as embedded FPGAs. Prior efforts mostly focus on quantizing matrix multiplications, leaving other layers like B…

Deep LearningQuantization

Cheetah: Mixed Low-Precision Hardware & Software Co-Design Framework for DNNs on the Edge

2019-08-06 · Hamed F. Langroudi, Zachariah Carmichael, David Pastuch, Dhireesha Kudithipudi

Low-precision DNNs have been extensively explored in order to reduce the size of DNN models for edge devices. Recently, the posit numerical format has shown promise for DNN data representation and compute with ultra-low …

Quantization

Quantization and Deployment of Deep Neural Networks on Microcontrollers

2021-05-27 · Pierre-Emmanuel Novac, Ghouthi Boukli Hacene, Alain Pegatoquet, Benoît Miramond 외

Embedding Artificial Intelligence onto low-power devices is a challenging task that has been partly overcome with recent advances in machine learning and hardware design. Presently, deep neural networks can be deployed o…

Activity RecognitionHuman Activity Recognitionobject-detectionObject Detection+3