paper-with-me

홈 › Papers

FAST: DNN Training Under Variable Precision Block Floating Point with Stochastic Rounding

2021-10-28 · Sai Qian Zhang, Bradley McDanel, H. T. Kung

Block Floating Point (BFP) can efficiently support quantization for Deep Neural Network (DNN) training by providing a wide dynamic range via a shared exponent across a group of values. In this paper, we propose a Fast First, Accurate Second Training (FAST) system for DNNs, where the weights, activations, and gradients are represented in BFP. FAST supports matrix multiplication with variable precision BFP input operands, enabling incremental increases in DNN precision throughout training. By increasing the BFP precision across both training iterations and DNN layers, FAST can greatly shorten the training time while reducing overall hardware resource usage. Our FAST Multipler-Accumulator (fMAC) supports dot product computations under multiple BFP precisions. We validate our FAST system on multiple DNNs with different datasets, demonstrating a 2-6$\times$ speedup in training on a single-chip platform over prior work based on \textbf{mixed-precision or block} floating point number systems while achieving similar performance in validation accuracy.

📄 PDF Abstract BibTeX arXiv:2110.15456

Code (0)

등록된 구현이 없습니다.

Tasks

Quantization

Similar Papers 제목 키워드 기반

Minimax Estimation of Bandable Precision Matrices

2017-10-19 · NeurIPS 2017 12 · Addison Hu, Sahand Negahban

The inverse covariance matrix provides considerable insight for understanding statistical models in the multivariate setting. In particular, when the distribution over variables is assumed to be multivariate normal, the …

A Forest Mixture Bound for Block-Free Parallel Inference

2018-05-17 · Neal Lawton, Aram Galstyan, Greg Ver Steeg

Coordinate ascent variational inference is an important algorithm for inference in probabilistic models, but it is slow because it updates only a single variable at a time. Block coordinate methods perform inference fast…

Variational Inference

Fast Projected Newton-like Method for Precision Matrix Estimation under Total Positivity

2023-09-21 · NeurIPS 2023 11

We study the problem of estimating precision matrices in Gaussian distributions that are multivariate totally positive of order two ($\mathrm{MTP}_2$). The precision matrix in such a distribution is an M-matrix. This pro…

Block-Wise Dynamic-Precision Neural Network Training Acceleration via Online Quantization Sensitivity Analytics

2022-10-31 · Ruoyang Liu, Chenhan Wei, Yixiong Yang, Wenxun Wang 외

Data quantization is an effective method to accelerate neural network training and reduce power consumption. However, it is challenging to perform low-bit quantized training: the conventional equal-precision quantization…

QuantizationSensitivity

Joint Estimation of Precision Matrices in Heterogeneous Populations

2016-01-02 · Takumi Saegusa, Ali Shojaie

We introduce a general framework for estimation of inverse covariance, or precision, matrices from heterogeneous populations. The proposed framework uses a Laplacian shrinkage penalty to encourage similarity among estima…

Clusteringparameter estimationVariable Selection