← Search

Ankur Agrawal

3 accepted papers

2020

Ultra-Low Precision 4-bit Training of Deep Neural Networks

NeurIPS 2020oral

In this paper, we propose a number of novel techniques and numerical representation formats that enable, for the very first time, the precision of training systems to be aggressively scaled from 8-bits to 4-bits. To enable this advance, we explore a novel adaptive Gradient Scaling technique (Gradsca…

2019

Accumulation Bit-Width Scaling For Ultra-Low Precision Training Of Deep Networks

ICLR 2019poster

Efforts to reduce the numerical precision of computations in deep learning training have yielded systems that aggressively quantize weights and activations, yet employ wide high-precision accumulators for partial sums in inner-product operations to preserve the quality of convergence. The absence of…

Cited by 43SourcePDFScholar
2015

Deep Learning with Limited Numerical Precision

ICML 2015poster

Training of large-scale deep neural networks is often constrained by the available computational resources. We study the effect of limited precision data representation and computation on neural network training. Within the context of low-precision fixed-point computations, we observe the rounding s…

Cited by 2815SourcePDFScholar