Fetching the paper…
Reading the bibliography…
To make deep neural networks feasible in resource-constrained environments (such as mobile devices), it is beneficial to quantize models by using low-precision weights.
Building a large annotated corpus of english: The penn treebank
M. P. Marcus, M. A. Marcinkiewicz, and B. Santorini · 1993
Earlier work this paper cites.
Regression shrinkage and selection via the lasso
R. Tibshirani · 1996
Earlier work this paper cites.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Dual averaging methods for regularized stochastic learning and online optimization
L. Xiao · 2010
Earlier work this paper cites.
Adam: A method for stochastic optimization
D. P. Kingma and J. Ba · 2014
Earlier work this paper cites.
Proximal algorithms
N. Parikh and S. Boyd · 2014
Earlier work this paper cites.
BinaryConnect: Training deep neural networks with binary weights during propagations
M. Courbariaux, Y. Bengio, and J.-P. David · 2015
Earlier work this paper cites.
S. Han, H. Mao, and W. J. Dally · 2015
Earlier work this paper cites.
Deep learning , volume 1
I. Goodfellow, Y. Bengio, A. Courville, and Y. Bengio · 2016
Earlier work this paper cites.
EIE: Efficient inference engine on compressed deep neural network
S. Han, X. Liu, H. Mao, J. Pu, A. Pedram, M. A. Horowitz, and W. J. Dally · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Cited alongside, same era.
F. Li and B. Liu · 2016
Cited alongside, same era.
Xnor-net: Imagenet classification using binary convolutional neural networks
M. Rastegari, V. Ordonez, J. Redmon, and A. Farhadi · 2016
Cited alongside, same era.
Quantization and training of low bit-width convolutional neural networks for object detection
P. Yin, S. Zhang, Y. Qi, and J. Xin · 2016
Cited alongside, same era.
Dorefa-net: Training low bitwidth convolutional neural networks with low bitwidth gradients
S. Zhou, Y. Wu, Z. Ni, X. Zhou, H. Wen, and Y. Zou · 2016
Cited alongside, same era.
M. A. Carreira-Perpinán and Y. Idelbayev · 2017
Later among the works it cites.
Loss-aware binarization of deep networks
L. Hou, Q. Yao, and J. T. Kwok · 2017
Later among the works it cites.
Quantized neural networks: Training neural networks with low precision weights and activations
I. Hubara, M. Courbariaux, D. Soudry, R. El-Yaniv, and Y. Bengio · 2017
Later among the works it cites.
Training quantized nets: A deeper understanding
H. Li, S. De, Z. Xu, C. Studer, H. Samet, and T. Goldstein · 2017
Later among the works it cites.
On the universal approximability of quantized relu neural networks
Y. Ding, J. Liu, and Y. Shi · 2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
C. Zhu, S. Han, H. Mao, and W. J. Dally · 2016
Cited alongside, same era.
The high-dimensional geometry of binary neural networks
A. G. Anderson and C. P. Berg · 2017
Cited alongside, same era.
M. Arjovsky, S. Chintala, and L. Bottou · 2017
Cited alongside, same era.
M. A. Carreira-Perpinán · 2017
Cited alongside, same era.
Loss-aware weight quantization of deep networks
L. Hou and J. T. Kwok · 2018
Closest in time.
Adversarial probabilistic regularization
J. Sun and X. Sun · 2018
Closest in time.
Alternating multi-bit quantization for recurrent neural networks
C. Xu, J. Yao, Z. Lin, W. Ou, Y. Cao, Z. Wang, and H. Zha · 2018
Closest in time.
Binaryrelax: A relaxation approach for training deep neural networks with quantized weights
P. Yin, S. Zhang, J. Lyu, S. Osher, Y. Qi, and J. Xin · 2018
Closest in time.