Fetching the paper…
Reading the bibliography…
We consider the problem of deep neural net compression by quantization: given a large, reference net, we want to quantize its real-valued weights using a codebook with $K$ entries so that the training loss of the quantized net is minimal.
Computers and Intractability: A Guide to the Theory of NP-Completeness
M. R. Garey and D. S. Johnson · 1979
Earlier work this paper cites.
A method of solving a convex programming problem with convergence rate 𝒪 ( 1 / k 2 ) \mathcal{O}(1/k^{2})
Y. Nesterov · 1983
Earlier work this paper cites.
Weight discretization paradigm for optical neural networks
E. Fiesler, A. Choudry, and H. J. Caulfield · 1990
Earlier work this paper cites.
Vector Quantization and Signal Compression
A. Gersho and R. M. Gray · 1992
Earlier work this paper cites.
Simplifying neural networks by soft weight-sharing
S. J. Nowlan and G. E. Hinton · 1992
Earlier work this paper cites.
Fast neural networks without multipliers
M. Marchesi, G. Orlandi, F. Piazza, and A. Uncini · 1993
Earlier work this paper cites.
Multilayer feedforward neural networks with single powers-of-two weights
C. Z. Tang and H. K. Kwan · 1993
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner · 1998
Earlier work this paper cites.
Numerical Optimization
J. Nocedal and S. J. Wright · 2006
Earlier work this paper cites.
k-means++
D. Arthur and S. Vassilvitskii · 2007
Earlier work this paper cites.
Rectified linear units improve restricted Boltzmann machines
V. Nair and G. E. Hinton · 2010
Earlier work this paper cites.
The K K -modes algorithm for clustering
M. Á. Carreira-Perpiñán and W. Wang · 2013
Cited alongside, same era.
Predicting parameters in deep learning
M. Denil, B. Shakibi, L. Dinh, M. Ranzato, and N. de Freitas · 2013
Cited alongside, same era.
Fixed-point feedforward deep neural network design using weights + 1 +1 , 0 0 , and − 1 -1
K. Hwang and W. Sung · 2014
Cited alongside, same era.
Caffe: Convolutional architecture for fast feature embedding
Y. Jia, E. Shelhamer, J. Donahue, S. Karayev, J. Long, R. Girshick, S. Guadarrama, and T. Darrell · 2014
Cited alongside, same era.
Dropout: A simple way to prevent neural networks from overfitting
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov · 2014
Cited alongside, same era.
Learning both weights and connections for efficient neural network
S. Han, J. Pool, J. Tran, and W. Dally · 2015
Later among the works it cites.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2015
Later among the works it cites.
Quantized neural networks: Training neural networks with low precision weights and activations
I. Hubara, M. Courbariaux, D. Soudry, R. El-Yaniv, and Y. Bengio · 2016
Later among the works it cites.
F. Li, B. Zhang, and B. Liu · 2016
Later among the works it cites.
XNOR-net: ImageNet classification using binary convolutional neural networks
M. Rastegari, V. Ordonez, J. Redmon, and A. Farhadi · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
W. Wang and M. Á. Carreira-Perpiñán · 2014
Cited alongside, same era.
BinaryConnect: Training deep neural networks with binary weights during propagations
M. Courbariaux, Y. Bengio, and J.-P. David · 2015
Cited alongside, same era.
Lasagne: First release, Aug. 2015
S. Dieleman, J. Schlüter, C. Raffel, E. Olson, S. K. Sœnderby, D. Nouri, D. Maturana, M. Thoma, E. Battenberg, J. Kelly, J. D. Fauw, M. Heilman, D. M. de Almeida, B. McFee, H. Weideman, G. Takács, P. de Rivaz, J. Crall, G. Sanders, K. Rasul, C. Liu, G. French, and J. Degrave · 2015
Cited alongside, same era.
Compressing deep convolutional networks using vector quantization
Y. Gong, L. Liu, M. Yang, and L. Bourdev · 2015
Cited alongside, same era.
Deep learning with limited numerical precision
S. Gupta, A. Agrawal, K. Gopalakrishnan, and P. Narayanan · 2015
Cited alongside, same era.
Theano Development Team · 2016
Later among the works it cites.
Dorefa-net: Training low bitwidth convolutional neural networks with low bitwidth gradients
S. Zhou, Z. Ni, X. Zhou, H. Wen, Y. Wu, and Y. Zou · 2016
Later among the works it cites.
M. Á. Carreira-Perpiñán · 2017
Closest in time.
Soft weight-sharing for neural network compression
K. Ullrich, E. Meeds, and M. Welling · 2017
Closest in time.
Trained ternary quantization
C. Zhu, S. Han, H. Mao, and W. J. Dally · 2017
Closest in time.