Fetching the paper…
Reading the bibliography…
Quantized deep neural networks (QDNNs) are attractive due to their much lower memory storage and faster inference speed than their regular full precision counterparts.
Cornell Aeronautical Laboratory (1957)
Rosenblatt, F.: The perceptron, a perceiving and recognizing automaton Project Para · 1957
Earlier work this paper cites.
Spartan Book (1962)
Rosenblatt, F.: Principles of neurodynamics · 1962
Earlier work this paper cites.
IEEE Trans. Info. Theory 28
Lloyd, S.: Least squares quantization in pcm · 1982
Earlier work this paper cites.
Proceedings of the IEEE 78
Widrow, B., Lehr, M.A.: 30 years of adaptive neural networks: perceptron, madaline, and backpropagation · 1990
Earlier work this paper cites.
SIAM Journal on Optimization 2
Gilbert, J.C., Nocedal, J.: Global convergence properties of conjugate gradient methods for optimization · 1992
Earlier work this paper cites.
Athena scientific Belmont (1999)
Bertsekas, D.P.: Nonlinear programming · 1999
Earlier work this paper cites.
Machine learning 37
Freund, Y., Schapire, R.E.: Large margin classification using the perceptron algorithm · 1999
Earlier work this paper cites.
In: IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp. 248–255 (2009)
Deng, J., Dong, W., Socher, R., Li, L., Li, K., Li, F.: Imagenet: A large-scale hierarchical image database · 2009
Earlier work this paper cites.
Tech Report (2009)
Krizhevsky, A.: Learning multiple layers of features from tiny images · 2009
Earlier work this paper cites.
Coursera, video lectures (2012)
Hinton, G.: Neural networks for machine learning, coursera · 2012
Earlier work this paper cites.
In: Advances in Neural Information Processing Systems (NIPS), pp. 1097–1105 (2012)
Krizhevsky, A., Sutskever, I., Hinton, G.: Imagenet classification with deep convolutional neural networks · 2012
Earlier work this paper cites.
arXiv preprint arXiv:1308.3432 (2013)
Bengio, Y., Léonard, N., Courville, A.: Estimating or propagating gradients through stochastic neurons for conditional computation · 2013
Earlier work this paper cites.
In: Advances in Neural Information Processing Systems (NIPS), p. 3123–3131 (2015)
Courbariaux, M., Bengio, Y., David, J.: Binaryconnect: Training deep neural networks with binary weights during propagations · 2015
Earlier work this paper cites.
arXiv preprint arXiv:1512.03385 (2015)
He, K., Zhang, X., Ren, S., Sun, J.: Deep residual learning for image recognition · 2015
Cited alongside, same era.
In: IEEE International Conference on Computer Vision (ICCV) (2015)
He, K., Zhang, X., Ren, S., Sun, J.: Delving deep into rectifiers: Surpassing human-level performance on imagenet classification · 2015
Cited alongside, same era.
arXiv preprint arXiv:1502.03167 (2015)
Ioffe, S., Szegedy, C.: Normalization: Accelerating deep network training by reducing internal covariate shift · 2015
Cited alongside, same era.
arXiv preprint arXiv:1409.1556 (2015)
Simonyan, K., Zisserman, A.: Very deep convolutional networks for large-scale image recognition · 2015
Cited alongside, same era.
Pure and Applied Functional Analysis 1
Combettes, P.L., Pesquet, J.C.: Stochastic approximations and perturbations in forward-backward splitting for monotone operators · 2016
In: NIPS, pp. 5813–5823 (2017)
Li, H., De, S., Xu, Z., Studer, C., Samet, H., Goldstein, T.: Training quantized nets: A deeper understanding · 2017
Later among the works it cites.
In: Advances in Neural Information Processing Systems, pp. 597–607 (2017)
Li, Y., Yuan, Y.: Convergence analysis of two-layer neural networks with relu activation · 2017
Later among the works it cites.
In: IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pp. 5456–5464 (2017)
Park, E., Ahn, J., Yoo, S.: Weighted-entropy-based quantization for deep neural networks · 2017
Later among the works it cites.
Tech Report (2017)
Paszke, A., Gross, S., Chintala, S., Chanan, G., Yang, E., DeVito, Z., Lin, Z., Desmaison, A., Antiga, L., Lerer, A.: Automatic differentiation in pytorch · 2017
Later among the works it cites.
arXiv preprint arXiv:1703.00560 (2017)
Tian, Y.: An analytical formula of population gradient for two-layered relu network and its applications in convergence and critical point analysis · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
arXiv preprint arXiv:1602.02830 (2016)
Hubara, I., Courbariaux, M., Soudry, D., El-Yaniv, R., Bengio, Y.: Binarized neural networks: Training neural networks with weights and activations constrained to +1 or -1 · 2016
Cited alongside, same era.
arXiv preprint arXiv:1605.04711 (2016)
Li, F., Zhang, B., Liu, B.: Ternary weight networks · 2016
Cited alongside, same era.
In: European Conference on Computer Vision (ECCV) (2016)
Rastegari, M., Ordonez, V., Redmon, J., Farhadi, A.: Xnor-net: Imagenet classification using binary convolutional neural networks · 2016
Cited alongside, same era.
arXiv preprint arXiv: 1606.06160 (2016)
Zhou, S., Wu, Y., Ni, Z., Zhou, X., Wen, H., Zou, Y.: Dorefa-net: Training low bitwidth convolutional neural networks with low bitwidth gradients · 2016
Cited alongside, same era.
arXiv preprint arXiv:1702.07966 (2017)
Brutzkus, A., Globerson, A.: Globally optimal gradient descent for a convnet with gaussian inputs · 2017
Cited alongside, same era.
In: IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2017)
Cai, Z., He, X., Sun, J., Vasconcelos, N.: Deep learning with low precision by half-wave gaussian quantization · 2017
Cited alongside, same era.
arXiv preprint arXiv:1707.01209 (2017)
Carreira-Perpinán, M.: Model compression as constrained optimization, with application to neural nets. part i: General framework · 2017
Cited alongside, same era.
arXiv preprint arXiv:1805.06085 (2018)
Choi, J., Wang, Z., Venkataramani, S., Chuang, P.I.J., Srinivasan, V., Gopalakrishnan, K.: Pact: Parameterized clipping activation for quantized neural networks · 2018
Closest in time.
arXiv preprint arXiv:1712.00779 (2018)
Du, S.S., Lee, J.D., Tian, Y., Poczos, B., Singh, A.: Gradient descent learns one-hidden-layer cnn: Don’t be afraid of spurious local minimum · 2018
Closest in time.
arXiv preprint arXiv:1807.03973 (2018)
He, J., Li, L., Xu, J., Zheng, C.: Relu deep neural networks and linear finite elements · 2018
Closest in time.
Journal of Machine Learning Research 18
Hubara, I., Courbariaux, M., Soudry, D., El-Yaniv, R., Bengio, Y.: Quantized neural networks: Training neural networks with low precision weights and activations · 2018
Closest in time.
arXiv preprint arXiv:1802.00168 (2018)
Wang, B., Luo, X., Li, Z., Zhu, W., Shi, Z., Osher, S.J.: Deep neural nets with interpolating function as output activation · 2018
Closest in time.
arXiv preprint arXiv:1801.06313; SIAM Journal on Imaging Sciences, to appear (2018)
Yin, P., Zhang, S., Lyu, J., Osher, S., Qi, Y., Xin, J.: Binaryrelax: A relaxation approach for training deep neural networks with quantized weights · 2018
Closest in time.
arXiv preprint arXiv:1612.06052; J. Comput. Math., to appear (2018)
Yin, P., Zhang, S., Qi, Y., Xin, J.: Quantization and training of low bit-width convolutional neural networks for object detection · 2018
Closest in time.