Fetching the paper…
Reading the bibliography…
Backpropagation algorithm is indispensable for the training of feedforward neural networks.
A stochastic approximation method
Robbins, H. and Monro, S · 1951
Earlier work this paper cites.
Learning representations by back-propagating errors
Rumelhart, D. E., Hinton, G. E., Williams, R. J., et al · 1988
Earlier work this paper cites.
On the momentum term in gradient descent learning algorithms
Qian, N · 1999
Earlier work this paper cites.
Learning deep architectures for ai
Bengio, Y. et al · 2009
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Krizhevsky, A. and Hinton, G · 2009
Earlier work this paper cites.
Neural networks for machine learning-lecture 6a-overview of mini-batch gradient descent, 2012
Hinton, G., Srivastava, N., and Swersky, K · 2012
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Krizhevsky, A., Sutskever, I., and Hinton, G. E · 2012
Earlier work this paper cites.
Multi-gpu training of convnets
Yadan, O., Adams, K., Taigman, Y., and Ranzato, M · 2013
Earlier work this paper cites.
Distributed optimization of deeply nested systems
Carreira-Perpinan, M. and Wang, W · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, D. and Ba, J · 2014
Cited alongside, same era.
Very deep convolutional networks for large-scale image recognition
Simonyan, K. and Zisserman, A · 2014
Cited alongside, same era.
Kickback cuts backprop’s red-tape: Biologically plausible credit assignment in neural networks
Balduzzi, D., Vanchinathan, H., and Buhmann, J. M · 2015
Cited alongside, same era.
Deep learning
LeCun, Y., Bengio, Y., and Hinton, G · 2015
Cited alongside, same era.
Going deeper with convolutions
Szegedy, C., Liu, W., Jia, Y., Sermanet, P., Reed, S., Anguelov, D., Erhan, D., Vanhoucke, V., and Rabinovich, A · 2015
Cited alongside, same era.
Densely connected convolutional networks
Huang, G., Liu, Z., Weinberger, K. Q., and van der Maaten, L · 2016
Later among the works it cites.
Decoupled neural interfaces using synthetic gradients
Jaderberg, M., Czarnecki, W. M., Osindero, S., Vinyals, O., Graves, A., and Kavukcuoglu, K · 2016
Later among the works it cites.
Direct feedback alignment provides learning in deep neural networks
Nøkland, A · 2016
Later among the works it cites.
Training neural networks without gradients: A scalable admm approach
Taylor, G., Burmeister, R., Xu, Z., Singh, B., Patel, A., and Goldstein, T · 2016
Later among the works it cites.
Benefits of depth in neural networks
Telgarsky, M · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Bottou, L., Curtis, F. E., and Nocedal, J · 2016
Cited alongside, same era.
The power of depth for feedforward neural networks
Eldan, R. and Shamir, O · 2016
Cited alongside, same era.
Deep residual learning for image recognition
He, K., Zhang, X., Ren, S., and Sun, J · 2016
Cited alongside, same era.
Understanding synthetic gradients and decoupled neural interfaces
Czarnecki, W. M., Świrszcz, G., Jaderberg, M., Osindero, S., Vinyals, O., and Kavukcuoglu, K · 2017
Later among the works it cites.
Benchmarks for popular cnn models
Johnson, J · 2017
Later among the works it cites.
Automatic differentiation in pytorch
Paszke, A., Gross, S., Chintala, S., Chanan, G., Yang, E., DeVito, Z., Lin, Z., Desmaison, A., Antiga, L., and Lerer, A · 2017
Later among the works it cites.