Cybernetic predicting devices
Ivakhnenko, A. G. and Lapa, V · 1965
Earlier work this paper cites.
Approximation by superpositions of a sigmoidal function
Cybenko, G · 1989
Earlier work this paper cites.
Optimal brain damage
LeCun, Y., Denker, J. S., and Solla, S. A · 1990
Earlier work this paper cites.
Approximation and estimation bounds for artificial neural networks
Barron, A. R · 1994
Earlier work this paper cites.
Training mlps layer by layer using an objective function for internal representations
Lengellé, R. and Denoeux, T · 1996
Earlier work this paper cites.
Approximation theory of the mlp model in neural networks
Pinkus, A · 1999
Earlier work this paper cites.
Convex neural networks
Bengio, Y., Roux, N. L., Vincent, P., Delalleau, O., and Marcotte, P · 2006
Earlier work this paper cites.
A fast learning algorithm for deep belief nets
Hinton, G. E., Osindero, S., and Teh, Y.-W · 2006
Earlier work this paper cites.
Greedy layer-wise training of deep networks
Bengio, Y., Lamblin, P., Popovici, D., and Larochelle, H · 2007
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Krizhevsky, A., Sutskever, I., and Hinton, G. E · 2012
Earlier work this paper cites.
Image classification with the fisher vector: Theory and practice
Sánchez, J., Perronnin, F., Mensink, T., and Verbeek, J · 2013
Earlier work this paper cites.
Provable bounds for learning some deep representations
Arora, S., Bhaskara, A., Ge, R., and Ma, T · 2014
Earlier work this paper cites.
Breaking the curse of dimensionality with convex neural networks
Original
Bach, F · 2014
Earlier work this paper cites.
Dark knowledge
Hinton, G., Vinyals, O., and Dean, J · 2014
Earlier work this paper cites.
How transferable are features in deep neural networks?
Yosinski, J., Clune, J., Bengio, Y., and Lipson, H · 2014
Earlier work this paper cites.
Visualizing and understanding convolutional networks
Zeiler, M. D. and Fergus, R · 2014
Earlier work this paper cites.
Beating the perils of non-convexity: Guaranteed training of neural networks using tensor methods
Original
Janzamin, M., Sedghi, H., and Anandkumar, A · 2015
Earlier work this paper cites.
Deeply-supervised nets
Lee, C.-Y., Xie, S., Gallagher, P., Zhang, Z., and Tu, Z · 2015
Earlier work this paper cites.
Deep roto-translation scattering for object classification
Oyallon, E. and Mallat, S · 2015
Earlier work this paper cites.