D. O. Hebb, The organization of behavior: A neuropsychological theory . New York: Wiley, Jun. 1949
1949
Earlier work this paper cites.
P. Smolensky, “Parallel distributed processing: Explorations in the microstructure of cognition, vol. 1,” D. E. Rumelhart, J. L. McClelland, and C. PDP Research Group, Eds. Cambridge, MA, USA: MIT Press, 1986, ch. Information Processing in Dynamical Systems: Foundations of Harmony Theory, pp. 194–281. [Online]. Available: http://dl.acm.org/citation.cfm?id=104279.104290
1986
Earlier work this paper cites.
D. S. Broomhead and D. Lowe, “Radial basis functions, multi-variable functional interpolation and adaptive networks,” Royal Signals and Radar Establishment Malvern (United Kingdom), Tech. Rep., 1988
1988
Earlier work this paper cites.
K. Hornik, M. Stinchcombe, and H. White, “Multilayer feedforward networks are universal approximators,” Neural Networks , vol. 2, no. 5, pp. 359 – 366, 1989. [Online]. Available: http://www.sciencedirect.com/science/article/pii/0893608089900208
1989
Earlier work this paper cites.
J. Park and I. W. Sandberg, “Universal approximation using radial-basis-function networks,” Neural Computation , vol. 3, no. 2, pp. 246–257, 1991. [Online]. Available: https://doi.org/10.1162/neco.1991.3.2.246
1991
Earlier work this paper cites.
S. Lowel and W. Singer, “Selection of intrinsic horizontal connections in the visual cortex by correlated neuronal activity,” Science , vol. 255, no. 5041, pp. 209–212, 1992. [Online]. Available: http://science.sciencemag.org/content/255/5041/209
1992
Earlier work this paper cites.
L. Xu, “Least mean square error reconstruction principle for self-organizing neural-nets,” Neural networks , vol. 6, no. 5, pp. 627–648, 1993
1993
Earlier work this paper cites.
J. Han and C. Moraga, “The influence of the sigmoid function parameters on the speed of backpropagation learning,” in Proceedings of the International Workshop on Artificial Neural Networks: From Natural to Artificial Neural Computation , ser. IWANN ’96. London, UK, UK: Springer-Verlag, 1995, pp. 195–201. [Online]. Available: http://dl.acm.org/citation.cfm?id=646366.689307
1995
Earlier work this paper cites.
Y. Lecun, L. Bottou, Y. Bengio, and P. Haffner, “Gradient-based learning applied to document recognition,” in Proceedings of the IEEE , 1998, pp. 2278–2324
1998
Earlier work this paper cites.
R. S. Sutton and A. G. Barto, Reinforcement Learning: An Introduction . MIT Press, 1998. [Online]. Available: http://www.cs.ualberta.ca/~sutton/book/the-book.html
1998
Earlier work this paper cites.
P. Viola and M. Jones, “Rapid object detection using a boosted cascade of simple features,” in Proceedings of the 2001 IEEE Computer Society Conference on Computer Vision and Pattern Recognition. CVPR 2001 , vol. 1, 2001, pp. I–511–I–518 vol.1
2001
Earlier work this paper cites.
G. E. Hinton, S. Osindero, and Y.-W. Teh, “A fast learning algorithm for deep belief nets,” Neural Comput. , vol. 18, no. 7, pp. 1527–1554, Jul. 2006. [Online]. Available: http://dx.doi.org/10.1162/neco.2006.18.7.1527
2006
Earlier work this paper cites.
G. E. Hinton and R. R. Salakhutdinov, “Reducing the dimensionality of data with neural networks,” science , vol. 313, no. 5786, pp. 504–507, 2006
2006
Earlier work this paper cites.
H. Grabner, P. M. Roth, and H. Bischof, “Eigenboosting: Combining discriminative and generative information,” in 2007 IEEE Conference on Computer Vision and Pattern Recognition , June 2007, pp. 1–8
2007
Earlier work this paper cites.
A. Krizhevsky and G. Hinton, “Learning multiple layers of features from tiny images,” Citeseer, Tech. Rep., 2009
2009
Earlier work this paper cites.