Fetching the paper…
Reading the bibliography…
Recent advances in training deep (multi-layer) architectures have inspired a renaissance in neural network use.
D. W. Marquardt and R. D. Snee, “Ridge regression in practice,” The American Statistician , vol. 29, pp. 3–20, 1975
1975
Earlier work this paper cites.
D. E. Rumelhart, G. E. Hinton, and R. J. Williams, “Learning representations by back-propagating errors,” Nature , vol. 323, pp. 533–536, 1986
1986
Earlier work this paper cites.
P. F. Schmidt, M. A. Kraaijveld, and R. P. W. Duin, “Feed forward neural networks with random weights,” in Proc. 11th IAPR Int. Conf. on Pattern Recognition, Volume II, Conf. B: Pattern Recognition Methodology and Systems (ICPR11, The Hague, Aug.30 - Sep.3), IEEE Computer Society Press, Los Alamitos, CA, 1992, 1-4 , 1992
1992
Earlier work this paper cites.
W. H. Press, S. A. Teukolsky, W. T. Vetterling, and B. P. Flannery, Numerical Recipes in C: The Art of Scientific Computing , 2nd ed. Cambridge University Press, Cambridge, UK, 1992
1992
Earlier work this paper cites.
C. L. P. Chen, “A rapid supervised learning neural network for function interpolation and approximation,” IEEE Transactions on Neural Networks , vol. 7, pp. 1220–1230, 1996
1996
Earlier work this paper cites.
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner, “Gradient-based learning applied to document recognition,” Proceedings of the IEEE , vol. 86, pp. 2278–2324, 1998
1998
Earlier work this paper cites.
C. Eliasmith and C. H. Anderson, “Developing and applying a toolkit from a general neurocomputational framework,” Neurocomputing , vol. 26, pp. 1013–1018, 1999
1999
Earlier work this paper cites.
——, Neural Engineering: Computation, Representation, and Dynamics in Neurobiological Systems . MIT Press, Cambridge, MA, 2003
2003
Earlier work this paper cites.
P. Y. Simard, D. Steinkraus, and J. C. Platt, “Best practices for convolutional neural networks applied to visual document analysis,” in Proceedings of the Seventh International Conference on Document Analysis and Recognition (ICDAR 2003) , 2003
2003
Earlier work this paper cites.
Y. LeCun, F. J. Huang, and L. Bottou, “Learning methods for generic object recognition with invariance to pose and lighting,” in Proceedings IEEE Computer Society Conference on Computer Vision and Pattern Recognition (CVPR) , vol. 2, 2004, pp. 97–104
2004
Earlier work this paper cites.
G. E. Hinton, S. Osindero, and Y.-W. Teh, “A fast learning algorithm for deep belief nets,” Neural computation , vol. 18, pp. 1527–1554, 2006
2006
Earlier work this paper cites.
G.-B. Huang, Q.-Y. Zhu, and C.-K. Siew, “Extreme learning machine: Theory and applications,” Neurocomputing , vol. 70, pp. 489–501, 2006
2006
Earlier work this paper cites.
N.-Y. Liang, G.-B. Huang, P. Saratchandran, and N. Sundararajan, “A fast and accurate online sequential learning algorithm for feedforward networks,” IEEE Transactions on Neural Networks , vol. 17, pp. 1411–1423, 2006
2006
Earlier work this paper cites.
K. Jarrett, K. Kavukcuoglu, M. Ranzato, and Y. LeCun, “What is the best multi-stage architecture for object recognition?” in In Proc. IEEE 12th International Conference on Computer Vision , 2009
2009
Cited alongside, same era.
D. Cireşan, U. Meier, L. M. Gambardella, and J. Schmidhuber, “Deep, big, simple neural nets for handwritten digit recognition,” Neural Computation , vol. 22, pp. 3207–3220, 2010
2010
Cited alongside, same era.
V. Nair and G. E. Hinton, “Rectified linear units improve restricted Boltzmann machines,” in Proceedings of the 27th International Conference on Machine Learning (ICML), Haifa, Israel , 2010
2010
Cited alongside, same era.
A. Coates, H. Lee, and A. Y. Ng, “An analysis of single-layer networks in unsupervised feature learning,” in Proc.14th International Conference on Artificial Intelligence and Statistics (AISTATS), 2011, Fort Lauderdale, FL, USA. Volume 15 of JMLR:W&CP 15 , 2011
2011
Cited alongside, same era.
L. L. C. Kasun, H. Zhou, G.-B. Huang, and C. M. Vong, “Representational learning with extreme learning machine for big data,” IEEE Intelligent Systems , vol. 28, pp. 31–34, 2013
2013
Later among the works it cites.
B. Widrow, A. Greenblatt, Y. Kim, and D. Park, “The No-Prop algorithm: A new learning algorithm for multilayer neural networks,” Neural Networks , vol. 37, pp. 182–188, 2013
2013
Later among the works it cites.
B. Widrow, “Reply to the comments on the “No-Prop” algorithm,” Neural Networks , vol. 48, p. 204, 2013
2013
Later among the works it cites.
J. Tapson and A. van Schaik, “Learning the pseudoinverse solution to network weights,” Neural Networks , vol. 45, pp. 94–100, 2013
2013
Later among the works it cites.
M.-H. Lim, “Comments on the “No-Prop” algorithm,” Neural Networks , vol. 48, pp. 59–60, 2013
2013
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
D. C. D. Cireşan, U. Meier, J. Masci, L. M. Gambardella, and J. Schmidhuber, “Flexible, high performance convolutional neural networks for image classification,” in Proceedings of the Twenty-Second International Joint Conference on Artificial Intelligence , 2011, pp. 1237–1242
2011
Cited alongside, same era.
D. Cireşan, U. Meier, and J. Schmidhuber, “Multi-column deep neural networks for image classification,” in Proc. CVPR , 2012, pp. 3642–3649
2012
Cited alongside, same era.
G.-B. Huang, H. Zhou, X. Ding, and R. Zhang, “Extreme learning machine for regression and multiclass classification,” IEEE Transactions on Systems, Man, and Cybernetics—Part B: Cybernetics , vol. 42, pp. 513–529, 2012
2012
Cited alongside, same era.
D. Yu and L. Ding, “Efficient and effective algorithms for training single-hidden-layer neural networks,” Pattern Recognition Letters , vol. 33, pp. 554–558, 2012
2012
Cited alongside, same era.
C. Eliasmith, T. C. Stewart, X. Choo, T. Bekolay, T. DeWolf, C. Tang, and D. Rasmussen, “A large-scale model of the functioning brain,” Science , vol. 338, pp. 1202–1205, 2012
2012
Cited alongside, same era.
L. Wan, M. Zeiler, S. Zhang, Y. LeCun, and R. Fergus, “Regularization of neural networks using DropConnect,” in Proceedings of the 30th International Conference on Machine Learning, Atlanta, Georgia, USA; JMLR: W&CP volume 28 , 2013
2013
Cited alongside, same era.
M. D. Zelier and R. Fergus, “Stochastic pooling for regularization of deep convolutional neural networks,” in In Proc. International Conference on Learning Representations, Scottsdale, USA, 2013 , 2013
2013
Cited alongside, same era.
I. J. Goodfellow, D. Warde-Farley, M. Mirza, A. Courville, and Y. Bengio, “Maxout networks,” in Proceedings of the 30th International Conference on Machine Learning, Atlanta, Georgia, USA; JMLR: W&CP volume 28 , 2013
2013
Cited alongside, same era.
Y. LeCun, C. Cortes, and C. J. C. Burges, “The MNIST database of handwritten digits,” Accessed August 2014, http://yann.lecun.com/exdb/mnist/
2014
Closest in time.
C.-Y. Lee, S. Xie, P. Gallagher, Z. Zhang, and Z. Tu, “Deeply-supervised nets,” in Deep Learning and Representation Learning Workshop, NIPS , 2014
2014
Closest in time.
2014
Closest in time.
G.-B. Huang, “An insight into extreme learning machines: Random neurons, random features and kernels,” Cognitive Computation , vol. 6, pp. 376–390, 2014
2014
Closest in time.
2014
Closest in time.
2015
Closest in time.
A. van Schaik and J.Tapson, “Online and adaptive pseudoinverse solutions for ELM weights,” Neurocomputing , vol. 149, pp. 233–238, 2015
2015
Closest in time.