Fetching the paper…
Reading the bibliography…
Machine learning has made tremendous progress in recent years and received large amounts of public attention.
Survey on automated machine learning
Zoeller, M. and Huber, M. (2019) · 1904
Earlier work this paper cites.
Dropout: A Simple Way to Prevent Neural Networks from Overfitting
Srivastava, N., Hinton, G., Krizhevsky, A., Sutskever, I., and Salakhutdinov, R. (2014) · 1958
Earlier work this paper cites.
Neural Net Pruning - Why and How
Sietsma, J. (1988) · 1988
Earlier work this paper cites.
Multilayer Feedforward Networks Are Universal Approximators
Hornik, K., Stinchcombe, M., and White, H. (1989) · 1989
Earlier work this paper cites.
Learning in Feedforward Layered Networks: The Tiling Algorithm
Mezard, M. and Nadal, J.-P. (1989) · 1989
Earlier work this paper cites.
Skeletonization: A Technique for Trimming the Fat from a Network via Relevance Assessment
Mozer, M. C. and Smolensky, P. (1989) · 1989
Earlier work this paper cites.
The Cascade-Correlation Learning Architecture
Fahlman, S. E. and Lebiere, C. (1990) · 1990
Earlier work this paper cites.
The Upstart Algorithm: A Method for Constructing and Training Feedforward Neural Networks
Frean, M. (1990) · 1990
Earlier work this paper cites.
Meiosis Networks
Hanson, S. J. (1990) · 1990
Earlier work this paper cites.
Optimal Brain Damage
LeCun, Y., Denker, J. S., and Solla, S. A. (1990) · 1990
Earlier work this paper cites.
The Recurrent Cascade-Correlation Architecture
Fahlman, S. E. (1991) · 1991
Earlier work this paper cites.
A Conceptual Approach to Generalisation in Dynamic Neural Networks
Sjogaard, S. (1991) · 1991
Earlier work this paper cites.
A Simple Weight Decay Can Improve Generalization
Krogh, A. and Hertz, J. A. (1992) · 1992
Earlier work this paper cites.
Cascade network architectures
Littmann, E. and Ritter, H. (1992) · 1992
Earlier work this paper cites.
Simplifying Neural Networks by Soft Weight-Sharing
Nowlan, S. J. and Hinton, G. E. (1992) · 1992
Earlier work this paper cites.
Node Splitting: A Constructive Algorithm for Feed-Forward Neural Networks
Wynne-Jones, M. (1992) · 1992
Earlier work this paper cites.
Optimal brain surgeon and general network pruning
Hassibi, B., Stork, D. G., and Wolff, G. J. (1993) · 1993
Earlier work this paper cites.
Generalization Abilities of Cascade Network Architecture
Littmann, E. and Ritter, H. (1993) · 1993
Earlier work this paper cites.
Pruning algorithms - a survey
Reed, R. (1993) · 1993
Earlier work this paper cites.
Optimal Brain Surgeon: Extensions and Performance Comparisons
Hassibi, B., Stork, D. G., and Wolff, G. (1994) · 1994
Earlier work this paper cites.
Fast Pruning Using Principal Components
Levin, A. U., Leen, T. K., and Moody, J. E. (1994) · 1994
Earlier work this paper cites.
A Comparison of the Computational Power of Sigmoid and Boolean Threshold Circuits
Maass, W., Schnitger, G., and Sontag, E. D. (1994) · 1994
Cited alongside, same era.
A Procedure for Determining the Topology of Multilayer Feedforward Neural Networks
Wang, Z., Di Massimo, C., Tham, M. T., and Morris, A. J. (1994) · 1994
Cited alongside, same era.
Dynamic learning algorithms
Waugh, S. (1994) · 1994
Cited alongside, same era.
Investigation of the Cascor Family of Learning Algorithms
Prechelt, L. (1997) · 1997
Cited alongside, same era.
Convolutional Networks for Images, Speech, and Time Series
LeCun, Y. and Bengio, Y. (1998) · 1998
Cited alongside, same era.
Gradient-Based Learning Applied to Document Recognition
LeCun, Y., Bottou, L., Bengio, Y., Haffner, P., et al. (1998) · 1998
Cited alongside, same era.
Very Deep Convolutional Networks for Text Classification
Conneau, A., Schwenk, H., Barrault, L., and Lecun, Y. (2016) · 2016
Later among the works it cites.
The Power of Depth for Feedforward Neural Networks
Eldan, R. and Shamir, O. (2016) · 2016
Later among the works it cites.
Deep Learning
Goodfellow, I., Bengio, Y., and Courville, A. (2016) · 2016
Later among the works it cites.
Deep Residual Learning for Image Recognition
He, K., Zhang, X., Ren, S., and Sun, J. (2016) · 2016
Later among the works it cites.
Ask Me Anything: Dynamic Memory Networks for Natural Language Processing
Kumar, A., Irsoy, O., Ondruska, P., Iyyer, M., Bradbury, J., Gulrajani, I., Zhong, V., Paulus, R., and Socher, R. (2016) · 2016
Later among the works it cites.
End-To-End Training of Deep Visuomotor Policies
Levine, S., Finn, C., Darrell, T., and Abbeel, P. (2016) · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Experiments with the Cascade-Correlation Algorithm
Yang, J. and Honavar, V. (1998) · 1998
Cited alongside, same era.
Almost Linear VC Dimension Bounds for Piecewise Polynomial Networks
Bartlett, P. L., Maiorov, V., and Meir, R. (1999) · 1999
Cited alongside, same era.
Evolving Artificial Neural Networks
Yao, X. (1999) · 1999
Cited alongside, same era.
An Empirical Evaluation of Deep Architectures on Problems with Many Factors of Variation
Larochelle, H., Erhan, D., Courville, A., Bergstra, J., and Bengio, Y. (2007) · 2007
Cited alongside, same era.
Random Search for Hyper-Parameter Optimization
Bergstra, J. and Bengio, Y. (2012) · 2012
Cited alongside, same era.
A Practical Guide to Training Restricted Boltzmann Machines
Hinton, G. E. (2012) · 2012
Cited alongside, same era.
Later among the works it cites.
Exponential Expressivity in Deep Neural Networks Through Transient Chaos
Poole, B., Lahiri, S., Raghu, M., Sohl-Dickstein, J., and Ganguli, S. (2016) · 2016
Later among the works it cites.
An Overview of Gradient Descent Optimization Algorithms
Ruder, S. (2016) · 2016
Later among the works it cites.
Google’s Neural Machine Translation System: Bridging the Gap Between Human and Machine Translation
Wu, Y., Schuster, M., Chen, Z., Le, Q. V., Norouzi, M., Macherey, W., Krikun, M., Cao, Y., Gao, Q., Macherey, K., et al. (2016) · 2016
Later among the works it cites.
Understanding Deep Learning Requires Rethinking Generalization
Zhang, C., Bengio, S., Hardt, M., Recht, B., and Vinyals, O. (2016) · 2016
Later among the works it cites.
Adanet: Adaptive Structural Learning of Artificial Neural Networks
Cortes, C., Gonzalvo, X., Kuznetsov, V., Mohri, M., and Yang, S. (2017) · 2017
Later among the works it cites.
Let’s evolve a neural network with a genetic algorithm — code included
Harvey, M. (2017) · 2017
Later among the works it cites.
Forward Thinking: Building and Training Neural Networks One Layer at a Time
Hettinger, C., Christensen, T., Ehlert, B., Humpherys, J., Jarvis, T., and Wade, S. (2017) · 2017
Later among the works it cites.
ImageNet Classification with Deep Convolutional Neural Networks
Krizhevsky, A., Sutskever, I., and Hinton, G. E. (2017) · 2017
Later among the works it cites.
On the Expressive Power of Deep Neural Networks
Raghu, M., Poole, B., Kleinberg, J., Ganguli, S., and Dickstein, J. S. (2017) · 2017
Later among the works it cites.
Regularized evolution for image classifier architecture search
Real, E., Aggarwal, A., Huang, Y., and Le, Q. V. (2018) · 2018
Later among the works it cites.
A General Reinforcement Learning Algorithm that Masters Chess, Shogi, and Go through Self-Play
Silver, D., Hubert, T., Schrittwieser, J., Antonoglou, I., Lai, M., Guez, A., Lanctot, M., Sifre, L., Kumaran, D., Graepel, T., et al. (2018) · 2018
Later among the works it cites.
Learning Transferable Architectures for Scalable Image Recognition
Zoph, B., Vasudevan, V., Shlens, J., and Le, Q. V. (2018) · 2018
Later among the works it cites.
Neural Architecture Search: A Survey
Elsken, T., Metzen, J. H., and Hutter, F. (2019) · 2019
Closest in time.
Simple Deep Neural Network on the MNIST Dataset
Keras (2019) · 2019
Closest in time.