Fetching the paper…
Reading the bibliography…
Deep Neural Networks have been shown to be beneficial for a variety of tasks, in particular allowing for end-to-end learning and reducing the requirement for manual design decisions.
Training feedforward neural networks using genetic algorithms
Montana, D.J., Davis, L.: · 1989
Earlier work this paper cites.
Approximation capabilities of multilayer feedforward networks
Hornik, K.: · 1991
Earlier work this paper cites.
Combinations of genetic algorithms and neural networks: A survey of the state of the art
Schaffer, J.D., Whitley, D., Eshelman, L.J.: · 1992
Earlier work this paper cites.
Pruning neural nets by genetic algorithm
Hancock, P.J.B.: · 1992
Earlier work this paper cites.
Stable function approximation in dynamic programming
Gordon, G.J.: · 1995
Earlier work this paper cites.
An Introduction to Genetic Algorithms
Mitchell, M.: · 1996
Earlier work this paper cites.
The vanishing gradient problem during learning recurrent neural nets and problem solutions
Hochreiter, S.: · 1998
Earlier work this paper cites.
An overview of evolutionary algorithms: Practical issues and common pitfalls
Whitley, D.: · 2001
Earlier work this paper cites.
Evolving neural networks through augmenting topologies
Stanley, K.O., Miikkulainen, R.: · 2002
Earlier work this paper cites.
Neuroevolution for reinforcement learning using evolution strategies
Igel, C.: · 2003
Earlier work this paper cites.
Evolutionary Computation: A Unified Approach
De Jong, K.A.: · 2006
Earlier work this paper cites.
A hypercube-based encoding for evolving large-scale neural networks
Stanley, K.O., D’Ambrosio, D.B., Gauci, J.: · 2009
Earlier work this paper cites.
Rectified linear units improve restricted boltzmann machines
Nair, V., Hinton, G.E.: · 2010
Cited alongside, same era.
An overview of some classical growing neural networks and new developments
Qiang, X., Cheng, G., Wang, Z.: · 2010
Cited alongside, same era.
Deep sparse rectifier neural networks
Glorot, X., Bordes, A., Bengio, Y.: · 2011
Cited alongside, same era.
On the importance of initialization and momentum in deep learning
Sutskever, I., Martens, J., Dahl, G., Hinton, G.: · 2013
Cited alongside, same era.
Rectifier nonlinearities improve neural network acoustic models
Maas, A.L., Hannun, A.Y., Ng, A.Y.: · 2013
Cited alongside, same era.
Deep learning
LeCun, Y., Bengio, Y., Hinton, G.: · 2015
Cited alongside, same era.
CMA-ES for hyperparameter optimization of deep neural networks
Loshchilov, I., Hutter, F.: · 2016
Later among the works it cites.
Identity mappings in deep residual networks
He, K., Zhang, X., Ren, S., Sun, J.: · 2016
Later among the works it cites.
All you need is a good init
Mishkin, D., Matas, J.: · 2017
Later among the works it cites.
Self-normalizing neural networks
Klambauer, G., Unterthiner, T., Mayr, A., Hochreiter, S.: · 2017
Later among the works it cites.
Miikkulainen, R., Liang, J.Z., Meyerson, E., Rawal, A., Fink, D., Francon, O., Raju, B., Shahrzad, H., Navruzyan, A., Duffy, N., Hodjat, B.: · 2017
Later among the works it cites.
A genetic programming approach to designing convolutional neural network architectures
Suganuma, M., Shirakawa, S., Nagao, T.: · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Very deep convolutional networks for large-scale image recognition
Simonyan, K., Zisserman, A.: · 2015
Cited alongside, same era.
Delving deep into rectifiers: Surpassing human-level performance on imagenet classification
He, K., Zhang, X., Ren, S., Sun, J.: · 2015
Cited alongside, same era.
Deep Learning
Goodfellow, I., Bengio, Y., Courville, A.: · 2016
Cited alongside, same era.
Batch normalized recurrent neural networks
Laurent, C., Pereyra, G., Brakel, P., Zhang, Y., Bengio, Y.: · 2016
Cited alongside, same era.
Fast and accurate deep network learning by exponential linear units (ELUs)
Clevert, D., Unterthiner, T., Hochreiter, S.: · 2016
Cited alongside, same era.
Noisy activation functions
Gulcehre, C., Moczulski, M., Denil, M., Bengio, Y.: · 2016
Cited alongside, same era.
Later among the works it cites.
Evolving parsimonious networks by mixing activation functions
Hagg, A., Mensing, M., Asteroth, A.: · 2017
Later among the works it cites.
Searching for activation functions
Ramachandran, P., Zoph, B., V. Le, Q.: · 2018
Closest in time.
Sigmoid-weighted linear units for neural network function approximation in reinforcement learning
Elfwing, S., Uchibe, E., Doya, K.: · 2018
Closest in time.
On the selection of initialization and activation function for deep neural networks
Hayou, S., Doucet, A., Rousseau, J.: · 2018
Closest in time.
Hierarchical representations for efficient architecture search
Liu, H., Simonyan, K., Vinyals, O., Fernando, C., Kavukcuoglu, K.: · 2018
Closest in time.