Fetching the paper…
Reading the bibliography…
Evolutionary computation methods have been successfully applied to neural networks since two decades ago, while those methods cannot scale well to the modern deep neural networks due to the complicated architectures and large quantities of connection weights.
A. L. Hodgkin and A. F. Huxley, “A quantitative description of membrane current and its application to conduction and excitation in nerve,” The Journal of physiology , vol. 117, no. 4, pp. 500–544, 1952
1952
Earlier work this paper cites.
D. H. Hubel and T. N. Wiesel, “Receptive fields, binocular interaction and functional architecture in the cat’s visual cortex,” The Journal of Physiology , vol. 160, no. 1, pp. 106–154, 1962
1962
Earlier work this paper cites.
J. Halton and G. Smith, “Radical inverse quasi-random point sequence, algorithm 247,” Communications of the ACM , vol. 7, p. 701, 1964
1964
Earlier work this paper cites.
J. Močkus, “On bayesian methods for seeking the extremum,” in Optimization Techniques IFIP Technical Conference . Springer, 1975, pp. 400–404
1975
Earlier work this paper cites.
J. Beck and T. Fiala, ““integer-making” theorems,” Discrete Applied Mathematics , vol. 3, no. 1, pp. 1–8, 1981
1981
Earlier work this paper cites.
A. Blumer, A. Ehrenfeucht, D. Haussler, and M. K. Warmuth, “Occam’s razor,” Information Processing Letters , vol. 24, no. 6, pp. 377–380, 1987
1987
Earlier work this paper cites.
D. E. Rumelhart, G. E. Hinton, and R. J. Williams, “Learning representations by back-propagating errors,” Cognitive Modeling , vol. 5, no. 3, p. 1, 1988
1988
Earlier work this paper cites.
H. Bourlard and Y. Kamp, “Auto-association by multilayer perceptrons and singular value decomposition,” Biological Cybernetics , vol. 59, no. 4-5, pp. 291–294, 1988
1988
Earlier work this paper cites.
Y. LeCun, B. Boser, J. S. Denker, D. Henderson, R. E. Howard, W. Hubbard, and L. D. Jackel, “Backpropagation applied to handwritten zip code recognition,” Neural Computation , vol. 1, no. 4, pp. 541–551, 1989
1989
Earlier work this paper cites.
G. Miller, “Designing neural networks using genetic algorithms,” Proceedings of ICGA-89 , 1989
1989
Earlier work this paper cites.
G. E. Hinton and R. S. Zemel, “Autoencoders, minimum description length, and helmholtz free energy,” Advances in Neural Information Processing Systems , pp. 3–3, 1994
1994
Earlier work this paper cites.
K. Deb and R. B. Agrawal, “Simulated binary crossover for continuous search space,” Complex Systems , vol. 9, no. 3, pp. 1–15, 1994
1994
Earlier work this paper cites.
T. Back, Evolutionary algorithms in theory and practice: evolution strategies, evolutionary programming, genetic algorithms . Oxford university press, 1996
1996
Earlier work this paper cites.
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner, “Gradient-based learning applied to document recognition,” Proceedings of the IEEE , vol. 86, no. 11, pp. 2278–2324, 1998
1998
Earlier work this paper cites.
M. Mitchell, An introduction to genetic algorithms . MIT press, 1998
1998
Earlier work this paper cites.
X. Yao, “Evolving artificial neural networks,” Proceedings of the IEEE , vol. 87, no. 9, pp. 1423–1447, 1999
1999
Earlier work this paper cites.
K. Deb, Multi-objective optimization using evolutionary algorithms . John Wiley & Sons, 2001, vol. 16
2001
Earlier work this paper cites.
K. O. Stanley and R. Miikkulainen, “Evolving neural networks through augmenting topologies,” Evolutionary computation , vol. 10, no. 2, pp. 99–127, 2002
2002
Earlier work this paper cites.
K. Deb, A. Pratap, S. Agarwal, and T. Meyarivan, “A fast and elitist multiobjective genetic algorithm: Nsga-ii,” IEEE Transactions on Evolutionary Computation , vol. 6, no. 2, pp. 182–197, 2002
2002
Earlier work this paper cites.
F. Ning, D. Delhomme, Y. LeCun, F. Piano, L. Bottou, and P. E. Barbano, “Toward automatic phenotyping of developing embryos from videos,” IEEE Transactions on Image Processing , vol. 14, no. 9, pp. 1360–1371, 2005
2005
Earlier work this paper cites.
G. E. Hinton and R. R. Salakhutdinov, “Reducing the dimensionality of data with neural networks,” Science , vol. 313, no. 5786, pp. 504–507, 2006
2006
Earlier work this paper cites.
C. E. Rasmussen and C. K. Williams, Gaussian processes for machine learning . MIT press Cambridge, 2006, vol. 1
2006
Cited alongside, same era.
D. Ashlock, Evolutionary computation for modeling and optimization . Springer Science & Business Media, 2006
2006
Cited alongside, same era.
K. O. Stanley, “Compositional pattern producing networks: A novel abstraction of development,” Genetic programming and evolvable machines , vol. 8, no. 2, pp. 131–162, 2007
2007
Cited alongside, same era.
H. Larochelle, D. Erhan, A. Courville, J. Bergstra, and Y. Bengio, “An empirical evaluation of deep architectures on problems with many factors of variation,” in Proceedings of the 24th International Conference on Machine Learning . ACM, 2007, pp. 473–480
2007
Cited alongside, same era.
E. I. Atanassov, A. Karaivanova, and S. Ivanovska, “Tuning the generation of sobol sequence with owen scrambling.” in Large-Scale Scientific Computing . Springer, 2009, pp. 459–466
K. Sohn, G. Zhou, C. Lee, and H. Lee, “Learning and selecting features jointly with point-wise gated boltzmann machines,” in International Conference on Machine Learning , 2013, pp. 217–225
2013
Later among the works it cites.
J. Bruna and S. Mallat, “Invariant scattering convolution networks,” IEEE transactions on Pattern Analysis and Machine Intelligence , vol. 35, no. 8, pp. 1872–1886, 2013
2013
Later among the works it cites.
2014
Later among the works it cites.
M. D. Zeiler and R. Fergus, “Visualizing and understanding convolutional networks,” in European Conference on Computer Vision . Springer, 2014, pp. 818–833
2014
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2009
Cited alongside, same era.
K. O. Stanley, D. B. D’Ambrosio, and J. Gauci, “A hypercube-based encoding for evolving large-scale neural networks,” Artificial life , vol. 15, no. 2, pp. 185–212, 2009
2009
Cited alongside, same era.
X. Glorot and Y. Bengio, “Understanding the difficulty of training deep feedforward neural networks,” in Proceedings of the 13th International Conference on Artificial Intelligence and Statistics , 2010, pp. 249–256
2010
Cited alongside, same era.
J. S. Bergstra, R. Bardenet, Y. Bengio, and B. Kégl, “Algorithms for hyper-parameter optimization,” in Advances in Neural Information Processing Systems , 2011, pp. 2546–2554
2011
Cited alongside, same era.
F. Hutter, H. H. Hoos, and K. Leyton-Brown, “Sequential model-based optimization for general algorithm configuration.” LION , vol. 5, pp. 507–523, 2011
2011
Cited alongside, same era.
X. Glorot, A. Bordes, and Y. Bengio, “Deep sparse rectifier neural networks,” in Proceedings of the Fourteenth International Conference on Artificial Intelligence and Statistics , 2011, pp. 315–323
2011
Cited alongside, same era.
Y. Bengio and O. Delalleau, “On the expressive power of deep architectures,” in International Conference on Algorithmic Learning Theory . Springer, 2011, pp. 18–36
2011
Cited alongside, same era.
O. Delalleau and Y. Bengio, “Shallow vs. deep sum-product networks,” in Advances in Neural Information Processing Systems , 2011, pp. 666–674
2011
Cited alongside, same era.
M. N. Omidvar, X. Li, Y. Mei, and X. Yao, “Cooperative co-evolution with differential grouping for large scale optimization,” IEEE Transactions on Evolutionary Computation , vol. 18, no. 3, pp. 378–393, 2014
2014
Later among the works it cites.
2014
Later among the works it cites.
Y. Jia, E. Shelhamer, J. Donahue, S. Karayev, J. Long, R. Girshick, S. Guadarrama, and T. Darrell, “Caffe: Convolutional architecture for fast feature embedding,” in Proceedings of the 22nd ACM international conference on Multimedia . ACM, 2014, pp. 675–678
2014
Later among the works it cites.
M. Kim and L. Rigazio, “Deep clustered convolutional kernels,” in Feature Extraction: Modern Questions and Challenges , 2015, pp. 160–172
2015
Later among the works it cites.
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich, “Going deeper with convolutions,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2015, pp. 1–9
2015
Later among the works it cites.
T.-H. Chan, K. Jia, S. Gao, J. Lu, Z. Zeng, and Y. Ma, “Pcanet: A simple deep learning baseline for image classification?” IEEE Transactions on Image Processing , vol. 24, no. 12, pp. 5017–5032, 2015
2015
Later among the works it cites.
S. Ioffe and C. Szegedy, “Batch normalization: Accelerating deep network training by reducing internal covariate shift,” in International Conference on Machine Learning , 2015, pp. 448–456
2015
Later among the works it cites.
2015
Later among the works it cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2016, pp. 770–778
2016
Later among the works it cites.
C. Fernando, D. Banarse, M. Reynolds, F. Besse, D. Pfau, M. Jaderberg, M. Lanctot, and D. Wierstra, “Convolution by evolution: Differentiable pattern producing networks,” in Proceedings of the 2016 on Genetic and Evolutionary Computation Conference . ACM, 2016, pp. 109–116
2016
Later among the works it cites.
2016
Later among the works it cites.
2016
Later among the works it cites.
2016
Later among the works it cites.
2016
Later among the works it cites.
B. Baker, O. Gupta, N. Naik, and R. Raskar, “Designing neural network architectures using reinforcement learning,” International Conference on Learning Representations , 2017
2017
Closest in time.
2017
Closest in time.