Fetching the paper…
Reading the bibliography…
Evolutionary computation has been shown to be a highly effective method for training neural networks, particularly when employed at scale on CPU clusters.
Complexity: Life at the edge of chaos
Roger Lewin. 1999 · 1999
Earlier work this paper cites.
The spmd model: Past, present and future. In European Parallel Virtual Machine/Message Passing Interface Users’ Group Meeting . Springer
Frederica Darema. 2001 · 2001
Earlier work this paper cites.
Parameter-exploring policy gradients
Frank Sehnke, Christian Osendorfer, Thomas Rückstieß, Alex Graves, Jan Peters, and Jürgen Schmidhuber. 2010 · 2010
Earlier work this paper cites.
Neurons are poised near the edge of chaos
Leon Chua, Valery Sbitnev, and Hyongsuk Kim. 2012 · 2012
Earlier work this paper cites.
Sequence to sequence learning with neural networks. In Advances in NIPS . 3104–3112
I. Sutskever, O. Vinyals, and Q. Le. 2014 · 2014
Earlier work this paper cites.
REINFORCEjs
Andrej Karpathy. 2015 · 2015
Earlier work this paper cites.
G. Brockman, V. Cheung, L. Pettersson, J. Schneider, J. Schulman, J. Tang, and W. Zaremba. 2016 · 2016
Earlier work this paper cites.
On large-batch training for deep learning: Generalization gap and sharp minima
Nitish Shirish Keskar, Dheevatsa Mudigere, Jorge Nocedal, Mikhail Smelyanskiy, and Ping Tak Peter Tang. 2016 · 2016
Earlier work this paper cites.
An overview of gradient descent optimization algorithms
Sebastian Ruder. 2016 · 2016
Earlier work this paper cites.
Population based training of neural networks
Max Jaderberg, Valentin Dalibard, Simon Osindero, Wojciech M Czarnecki, Jeff Donahue, Ali Razavi, Oriol Vinyals, Tim Green, Iain Dunning, Karen Simonyan, et al · 2017
Earlier work this paper cites.
Evolution strategies as a scalable alternative to reinforcement learning
Tim Salimans, Jonathan Ho, Xi Chen, Szymon Sidor, and Ilya Sutskever. 2017 · 2017
Earlier work this paper cites.
Felipe Petroski Such, Vashisht Madhavan, Edoardo Conti, Joel Lehman, Kenneth O Stanley, and Jeff Clune. 2017 · 2017
Cited alongside, same era.
L2 regularization versus batch and weight normalization
Twan Van Laarhoven. 2017 · 2017
Cited alongside, same era.
JAX: composable transformations of Python+NumPy programs
James Bradbury, Roy Frostig, Peter Hawkins, Matthew James Johnson, Chris Leary, Dougal Maclaurin, George Necula, Adam Paszke, Jake VanderPlas, Skye Wanderman-Milne, and Qiao Zhang. 2018 · 2018
Cited alongside, same era.
Simple random search of static linear policies is competitive for reinforcement learning. In The 32nd Conference on Neural Information Processing Systems . 1805–1814
Horia Mania, Aurelia Guy, and Benjamin Recht. 2018 · 2018
Cited alongside, same era.
Flax: A neural network library and ecosystem for JAX
Jonathan Heek, Anselm Levskaya, Avital Oliver, Marvin Ritter, Bertrand Rondepierre, Andreas Steiner, and Marc van Zee. 2020 · 2020
Later among the works it cites.
Optax: composable gradient transformation and optimisation, in JAX!
Matteo Hessel, David Budden, Fabio Viola, Mihaela Rosca, Eren Sezener, and Tom Hennigan. 2020 · 2020
Later among the works it cites.
Learning agile locomotion via adversarial training. In 2020 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 6098–6105
Yujin Tang, Jie Tan, and Tatsuya Harada. 2020b · 2020
Later among the works it cites.
ClipUp: A Simple and Powerful Optimizer for Distribution-Based Policy Evolution. In International Conference on Parallel Problem Solving from Nature . 515–527
Nihat Engin Toklu, Paweł Liskowski, and Rupesh Kumar Srivastava. 2020 · 2020
Later among the works it cites.
Brax - A Differentiable Physics Engine for Large Scale Rigid Body Simulation
C. Daniel Freeman, Erik Frey, Anton Raichuk, Sertan Girgin, Igor Mordatch, and Olivier Bachem. 2021 · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Felipe Such. 2018 · 2018
Cited alongside, same era.
Learning to Predict Without Looking Ahead: World Models Without Forward Prediction. In Advances in Neural Information Processing Systems , Vol. 32. Curran Associates, Inc
Daniel Freeman, David Ha, and Luke Metz. 2019 · 2019
Cited alongside, same era.
HARK Side of Deep Learning–From Grad Student Descent to Automated Machine Learning
Oguzhan Gencoglu, Mark van Gils, Esin Guldogan, Chamin Morikawa, Mehmet Süzen, Mathias Gruber, Jussi Leinonen, and Heikki Huttunen. 2019 · 2019
Cited alongside, same era.
Deep neuroevolution of recurrent and discrete world models. In Proceedings of GECCO . 456–462
Sebastian Risi and Kenneth O Stanley. 2019 · 2019
Cited alongside, same era.
Rui Wang, Joel Lehman, Jeff Clune, and Kenneth O Stanley. 2019 · 2019
Cited alongside, same era.
How does learning rate decay help modern neural networks?
Kaichao You, Mingsheng Long, Jianmin Wang, and Michael I Jordan. 2019 · 2019
Cited alongside, same era.
Slime Volleyball Gym Environment
David Ha. 2020 · 2020
Cited alongside, same era.
Neuroevolution of Self-Interpretable Agents. In Genetic and Evolutionary Computation Conference
Yujin Tang, Duong Nguyen, and David Ha. 2020a
Cited in the paper.
Later among the works it cites.
Collective Intelligence for Deep Learning: A Survey of Recent Developments
David Ha and Yujin Tang. 2021 · 2021
Later among the works it cites.
The hardware lottery
Sara Hooker. 2021 · 2021
Later among the works it cites.
Gradients are Not All You Need
Luke Metz, C Daniel Freeman, Samuel S Schoenholz, and Tal Kachman. 2021 · 2021
Later among the works it cites.
The Future of Artificial Intelligence is Self-Organizing and Self-Assembling
Sebastian Risi. 2021 · 2021
Later among the works it cites.
The Sensory Neuron as a Transformer: Permutation-Invariant Neural Networks for Reinforcement Learning. In The 35th Conference on Neural Information Processing Systems
Yujin Tang and David Ha. 2021 · 2021
Later among the works it cites.
Modern Evolution Strategies for Creativity: Fitting Concrete Images and Abstract Concepts
Yingtao Tian and David Ha. 2021 · 2021
Later among the works it cites.