Fetching the paper…
Reading the bibliography…
An evolution strategy (ES) variant based on a simplification of a natural evolution strategy recently attracted attention because it performs surprisingly well in challenging deep reinforcement learning domains.
The approximate arithmetical solution by finite differences of physical problems involving differential equations, with an application to the stresses in a masonry dam
Richardson, L. F. (1911) · 1911
Earlier work this paper cites.
Likelilood ratio gradient estimation: an overview
Glynn, P. W. (1987) · 1987
Earlier work this paper cites.
Multivariate stochastic approximation using a simultaneous perturbation gradient approximation
Spall, J. C. (1992) · 1992
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Williams, R. J. (1992) · 1992
Earlier work this paper cites.
Evolution and optimum seeking: the sixth generation
Schwefel, H.-P. P. (1993) · 1993
Earlier work this paper cites.
Perspective: complex adaptations and the evolution of evolvability
Wagner, G. P. and Altenberg, L. (1996) · 1996
Earlier work this paper cites.
Evolvability
Kirschner, M. and Gerhart, J. (1998) · 1998
Earlier work this paper cites.
Statistical machine learning and combinatorial optimization
Berny, A. (2000) · 2000
Earlier work this paper cites.
Evolution of digital organisms at high mutation rates leads to survival of the flattest
Wilke, C. O., nad Charles Ofria, J. L. W., Lenski, R. E., and Adami, C. (2001) · 2001
Earlier work this paper cites.
Evolution strategies: A comprehensive introduction
Beyer, H.-G. and Schwefel, H.-P. (2002) · 2002
Earlier work this paper cites.
Evolutionary Computation: A Unified Perspective
De Jong, K. A. (2002) · 2002
Earlier work this paper cites.
Balancing robustness and evolvability
Lenski, R. E., Barrick, J. E., and Ofria, C. (2006) · 2006
Cited alongside, same era.
Self-adaptation in evolutionary algorithms
Meyer-Nieberg, S. and Beyer, H.-G. (2007) · 2007
Cited alongside, same era.
Natural selection fails to optimize mutation rates for long-term adaptation on rugged fitness landscapes
Clune, J., Misevic, D., Ofria, C., Lenski, R. E., Elena, S. F., and Sanjuán, R. (2008) · 2008
Cited alongside, same era.
Robustness and evolvability: a paradox resolved
Wagner, A. (2008) · 2008
Cited alongside, same era.
Parameter-exploring policy gradients
Sehnke, F., Osendorfer, C., Rückstieß, T., Graves, A., Peters, J., and Schmidhuber, J. (2010) · 2010
Cited alongside, same era.
Improving evolvability through novelty search and self-adaptation
Lehman, J. and Stanley, K. O. (2011b) · 2011
Cited alongside, same era.
Trust region policy optimization
Schulman, J., Levine, S., Abbeel, P., Jordan, M., and Moritz, P. (2015) · 2015
Later among the works it cites.
OpenAI gym
Brockman, G., Cheung, V., Pettersson, L., Schneider, J., Schulman, J., Tang, J., and Zaremba, W. (2016) · 2016
Later among the works it cites.
Kounios, L., Clune, J., Kouvaris, K., Wagner, G. P., Pavlicev, M., Weinreich, D. M., and Watson, R. A. (2016) · 2016
Later among the works it cites.
Quality diversity: A new frontier for evolutionary computation
Pugh, J. K., Soros, L. B., and Stanley, K. O. (2016) · 2016
Later among the works it cites.
Conti, E., Madhavan, V., Petroski Such, F., Lehman, J., Stanley, K. O., and Clune, J. (2017) · 2017
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Bayesian learning via stochastic gradient langevin dynamics
Welling, M. and Teh, Y. W. (2011) · 2011
Cited alongside, same era.
Synthesis and stabilization of complex behaviors through online trajectory optimization
Tassa, Y., Erez, T., and Todorov, E. (2012) · 2012
Cited alongside, same era.
Mujoco: A physics engine for model-based control
Todorov, E., Erez, T., and Tassa, Y. (2012) · 2012
Cited alongside, same era.
Evolvability is inevitable: Increasing evolvability without the pressure to adapt
Lehman, J. and Stanley, K. O. (2013) · 2013
Cited alongside, same era.
Natural evolution strategies
Wierstra, D., Schaul, T., Glasmachers, T., Sun, Y., Peters, J., and Schmidhuber, J. (2014) · 2014
Cited alongside, same era.
Abandoning objectives: Evolution through the search for novelty alone
Lehman, J. and Stanley, K. O. (2011a)
Cited in the paper.
Openai baselines
Dhariwal, P., Hesse, C., Plappert, M., Radford, A., Schulman, J., Sidor, S., and Wu, Y. (2017) · 2017
Closest in time.
Gradient-free policy architecture search and adaptation
Ebrahimi, S., Rohrbach, A., and Darrell, T. (2017) · 2017
Closest in time.
Petroski Such, F., Madhavan, V., Conti, E., Lehman, J., Stanley, K. O., and Clune, J. (2017) · 2017
Closest in time.
Parameter space noise for exploration
Plappert, M., Houthooft, R., Dhariwal, P., Sidor, S., Chen, R. Y., Chen, X., Asfour, T., Abbeel, P., and Andrychowicz, M. (2017) · 2017
Closest in time.
Evolution strategies as a scalable alternative to reinforcement learning
Salimans, T., Ho, J., Chen, X., and Sutskever, I. (2017) · 2017
Closest in time.
On the relationship between the openai evolution strategy and stochastic gradient descent
Zhang, X., Clune, J., and Stanley, K. O. (2017) · 2017
Closest in time.