Fetching the paper…
Reading the bibliography…
In many reinforcement learning tasks, the goal is to learn a policy to manipulate an agent, whose design is fixed, to maximize some notion of cumulative reward.
Evolutionsstrategien
I. Rechenberg · 1978
Earlier work this paper cites.
Numerical optimization of computer models
H.-P. Schwefel · 1981
Earlier work this paper cites.
Passive walking with knees
T. McGeer · 1990
Earlier work this paper cites.
Exercise and mental health
J. S. Raglin · 1990
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
R. J. Williams · 1992
Earlier work this paper cites.
Evolving 3d morphology and behavior by competition
K. Sims · 1994
Earlier work this paper cites.
Evolving virtual creatures
K. Sims · 1994
Earlier work this paper cites.
The embodied mind: Cognitive science and human experience (book)
M. R. Gover · 1996
Earlier work this paper cites.
Long short-term memory
S. Hochreiter and J. Schmidhuber · 1997
Earlier work this paper cites.
Medicine and science in the life of luigi galvani (1737–1798)
M. Bresadola · 1998
Earlier work this paper cites.
Automated synthesis and optimization of robot configurations: an evolutionary approach
C. Leger et al · 1999
Earlier work this paper cites.
Automatic design and manufacture of robotic lifeforms
H. Lipson and J. B. Pollack · 2000
Earlier work this paper cites.
A three-dimensional passive-dynamic walking robot with two legs and knees
S. H. Collins, M. Wisse, and A. Ruina · 2001
Earlier work this paper cites.
Completely derandomized self-adaptation in evolution strategies
N. Hansen and A. Ostermeier · 2001
Earlier work this paper cites.
Evolving neural networks through augmenting topologies
K. O. Stanley and R. Miikkulainen · 2002
Earlier work this paper cites.
Six views of embodied cognition
M. Wilson · 2002
Earlier work this paper cites.
Embodied cognition: A field guide
M. L. Anderson · 2003
Earlier work this paper cites.
Evolving control for modular robotic units
E. H. Ostergaard and H. H. Lund · 2003
Earlier work this paper cites.
Morphology and computation
C. Paul · 2004
Earlier work this paper cites.
Efficient bipedal robots based on passive-dynamic walkers
S. Collins, A. Ruina, R. Tedrake, and M. Wisse · 2005
Earlier work this paper cites.
Short-term effects on lower-body functional power development: weightlifting vs. vertical jump training programs
V. Tricoli, L. Lamas, R. Carnevale, and C. Ugrinowitsch · 2005
Earlier work this paper cites.
Passive propulsion in vortex wakes
D. Beal, F. Hover, M. Triantafyllou, J. Liao, and G. Lauder · 2006
Earlier work this paper cites.
How the body shapes the way we think: a new view of intelligence
R. Pfeifer and J. Bongard · 2006
Earlier work this paper cites.
Evolving spatiotemporal coordination in a modular robotic system
M. Prokopenko, V. Gerasimov, and I. Tanev · 2006
Earlier work this paper cites.
Evolved and designed self-reproducing modular robotics
V. Zykov, E. Mytilinaios, M. Desnoyer, and H. Lipson · 2007
Earlier work this paper cites.
Strandbeests
T. Jansen · 2008
Cited alongside, same era.
A critical look at the embodied cognition hypothesis and a new proposal for grounding conceptual content
B. Z. Mahon and A. Caramazza · 2008
Cited alongside, same era.
An experiment in automatic game design
J. Togelius and J. Schmidhuber · 2008
Cited alongside, same era.
Natural evolution strategies
D. Wierstra, T. Schaul, J. Peters, and J. Schmidhuber · 2008
Cited alongside, same era.
Exercise and mental health: many reasons to move
A. Deslandes, H. Moraes, C. Ferreira, H. Veiga, H. Silveira, R. Mouta, F. A. Pompeu, E. S. F. Coutinho, and J. Laks · 2009
Cited alongside, same era.
Artificial intelligence for games
I. Millington and J. Funge · 2009
Cited alongside, same era.
Steps toward a modular library for turning any evolutionary domain into an online interactive platform
P. A. Szerlip and K. O. Stanley · 2014
Later among the works it cites.
Trust region policy optimization
J. Schulman, S. Levine, P. Abbeel, M. Jordan, and P. Moritz · 2015
Later among the works it cites.
G. Brockman, V. Cheung, L. Pettersson, J. Schneider, J. Schulman, J. Tang, and W. Zaremba · 2016
Later among the works it cites.
Experiments in handwriting with a neural network
S. Carter, D. Ha, I. Johnson, and C. Olah · 2016
Later among the works it cites.
BipedalWalker-v2, 2016
O. Klimov · 2016
Later among the works it cites.
BipedalWalkerHardcore-v2, 2016
O. Klimov · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. E. Auerbach and J. C. Bongard · 2010
Cited alongside, same era.
Evolving cppns to grow three-dimensional physical structures
J. E. Auerbach and J. C. Bongard · 2010
Cited alongside, same era.
Autonomous evolution of topographic regularities in artificial neural networks
J. Gauci and K. O. Stanley · 2010
Cited alongside, same era.
Parameter-exploring policy gradients
F. Sehnke, C. Osendorfer, T. Rückstieß, A. Graves, J. Peters, and J. Schmidhuber · 2010
Cited alongside, same era.
Embodied cognition
L. Shapiro · 2010
Cited alongside, same era.
BoxCar2D
R. Weber · 2010
Cited alongside, same era.
Asynchronous methods for deep reinforcement learning
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu · 2016
Later among the works it cites.
Neural architecture search with reinforcement learning
B. Zoph and Q. V. Le · 2016
Later among the works it cites.
Using artificial intelligence to augment human intelligence
S. Carter and M. Nielsen · 2017
Later among the works it cites.
PyBullet Physics Environment, 2017
E. Coumans · 2017
Later among the works it cites.
Evolving stable strategies, 2017
D. Ha · 2017
Later among the works it cites.
Joint optimization of robot design and motion parameters using the implicit function theorem
S. Ha, S. Coros, A. Alspach, J. Kim, and K. Yamane · 2017
Later among the works it cites.
Roboschool, May 2017
O. Klimov and J. Schulman · 2017
Later among the works it cites.
Evolution strategies as a scalable alternative to reinforcement learning
T. Salimans, J. Ho, X. Chen, S. Sidor, and I. Sutskever · 2017
Later among the works it cites.
Proximal policy optimization algorithms
J. Schulman, F. Wolski, P. Dhariwal, A. Radford, and O. Klimov · 2017
Later among the works it cites.
F. P. Such, V. Madhavan, E. Conti, J. Lehman, K. O. Stanley, and J. Clune · 2017
Later among the works it cites.
Evolving soft locomotion in aquatic and terrestrial environments: effects of material properties and environmental transitions
F. Corucci, N. Cheney, F. Giorgio-Serchi, J. Bongard, and C. Laschi · 2018
Closest in time.
RL A3C Pytorch Experiments, 2018
D. Griffis · 2018
Closest in time.
Co-creative level design via machine learning
M. Guzdial, N. Liao, and M. Riedl · 2018
Closest in time.
Computational co-optimization of design parameters and motion trajectories for robotic systems
S. Ha, S. Coros, A. Alspach, J. Kim, and K. Yamane · 2018
Closest in time.
J. Lehman, J. Clune, D. Misevic, C. Adami, J. Beaulieu, P. J. Bentley, S. Bernard, G. Belson, D. M. Bryson, N. Cheney, et al · 2018
Closest in time.
Jointly learning to construct and control agents using deep reinforcement learning
C. Schaff, D. Yunis, A. Chakrabarti, and M. R. Walter · 2018
Closest in time.
Jointly learning to construct and control agents using deep reinforcement learning
C. Schaff, D. Yunis, A. Chakrabarti, and M. R. Walter · 2018
Closest in time.
Procedural content generation via machine learning (pcgml)
A. Summerville, S. Snodgrass, M. Guzdial, C. Holmgard, A. K. Hoover, A. Isaksen, A. Nealen, and J. Togelius · 2018
Closest in time.
Evolving mario levels in the latent space of a deep convolutional generative adversarial network
V. Volz, J. Schrum, J. Liu, S. M. Lucas, A. Smith, and S. Risi · 2018
Closest in time.