Fetching the paper…
Reading the bibliography…
Deep reinforcement learning approaches have shown impressive results in a variety of different domains, however, more complex heterogeneous architectures such as world models require the different neural components to be trained separately instead of end-to-end.
Weight Agnostic Neural Networks
Gaier, A.; and Ha, D. 2019 · 1906
Earlier work this paper cites.
Steps toward artificial intelligence
Minsky, M. 1961 · 1961
Earlier work this paper cites.
An on-line algorithm for dynamic reinforcement learning and planning in reactive environments
Schmidhuber, J. 1990 · 1990
Earlier work this paper cites.
Long short-term memory
Hochreiter, S.; and Schmidhuber, J. 1997 · 1997
Earlier work this paper cites.
A study on overfitting in deep reinforcement learning
Zhang, C.; Vinyals, O.; Munos, R.; and Bengio, S. 2018 · 2001
Earlier work this paper cites.
A fast and elitist multiobjective genetic algorithm: NSGA-II
Deb, K.; Pratap, A.; Agarwal, S.; and Meyarivan, T. 2002 · 2002
Earlier work this paper cites.
Evolving neural networks through augmenting topologies
Stanley, K. O.; and Miikkulainen, R. 2002 · 2002
Earlier work this paper cites.
ALPS: the age-layered population structure for reducing the problem of premature convergence
Hornby, G. S. 2006 · 2006
Earlier work this paper cites.
Neuroevolution: from architectures to learning
Floreano, D.; Dürr, P.; and Mattiussi, C. 2008 · 2008
Earlier work this paper cites.
Exploiting open-endedness to solve problems through the search for novelty
Lehman, J.; and Stanley, K. O. 2008 · 2008
Earlier work this paper cites.
Visualizing data using t-SNE
Maaten, L. v. d.; and Hinton, G. 2008 · 2008
Earlier work this paper cites.
PILCO: A model-based and data-efficient approach to policy search
Deisenroth, M.; and Rasmussen, C. E. 2011 · 2011
Earlier work this paper cites.
Age-fitness pareto optimization
Schmidt, M.; and Lipson, H. 2011 · 2011
Earlier work this paper cites.
Encouraging behavioral diversity in evolutionary robotics: An empirical study
Mouret, J.-B.; and Doncieux, S. 2012 · 2012
Earlier work this paper cites.
An enhanced hypercube-based encoding for evolving the placement, density, and connectivity of neurons
Risi, S.; and Stanley, K. O. 2012 · 2012
Earlier work this paper cites.
Evolving large-scale neural networks for vision-based reinforcement learning
Koutník, J.; Cuccu, G.; Schmidhuber, J.; and Gomez, F. 2013 · 2013
Earlier work this paper cites.
Delving deep into rectifiers: Surpassing human-level performance on imagenet classification
He, K.; Zhang, X.; Ren, S.; and Sun, J. 2015 · 2015
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Mnih, V.; Kavukcuoglu, K.; Silver, D.; Rusu, A. A.; Veness, J.; Bellemare, M. G.; Graves, A.; Riedmiller, M.; Fidjeland, A. K.; Ostrovski, G.; et al. 2015 · 2015
Cited alongside, same era.
Deep learning in neural networks: An overview
Schmidhuber, J. 2015 · 2015
Cited alongside, same era.
Trust region policy optimization
Schulman, J.; Levine, S.; Abbeel, P.; Jordan, M.; and Moritz, P. 2015 · 2015
Cited alongside, same era.
From pixels to torques: Policy learning with deep dynamical models
Wahlström, N.; Schön, T. B.; and Deisenroth, M. P. 2015 · 2015
Cited alongside, same era.
Embed to control: A locally linear latent dynamics model for control from raw images
Watter, M.; Springenberg, J.; Boedecker, J.; and Riedmiller, M. 2015 · 2015
Cited alongside, same era.
Svcca: Singular vector canonical correlation analysis for deep learning dynamics and interpretability
Raghu, M.; Gilmer, J.; Yosinski, J.; and Sohl-Dickstein, J. 2017 · 2017
Later among the works it cites.
Neuroevolution in games: State of the art and open challenges
Risi, S.; and Togelius, J. 2017 · 2017
Later among the works it cites.
Evolution strategies as a scalable alternative to reinforcement learning
Salimans, T.; Ho, J.; Chen, X.; Sidor, S.; and Sutskever, I. 2017 · 2017
Later among the works it cites.
Proximal policy optimization algorithms
Schulman, J.; Wolski, F.; Dhariwal, P.; Radford, A.; and Klimov, O. 2017 · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kempka, M.; Wydmuch, M.; Runc, G.; Toczek, J.; and Jaśkowski, W. 2016 · 2016
Cited alongside, same era.
Carracing-v0
Klimov, O. 2016 · 2016
Cited alongside, same era.
Neuromodulation improves the evolution of forward models
Norouzzadeh, M. S.; and Clune, J. 2016 · 2016
Cited alongside, same era.
Quality diversity: A new frontier for evolutionary computation
Pugh, J. K.; Soros, L. B.; and Stanley, K. O. 2016 · 2016
Cited alongside, same era.
Autoencoder-augmented neuroevolution for visual Doom playing
Alvernaz, S.; and Togelius, J. 2017 · 2017
Cited alongside, same era.
Black-box data-efficient policy search for robotics
Chatzilygeroudis, K.; Rama, R.; Kaushik, R.; Goepp, D.; Vassiliades, V.; and Mouret, J.-B. 2017 · 2017
Cited alongside, same era.
Scalable co-optimization of morphology and control in embodied machines
Cheney, N.; Bongard, J.; SunSpiral, V.; and Lipson, H. 2018 · 2017
Cited alongside, same era.
Such, F. P.; Madhavan, V.; Conti, E.; Lehman, J.; Stanley, K. O.; and Clune, J. 2017 · 2017
Later among the works it cites.
Quantifying generalization in reinforcement learning
Cobbe, K.; Klimov, O.; Hesse, C.; Kim, T.; and Schulman, J. 2018 · 2018
Later among the works it cites.
Recurrent world models facilitate policy evolution
Ha, D.; and Schmidhuber, J. 2018 · 2018
Later among the works it cites.
Learning latent dynamics for planning from pixels
Hafner, D.; Lillicrap, T.; Fischer, I.; Villegas, R.; Ha, D.; Lee, H.; and Davidson, J. 2018 · 2018
Later among the works it cites.
Illuminating Generalization in Deep Reinforcement Learning through Procedural Level Generation
Justesen, N.; Torrado, R. R.; Bontrager, P.; Khalifa, A.; Togelius, J.; and Risi, S. 2018 · 2018
Later among the works it cites.
Safe mutations for deep and recurrent neural networks through output gradients
Lehman, J.; Chen, J.; Clune, J.; and Stanley, K. O. 2018 · 2018
Later among the works it cites.
An Atari model zoo for analyzing, visualizing, and comparing deep reinforcement learning agents
Such, F. P.; Madhavan, V.; Liu, R.; Wang, R.; Castro, P. S.; Li, Y.; Schubert, L.; Bellemare, M.; Clune, J.; and Lehman, J. 2018 · 2018
Later among the works it cites.
Unsupervised Predictive Memory in a Goal-Directed Agent
Wayne, G.; Hung, C.-C.; Amos, D.; Mirza, M.; Ahuja, A.; Grabska-Barwinska, A.; Rae, J.; Mirowski, P.; Leibo, J. Z.; Santoro, A.; et al. 2018 · 2018
Later among the works it cites.
Deep Neuroevolution of Recurrent and Discrete World Models
Risi, S.; and Stanley, K. O. 2019 · 2019
Closest in time.
Designing neural networks through neuroevolution
Stanley, K. O.; Clune, J.; Lehman, J.; and Miikkulainen, R. 2019 · 2019
Closest in time.
Increasing generality in machine learning through procedural content generation
Risi, S.; and Togelius, J. 2020 · 2020
Closest in time.
Neuroevolution of self-interpretable agents
Tang, Y.; Nguyen, D.; and Ha, D. 2020 · 2020
Closest in time.