Fetching the paper…
Reading the bibliography…
Quality-Diversity (QD) algorithms, and MAP-Elites (ME) in particular, have proven very useful for a broad range of applications including enabling real robots to recover quickly from joint damage, solving strongly deceptive maze tasks or evolving robot morphologies to discover new gaits.
Evolutionsstrategie—Optimierung technischer Systeme nach Prinzipien der biologischen Information
Ingo Rechenberg. 1973 · 1973
Earlier work this paper cites.
A survey of evolution strategies. In Proceedings of the fourth international conference on genetic algorithms , Vol. 2. Morgan Kaufmann Publishers San Mateo, CA
Thomas Back, Frank Hoffmeister, and Hans-Paul Schwefel. 1991 · 1991
Earlier work this paper cites.
A possibility for implementing curiosity and boredom in model-building neural controllers. In Proc. of the international conference on simulation of adaptive behavior: From animals to animats . 222–227
Jürgen Schmidhuber. 1991 · 1991
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald J Williams. 1992 · 1992
Earlier work this paper cites.
Intrinsic motivation systems for autonomous mental development
Pierre-Yves Oudeyer, Frdric Kaplan, and Verena V Hafner. 2007 · 2007
Earlier work this paper cites.
Exploiting open-endedness to solve problems through the search for novelty.. In ALIFE . 329–336
Joel Lehman and Kenneth O Stanley. 2008 · 2008
Earlier work this paper cites.
Natural evolution strategies. In 2008 IEEE Congress on Evolutionary Computation (IEEE World Congress on Computational Intelligence) . IEEE, 3381–3387
Daan Wierstra, Tom Schaul, Jan Peters, and Juergen Schmidhuber. 2008 · 2008
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks. In Proceedings of the thirteenth international conference on artificial intelligence and statistics . 249–256
Xavier Glorot and Yoshua Bengio. 2010 · 2010
Earlier work this paper cites.
Parameter-exploring policy gradients
Frank Sehnke, Christian Osendorfer, Thomas Rückstieß, Alex Graves, Jan Peters, and Jürgen Schmidhuber. 2010 · 2010
Earlier work this paper cites.
Evolving a diversity of virtual creatures through novelty search and local competition. In Proceedings of the 13th annual conference on Genetic and evolutionary computation . ACM, 211–218
Joel Lehman and Kenneth O Stanley. 2011 · 2011
Earlier work this paper cites.
Active learning of inverse models with intrinsically motivated goal exploration in robots
Adrien Baranes and Pierre-Yves Oudeyer. 2013 · 2013
Earlier work this paper cites.
Intrinsic motivation and reinforcement learning
Andrew G Barto. 2013 · 2013
Earlier work this paper cites.
The arcade learning environment: An evaluation platform for general agents
Marc G Bellemare, Yavar Naddaf, Joel Veness, and Michael Bowling. 2013 · 2013
Earlier work this paper cites.
Playing atari with deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, and Martin Riedmiller. 2013 · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba. 2014 · 2014
Earlier work this paper cites.
Robots that can adapt like animals
Antoine Cully, Jeff Clune, Danesh Tarapore, and Jean-Baptiste Mouret. 2015 · 2015
Earlier work this paper cites.
Continuous control with deep reinforcement learning
Timothy P Lillicrap, Jonathan J Hunt, Alexander Pritzel, Nicolas Heess, Tom Erez, Yuval Tassa, David Silver, and Daan Wierstra. 2015 · 2015
Earlier work this paper cites.
Illuminating search spaces by mapping elites
Jean-Baptiste Mouret and Jeff Clune. 2015 · 2015
Earlier work this paper cites.
Innovation engines: Automated creativity and improved stochastic optimization via deep learning. In Proceedings of the 2015 Annual Conference on Genetic and Evolutionary Computation . 959–966
Anh Mai Nguyen, Jason Yosinski, and Jeff Clune. 2015 · 2015
Cited alongside, same era.
Confronting the challenge of quality diversity. In Proceedings of the 2015 Annual Conference on Genetic and Evolutionary Computation . 967–974
Justin K Pugh, Lisa B Soros, Paul A Szerlip, and Kenneth O Stanley. 2015 · 2015
Cited alongside, same era.
Unifying count-based exploration and intrinsic motivation. In Advances in neural information processing systems . 1471–1479
Marc Bellemare, Sriram Srinivasan, Georg Ostrovski, Tom Schaul, David Saxton, and Remi Munos. 2016 · 2016
Cited alongside, same era.
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba. 2016 · 2016
Cited alongside, same era.
Using centroidal voronoi tessellations to scale up the multidimensional archive of phenotypic elites algorithm
Vassilis Vassiliades, Konstantinos Chatzilygeroudis, and Jean-Baptiste Mouret. 2017 · 2017
Later among the works it cites.
Exploration by random network distillation
Yuri Burda, Harrison Edwards, Amos Storkey, and Oleg Klimov. 2018 · 2018
Later among the works it cites.
Reset-free trial-and-error learning for robot damage recovery
Konstantinos Chatzilygeroudis, Vassilis Vassiliades, and Jean-Baptiste Mouret. 2018 · 2018
Later among the works it cites.
Gep-pg: Decoupling exploration and exploitation in deep reinforcement learning algorithms
Cédric Colas, Olivier Sigaud, and Pierre-Yves Oudeyer. 2018 · 2018
Later among the works it cites.
Improving exploration in evolution strategies for deep reinforcement learning via a population of novelty-seeking agents. In Advances in Neural Information Processing Systems . 5027–5038
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Nikolaus Hansen. 2016 · 2016
Cited alongside, same era.
Vime: Variational information maximizing exploration. In Advances in Neural Information Processing Systems . 1109–1117
Rein Houthooft, Xi Chen, Yan Duan, John Schulman, Filip De Turck, and Pieter Abbeel. 2016 · 2016
Cited alongside, same era.
Creative generation of 3D objects with deep learning and innovation engines. In Proceedings of the 7th International Conference on Computational Creativity
Joel Lehman, Sebastian Risi, and Jeff Clune. 2016 · 2016
Cited alongside, same era.
Searching for quality diversity when diversity is unaligned with quality. In International Conference on Parallel Problem Solving from Nature . Springer, 880–889
Justin K Pugh, Lisa B Soros, and Kenneth O Stanley. 2016 · 2016
Cited alongside, same era.
Surprise-based intrinsic motivation for deep reinforcement learning
Joshua Achiam and Shankar Sastry. 2017 · 2017
Cited alongside, same era.
Quality and diversity optimization: A unifying modular framework
Antoine Cully and Yiannis Demiris. 2017 · 2017
Cited alongside, same era.
Model-agnostic meta-learning for fast adaptation of deep networks. In Proceedings of the 34th International Conference on Machine Learning-Volume 70 . JMLR. org, 1126–1135
Chelsea Finn, Pieter Abbeel, and Sergey Levine. 2017 · 2017
Cited alongside, same era.
Intrinsically motivated goal exploration processes with automatic curriculum learning
Sébastien Forestier, Yoan Mollard, and Pierre-Yves Oudeyer. 2017 · 2017
Cited alongside, same era.
Edoardo Conti, Vashisht Madhavan, Felipe Petroski Such, Joel Lehman, Kenneth Stanley, and Jeff Clune. 2018 · 2018
Later among the works it cites.
Addressing function approximation error in actor-critic methods
Scott Fujimoto, Herke Van Hoof, and David Meger. 2018 · 2018
Later among the works it cites.
Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Tuomas Haarnoja, Aurick Zhou, Pieter Abbeel, and Sergey Levine. 2018 · 2018
Later among the works it cites.
Gotta learn fast: A new benchmark for generalization in rl
Alex Nichol, Vicki Pfau, Christopher Hesse, Oleg Klimov, and John Schulman. 2018 · 2018
Later among the works it cites.
Model-based active exploration
Pranav Shyam, Wojciech Jaśkowski, and Faustino Gomez. 2018 · 2018
Later among the works it cites.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto. 2018 · 2018
Later among the works it cites.
MULEX: Disentangling Exploitation from Exploration in Deep RL
Lucas Beyer, Damien Vincent, Olivier Teboul, Sylvain Gelly, Matthieu Geist, and Olivier Pietquin. 2019 · 2019
Later among the works it cites.
Go-explore: a new approach for hard-exploration problems
Adrien Ecoffet, Joost Huizinga, Joel Lehman, Kenneth O Stanley, and Jeff Clune. 2019 · 2019
Later among the works it cites.
Covariance Matrix Adaptation for the Rapid Illumination of Behavior Space
Matthew C Fontaine, Julian Togelius, Stefanos Nikolaidis, and Amy K Hoover. 2019 · 2019
Later among the works it cites.
Evolvability ES: scalable and direct optimization of evolvability. In Proceedings of the Genetic and Evolutionary Computation Conference . 107–115
Alexander Gajewski, Jeff Clune, Kenneth O Stanley, and Joel Lehman. 2019 · 2019
Later among the works it cites.
Adversarial policies: Attacking deep reinforcement learning
Adam Gleave, Michael Dennis, Neel Kant, Cody Wild, Sergey Levine, and Stuart Russell. 2019 · 2019
Later among the works it cites.
Teacher algorithms for curriculum learning of Deep RL in continuously parameterized environments
Rémy Portelas, Cédric Colas, Katja Hofmann, and Pierre-Yves Oudeyer. 2019 · 2019
Later among the works it cites.
Scheduled intrinsic drive: A hierarchical take on intrinsically motivated exploration
Jingwei Zhang, Niklas Wetzel, Nicolai Dorka, Joschka Boedecker, and Wolfram Burgard. 2019 · 2019
Later among the works it cites.