Neural architecture search with reinforcement learning
Original
Barret Zoph and Quoc V Le · 2016
Later among the works it cites.
Accelerating neural architecture search using performance prediction
Original
Bowen Baker, Otkrist Gupta, Ramesh Raskar, and Nikhil Naik · 2017
Later among the works it cites.
Emergent complexity via multi-agent competition
Original
Trapit Bansal, Jakub Pachocki, Szymon Sidor, Ilya Sutskever, and Igor Mordatch · 2017
Later among the works it cites.
Building machines that learn and think for themselves: Commentary on lake et al., behavioral and brain sciences, 2017
Original
Matthew Botvinick, David GT Barrett, Peter Battaglia, Nando de Freitas, Dharshan Kumaran, Joel Z Leibo, Tim Lillicrap, Joseph Modayil, S Mohamed, Neil C Rabinowitz, et al · 2017
Later among the works it cites.
Minimal criterion coevolution: a new approach to open-ended search
Jonathan C Brant and Kenneth O Stanley · 2017
Later among the works it cites.
Meta-learning and universality: Deep representations and gradient descent can approximate any learning algorithm
Original
Chelsea Finn and Sergey Levine · 2017
Later among the works it cites.
Model-agnostic meta-learning for fast adaptation of deep networks
Chelsea Finn, Pieter Abbeel, and Sergey Levine · 2017
Later among the works it cites.
Noisy networks for exploration
Original
Meire Fortunato, Mohammad Gheshlaghi Azar, Bilal Piot, Jacob Menick, Ian Osband, Alex Graves, Vlad Mnih, Remi Munos, Demis Hassabis, Olivier Pietquin, et al · 2017
Later among the works it cites.
Visualizing and understanding atari agents
Original
Sam Greydanus, Anurag Koul, Jonathan Dodge, and Alan Fern · 2017
Later among the works it cites.
Rainbow: Combining improvements in deep reinforcement learning
Original
Matteo Hessel, Joseph Modayil, Hado Van Hasselt, Tom Schaul, Georg Ostrovski, Will Dabney, Dan Horgan, Bilal Piot, Mohammad Azar, and David Silver · 2017
Later among the works it cites.
Densely connected convolutional networks
Gao Huang, Zhuang Liu, Laurens Van Der Maaten, and Kilian Q Weinberger · 2017
Later among the works it cites.
Population based training of neural networks
Original
Max Jaderberg, Valentin Dalibard, Simon Osindero, Wojciech M Czarnecki, Jeff Donahue, Ali Razavi, Oriol Vinyals, Tim Green, Iain Dunning, Karen Simonyan, et al · 2017
Later among the works it cites.
Overcoming catastrophic forgetting in neural networks
James Kirkpatrick, Razvan Pascanu, Neil Rabinowitz, Joel Veness, Guillaume Desjardins, Andrei A Rusu, Kieran Milan, John Quan, Tiago Ramalho, Agnieszka Grabska-Barwinska, et al · 2017
Later among the works it cites.
Learning curve prediction with bayesian neural networks
Aaron Klein, Stefan Falkner, Jost Tobias Springenberg, and Frank Hutter · 2017
Later among the works it cites.
How evolution learns to generalise: Using the principles of learning theory to understand the evolution of developmental organisation
Kostas Kouvaris, Jeff Clune, Louis Kounios, Markus Brede, and Richard A Watson · 2017
Later among the works it cites.
Evolving deep neural networks
Original
Risto Miikkulainen, Jason Liang, Elliot Meyerson, Aditya Rawal, Dan Fink, Olivier Francon, Bala Raju, Hormoz Shahrzad, Arshak Navruzyan, Nigel Duffy, and Babak Hodjat · 2017
Later among the works it cites.
Plug & play generative networks: Conditional iterative generation of images in latent space
Anh Nguyen, Jeff Clune, Yoshua Bengio, Alexey Dosovitskiy, and Jason Yosinski · 2017
Later among the works it cites.
Curiosity-driven exploration by self-supervised prediction
Deepak Pathak, Pulkit Agrawal, Alexei A Efros, and Trevor Darrell · 2017
Later among the works it cites.
Parameter space noise for exploration
Original
Matthias Plappert, Rein Houthooft, Prafulla Dhariwal, Szymon Sidor, Richard Y Chen, Xi Chen, Tamim Asfour, Pieter Abbeel, and Marcin Andrychowicz · 2017
Later among the works it cites.
Optimization as a model for few-shot learning
Sachin Ravi and Hugo Larochelle · 2017
Later among the works it cites.
Dynamic routing between capsules
Sara Sabour, Nicholas Frosst, and Geoffrey E Hinton · 2017
Later among the works it cites.
Evolution strategies as a scalable alternative to reinforcement learning
Original
Tim Salimans, Jonathan Ho, Xi Chen, Szymon Sidor, and Ilya Sutskever · 2017
Later among the works it cites.
Mastering chess and shogi by self-play with a general reinforcement learning algorithm
Original
David Silver, Thomas Hubert, Julian Schrittwieser, Ioannis Antonoglou, Matthew Lai, Arthur Guez, Marc Lanctot, Laurent Sifre, Dharshan Kumaran, Thore Graepel, et al · 2017
Later among the works it cites.
Third-person imitation learning
Original
Bradly C Stadie, Pieter Abbeel, and Ilya Sutskever · 2017
Later among the works it cites.
Open-endedness: The last grand challenge you’ve never heard of
KO Stanley, J Lehman, and L Soros · 2017
Later among the works it cites.
Deep neuroevolution: Genetic algorithms are a competitive alternative for training deep neural networks for reinforcement learning
Original
Felipe Petroski Such, Vashisht Madhavan, Edoardo Conti, Joel Lehman, Kenneth O Stanley, and Jeff Clune · 2017
Later among the works it cites.
Inception-v4, inception-resnet and the impact of residual connections on learning
Christian Szegedy, Sergey Ioffe, and Vincent Vanhoucke · 2017
Later among the works it cites.
Diffusion-based neuromodulation can eliminate catastrophic forgetting in simple neural networks
Roby Velez and Jeff Clune · 2017
Later among the works it cites.
Feudal networks for hierarchical reinforcement learning
Alexander Sasha Vezhnevets, Simon Osindero, Tom Schaul, Nicolas Heess, Max Jaderberg, David Silver, and Koray Kavukcuoglu · 2017
Later among the works it cites.
Continual learning through synaptic intelligence
Friedemann Zenke, Ben Poole, and Surya Ganguli · 2017
Later among the works it cites.
Playing hard exploration games by watching youtube
Yusuf Aytar, Tobias Pfaff, David Budden, Thomas Paine, Ziyu Wang, and Nando de Freitas · 2018
Later among the works it cites.
Learning from demonstration in the wild
Original
Feryal Behbahani, Kyriacos Shiarlis, Xi Chen, Vitaly Kurin, Sudhanshu Kasewa, Ciprian Stirbu, João Gomes, Supratik Paul, Frans A Oliehoek, João Messias, et al · 2018
Later among the works it cites.
Large-scale study of curiosity-driven learning
Original
Yuri Burda, Harri Edwards, Deepak Pathak, Amos Storkey, Trevor Darrell, and Alexei A Efros · 2018
Later among the works it cites.
Improving exploration in evolution strategies for deep reinforcement learning via a population of novelty-seeking agents
Edoardo Conti, Vashisht Madhavan, Felipe Petroski Such, Joel Lehman, Kenneth Stanley, and Jeff Clune · 2018
Later among the works it cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Original
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2018
Later among the works it cites.
Agi safety literature review
Original
Tom Everitt, Gary Lea, and Marcus Hutter · 2018
Later among the works it cites.
Diversity is all you need: Learning skills without a reward function
Original
Benjamin Eysenbach, Abhishek Gupta, Julian Ibarz, and Sergey Levine · 2018
Later among the works it cites.
Unsupervised meta-learning for reinforcement learning
Original
Abhishek Gupta, Benjamin Eysenbach, Chelsea Finn, and Sergey Levine · 2018
Later among the works it cites.
Reinforcement learning for improving agent design
Original
David Ha · 2018
Later among the works it cites.
Learning an embedding space for transferable robot skills
Karol Hausman, Jost Tobias Springenberg, Ziyu Wang, Nicolas Heess, and Martin Riedmiller · 2018
Later among the works it cites.
Evolving multimodal robot behavior via many stepping stones with the combinatorial multi-objective evolutionary algorithm
Original
Joost Huizinga and Jeff Clune · 2018
Later among the works it cites.
The emergence of canalization and evolvability in an open-ended, interactive evolutionary system
Joost Huizinga, Kenneth O Stanley, and Jeff Clune · 2018
Later among the works it cites.
Human-level performance in first-person multiplayer games with population-based deep reinforcement learning
Original
Max Jaderberg, Wojciech M Czarnecki, Iain Dunning, Luke Marris, Guy Lever, Antonio Garcia Castaneda, Charles Beattie, Neil C Rabinowitz, Ari S Morcos, Avraham Ruderman, et al · 2018
Later among the works it cites.
Illuminating generalization in deep reinforcement learning through procedural level generation
Original
Niels Justesen, Ruben Rodriguez Torrado, Philip Bontrager, Ahmed Khalifa, Julian Togelius, and Sebastian Risi · 2018
Later among the works it cites.
Neural architecture search with bayesian optimisation and optimal transport
Kirthevasan Kandasamy, Willie Neiswanger, Jeff Schneider, Barnabas Poczos, and Eric P Xing · 2018
Later among the works it cites.
Progressive neural architecture search
Chenxi Liu, Barret Zoph, Maxim Neumann, Jonathon Shlens, Wei Hua, Li-Jia Li, Li Fei-Fei, Alan Yuille, Jonathan Huang, and Kevin Murphy · 2018
Later among the works it cites.
Differentiable plasticity: training plastic neural networks with backpropagation
Thomas Miconi, Jeff Clune, and Kenneth O Stanley · 2018
Later among the works it cites.
From nodes to networks: Evolving recurrent neural networks
Original
Aditya Rawal and Risto Miikkulainen · 2018
Later among the works it cites.
Learning montezuma’s revenge from a single demonstration
Original
Tim Salimans and Richard Chen · 2018
Later among the works it cites.
Episodic curiosity through reachability
Original
Nikolay Savinov, Anton Raichuk, Raphaël Marinier, Damien Vincent, Marc Pollefeys, Timothy Lillicrap, and Sylvain Gelly · 2018
Later among the works it cites.
Born to learn: the inspiration, progress, and future of evolved plastic artificial neural networks
Andrea Soltoggio, Kenneth O Stanley, and Sebastian Risi · 2018
Later among the works it cites.
Deep curiosity search: Intra-life exploration improves performance on challenging deep reinforcement problems
Christopher Stanton and Jeff Clune · 2018
Later among the works it cites.
Behavioral cloning from observation
Original
Faraz Torabi, Garrett Warnell, and Peter Stone · 2018
Later among the works it cites.
Maximum individual complexity is indefinitely scalable in geb
Alastair Channon · 2019
Closest in time.
Go-explore: a new approach for hard-exploration problems
Original
Adrien Ecoffet, Joost Huizinga, Joel Lehman, Kenneth O Stanley, and Jeff Clune · 2019
Closest in time.
Imitating latent policies from observation
Ashley D Edwards, Himanshu Sahni, Yannick Schroeker, and Charles L Isbell · 2019
Closest in time.
Backpropamine: training self-modifying neural networks with differentiable neuromodulated plasticity
Thomas Miconi, Aditya Rawal, Jeff Clune, and Kenneth O Stanley · 2019
Closest in time.
Generative teaching networks: learning to teach by generating synthetic training data
Felipe Petroski Such, Aditya Rawal, Joel Lehman, Kenneth O Stanley, and Jeff Clune · 2019
Closest in time.
Evolving images for visual neurons using a deep generative network reveals coding principles and neuronal preferences
Carlos R. Ponce, Will Xiao, Peter F. Schade, Till S. Hartmann, Gabriel Kreiman, and Margaret S. Livingstone · 2019
Closest in time.
Designing neural networks through neuroevolution
Kenneth O Stanley, Jeff Clune, Joel Lehman, and Risto Miikkulainen · 2019
Closest in time.
The bitter lesson, 2019
Rich Sutton · 2019
Closest in time.
Paired open-ended trailblazer (poet): Endlessly generating increasingly complex and diverse learning environments and their solutions
Original
Rui Wang, Joel Lehman, Jeff Clune, and Kenneth O Stanley · 2019
Closest in time.
Learning to continually learn
S. Beaulieu, L. Frati, T. Miconi, J. Lehman, K. Stanley, J. Clune, and N. Cheney · 2020
Closest in time.