Illuminating generalization in deep reinforcement learning through procedural level generation
Original
Niels Justesen, Ruben Rodriguez Torrado, Philip Bontrager, Ahmed Khalifa, Julian Togelius, and Sebastian Risi · 2018
Later among the works it cites.
Super Mario Bros for OpenAI Gym
Christian Kauten · 2018
Later among the works it cites.
State representation learning for control: An overview
Timothée Lesort, Natalia Díaz Rodríguez, Jean-François Goudou, and David Filliat · 2018
Later among the works it cites.
Gotta learn fast: A new benchmark for generalization in rl
Original
Alex Nichol, Vicki Pfau, Christopher Hesse, Oleg Klimov, and John Schulman · 2018
Later among the works it cites.
The uncertainty bellman equation and exploration
Brendan O’Donoghue, Ian Osband, Rémi Munos, and Volodymyr Mnih · 2018
Later among the works it cites.
Assessing generalization in deep reinforcement learning
Original
Charles Packer, Katelyn Gao, Jernej Kos, Philipp Krähenbühl, Vladlen Koltun, and Dawn Song · 2018
Later among the works it cites.
Deep curiosity search: Intra-life exploration can improve performance on challenging deep reinforcement learning problems
Original
Christopher Stanton and Jeff Clune · 2018
Later among the works it cites.
Large-scale study of curiosity-driven learning
Yuri Burda, Harrison Edwards, Deepak Pathak, Amos J. Storkey, Trevor Darrell, and Alexei A. Efros · 2019
Later among the works it cites.
Exploration by random network distillation
Yuri Burda, Harrison Edwards, Amos J. Storkey, and Oleg Klimov · 2019
Later among the works it cites.
Contingency-aware exploration in reinforcement learning
Jongwook Choi, Yijie Guo, Marcin Moczulski, Junhyuk Oh, Neal Wu, Mohammad Norouzi, and Honglak Lee · 2019
Later among the works it cites.
Quantifying generalization in reinforcement learning
Karl Cobbe, Oleg Klimov, Christopher Hesse, Taehoon Kim, and John Schulman · 2019
Later among the works it cites.
Feature control as intrinsic motivation for hierarchical reinforcement learning
Nat Dilokthanakul, Christos Kaplanis, Nick Pawlowski, and Murray Shanahan · 2019
Later among the works it cites.
Go-explore: a new approach for hard-exploration problems
Original
Adrien Ecoffet, Joost Huizinga, Joel Lehman, Kenneth O Stanley, and Jeff Clune · 2019
Later among the works it cites.
Diversity is all you need: Learning skills without a reward function
Benjamin Eysenbach, Abhishek Gupta, Julian Ibarz, and Sergey Levine · 2019
Later among the works it cites.
Infobot: Transfer and exploration via the information bottleneck
Anirudh Goyal, Riashat Islam, Daniel Strouse, Zafarali Ahmed, Hugo Larochelle, Matthew Botvinick, Yoshua Bengio, and Sergey Levine · 2019
Later among the works it cites.
Human-level performance in 3d multiplayer games with population-based reinforcement learning
Max Jaderberg, Wojciech M Czarnecki, Iain Dunning, Luke Marris, Guy Lever, Antonio Garcia Castaneda, Charles Beattie, Neil C Rabinowitz, Ari S Morcos, Avraham Ruderman, et al · 2019
Later among the works it cites.
Obstacle tower: A generalization challenge in vision, control, and planning
Arthur Juliani, Ahmed Khalifa, Vincent-Pierre Berges, Jonathan Harper, Ervin Teng, Hunter Henry, Adam Crespi, Julian Togelius, and Danny Lange · 2019
Later among the works it cites.
TorchBeast: A PyTorch Platform for Distributed RL
Original
Heinrich Küttler, Nantas Nardelli, Thibaut Lavril, Marco Selvatici, Viswanath Sivakumar, Tim Rocktäschel, and Edward Grefenstette · 2019
Later among the works it cites.
Hierarchical RL using an ensemble of proprioceptive periodic policies
Kenneth Marino, Abhinav Gupta, Rob Fergus, and Arthur Szlam · 2019
Later among the works it cites.
Scheduled intrinsic drive: A hierarchical take on intrinsically motivated exploration
Original
Jingwei Zhang, Niklas Wetzel, Nicolai Dorka, Joschka Boedecker, and Wolfram Burgard · 2019
Later among the works it cites.