Evaluation in artificial intelligence: from task-oriented to ability-oriented measurement
José Hernández-Orallo · 2017
Later among the works it cites.
The Measure of All Minds: Evaluating Natural and Artificial Intelligence
José Hernández-Orallo · 2017
Later among the works it cites.
Measuring the tendency of cnns to learn surface statistical regularities
Jason Jo and Yoshua Bengio · 2017
Later among the works it cites.
Exploring generalization in deep learning
Behnam Neyshabur, Srinadh Bhojanapalli, David McAllester, and Nati Srebro · 2017
Later among the works it cites.
Towards generalization and simplicity in continuous control
Lowrey K. Todorov E. V. Rajeswaran, A. and S. M. Kakade · 2017
Later among the works it cites.
Mastering chess and shogi by self-play with a general reinforcement learning algorithm
Original
David Silver, Thomas Hubert, Julian Schrittwieser, Ioannis Antonoglou, Matthew Lai, Arthur Guez, Marc Lanctot, Laurent Sifre, Dharshan Kumaran, Thore Graepel, et al · 2017
Later among the works it cites.
Mastering the game of go without human knowledge
David Silver, Julian Schrittwieser, Karen Simonyan, Ioannis Antonoglou, Aja Huang, Arthur Guez, Thomas Hubert, Lucas Baker, Matthew Lai, Adrian Bolton, et al · 2017
Later among the works it cites.
Starcraft ii: A new challenge for reinforcement learning
Original
Oriol Vinyals, Timo Ewalds, Sergey Bartunov, Petko Georgiev, Alexander Sasha Vezhnevets, Michelle Yeo, Alireza Makhzani, Heinrich Küttler, John Agapiou, Julian Schrittwieser, John Quan, Stephen Gaffney, Stig Petersen, Karen Simonyan, Tom Schaul, Hado van Hasselt, David Silver, Timothy P. Lillicrap, Kevin Calderone, Paul Keet, Anthony Brunasso, David Lawrence, Anders Ekermo, Jacob Repp, and Rodney Tsing · 2017
Later among the works it cites.
Understanding deep learning requires rethinking generalization
Chiyuan Zhang, Samy Bengio, Moritz Hardt, Benjamin Recht, and Oriol Vinyals · 2017
Later among the works it cites.
Chauffeurnet: Learning to drive by imitating the best and synthesizing the worst
Original
Mayank Bansal, Alex Krizhevsky, and Abhijit Ogale · 2018
Later among the works it cites.
Sample-efficient reinforcement learning with stochastic ensemble value expansion, 2018
Jacob Buckman, Danijar Hafner, George Tucker, Eugene Brevdo, and Honglak Lee · 2018
Later among the works it cites.
Quantifying generalization in reinforcement learning
Karl Cobbe, Oleg Klimov, Christopher Hesse, Taehoon Kim, and John Schulman · 2018
Later among the works it cites.
Deep reinforcement learning that matters
Islam R. Bachman P. Pineau J. Precup D. Henderson, P. and D. Meger · 2018
Later among the works it cites.
Predicting the generalization gap in deep networks with margin distributions
Yiding Jiang, Dilip Krishnan, Hossein Mobahi, and Samy Bengio · 2018
Later among the works it cites.
Illuminating generalization in deep reinforcement learning through procedural level generation
Original
Niels Justesen, Ruben Rodriguez Torrado, Philip Bontrager, Ahmed Khalifa, Julian Togelius, and Sebastian Risi · 2018
Later among the works it cites.
Deep learning: A critical appraisal
Original
Gary Marcus · 2018
Later among the works it cites.
Assessing generalization in deep reinforcement learning
Charles Packer, Katelyn Gao, Jernej Kos, Philipp Krähenbühl, Vladlen Koltun, and Dawn Xiaodong Song · 2018
Later among the works it cites.
General video game ai: a multi-track framework for evaluating agents, games and content generation algorithms
Original
Diego Perez-Liebana, Jialin Liu, Ahmed Khalifa, Raluca D Gaina, Julian Togelius, and Simon M Lucas · 2018
Later among the works it cites.
Reproducible, Reusable, and Robust Reinforcement Learning
Joelle Pineau · 2018
Later among the works it cites.
Reinforcement Learning: An Introduction (Second Edition)
Richard S. Sutton and Andrew G. Barto · 2018
Later among the works it cites.
Glue: A multi-task benchmark and analysis platform for natural language understanding
Alex Wang, Amanpreet Singh, Julian Michael, Felix Hill, Omer Levy, and Samuel R Bowman · 2018
Later among the works it cites.
A dissection of overfitting and generalization in continuous reinforcement learning
Original
Amy Zhang, Nicolas Ballas, and Joelle Pineau · 2018
Later among the works it cites.
The animal-ai environment: Training and testing animal-like artificial cognition, 2019
Benjamin Beyret, José Hernández-Orallo, Lucy Cheke, Marta Halina, Murray Shanahan, and Matthew Crosby · 2019
Closest in time.
The minerl competition on sample efficient reinforcement learning using human priors
William H. Guss, Cayden Codel, Katja Hofmann, Brandon Houghton, Noburu Kuno, Stephanie Milani, Sharada Prasanna Mohanty, Diego Perez Liebana, Ruslan Salakhutdinov, Nicholay Topin, Manuela Veloso, and Phillip Wang · 2019
Closest in time.
Obstacle tower: A generalization challenge in vision, control, and planning
Arthur Juliani, Ahmed Khalifa, Vincent-Pierre Berges, Jonathan Harper, Ervin Teng, Hunter Henry, Adam Crespi, Julian Togelius, and Danny Lange · 2019
Closest in time.
Behaviour suite for reinforcement learning
Original
Ian Osband, Yotam Doron, Matteo Hessel, John Aslanides, Eren Sezener, Andre Saraiva, Katrina McKinney, Tor Lattimore, Csaba Szepezvari, Satinder Singh, et al · 2019
Closest in time.
The multi-agent reinforcement learning in malmÖ (marlÖ) competition
Diego Perez-Liebana, Katja Hofmann, Sharada Prasanna Mohanty, Noboru Sean Kuno, Andre Kramer, Sam Devlin, Raluca D. Gaina, and Daniel Ionita · 2019
Closest in time.
OpenAI Five
OpenAI team · 2019
Closest in time.
OpenAI Five Arena Results
OpenAI team · 2019
Closest in time.
Superglue: A stickier benchmark for general-purpose language understanding systems
Alex Wang, Yada Pruksachatkun, Nikita Nangia, Amanpreet Singh, Julian Michael, Felix Hill, Omer Levy, and Samuel R Bowman · 2019
Closest in time.
Paired open-ended trailblazer (poet): Endlessly generating increasingly complex and diverse learning environments and their solutions
Original
Rui Wang, Joel Lehman, Jeff Clune, and Kenneth O. Stanley · 2019
Closest in time.