Fetching the paper…
Reading the bibliography…
Reinforcement learning has enjoyed multiple successes in recent years.
Brains, behavior
JS Albus · 1981
Earlier work this paper cites.
Learning to predict by the methods of temporal differences
Richard S Sutton · 1988
Earlier work this paper cites.
Q-learning
Christopher JCH Watkins and Peter Dayan · 1992
Earlier work this paper cites.
C4.5: Programs for Machine Learning
Ross Quinlan · 1993
Earlier work this paper cites.
On-line Q-learning using connectionist systems
Gavin A Rummery and Mahesan Niranjan · 1994
Earlier work this paper cites.
Reinforcement learning with replacing eligibility traces
Satinder P Singh and Richard S Sutton · 1996
Earlier work this paper cites.
Learning from demonstration
Stefan Schaal · 1997
Earlier work this paper cites.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto · 1998
Earlier work this paper cites.
Convergence of Q-learning: A simple proof
Francisco S Melo · 2001
Cited alongside, same era.
Data Mining: Practical Machine Learning Tools and Techniques
Ian H. Witten and Eibe Frank · 2005
Cited alongside, same era.
Pattern recognition
Christopher M Bishop · 2006
Cited alongside, same era.
Probabilistic policy reuse in a reinforcement learning agent
Fernando Fernández and Manuela Veloso · 2006
Cited alongside, same era.
Confidence-based policy learning from demonstration using gaussian mixture models
Sonia Chernova and Manuela Veloso · 2007
Cited alongside, same era.
Cross-domain transfer for reinforcement learning
Matthew E Taylor and Peter Stone · 2007
Cited alongside, same era.
Using spatial hints to improve policy reuse in a reinforcement learning agent
Bruno N Da Silva and Alan Mackworth · 2010
Later among the works it cites.
Integrating reinforcement learning with human demonstrations of varying ability
Matthew E. Taylor, Halit Bener Suay, and Sonia Chernova · 2011
Later among the works it cites.
The Mario AI benchmark and competitions
Sergey Karakovskiy and Julian Togelius · 2012
Later among the works it cites.
Policy transfer using reward shaping
Tim Brys, Anna Harutyunyan, Matthew E Taylor, and Ann Nowé · 2015
Later among the works it cites.
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba · 2016
Later among the works it cites.
Learning from demonstration for shaping through inverse reinforcement learning
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Brenna D Argall, Sonia Chernova, Manuela Veloso, and Brett Browning · 2009
Cited alongside, same era.
Transfer Learning for Reinforcement Learning Domains: A Survey
Matthew E. Taylor and Peter Stone · 2009
Cited alongside, same era.
Halit Bener Suay, Tim Brys, Matthew E Taylor, and Sonia Chernova · 2016
Later among the works it cites.
Multi-objectivization and ensembles of shapings in reinforcement learning
Tim Brys, Anna Harutyunyan, Peter Vrancx, Ann Nowé, and Matthew E Taylor · 2017
Later among the works it cites.
Improving Reinforcement Learning with Confidence-Based Demonstrations
Zhaodong Wang and Matthew E. Taylor · 2017
Later among the works it cites.