Fetching the paper…
Reading the bibliography…
Sequential decision making problems, such as structured prediction, robotic control, and game playing, require a combination of planning policies and generalisation of those plans.
Heuristic and Analytic Processes in Reasoning
J. St B. T. Evans · 1984
Earlier work this paper cites.
Simple Statistical Gradient-Following Algorithms for Connectionist Reinforcement Learning
R. J. Williams · 1992
Earlier work this paper cites.
Maps of Bounded Rationality: Psychology for Behavioral Economics
Daniel Kahneman · 2003
Earlier work this paper cites.
General Game Playing: Overview of the AAAI Competition
M. Genesereth, N. Love, and B. Pell · 2005
Earlier work this paper cites.
Bayeselo
R Coulom · 2005
Earlier work this paper cites.
Bandit Based Monte-Carlo Planning
L. Kocsis and C. Szepesvári · 2006
Earlier work this paper cites.
Combining Online and Offline Knowledge in UCT
S. Gelly and D. Silver · 2007
Earlier work this paper cites.
Search-based Structured Prediction
H. Daumé III, J. Langford, and D. Marcu · 2009
Cited alongside, same era.
Monte Carlo Tree Search in Hex
B. Arneson, R. Hayward, and P. Hednerson · 2010
Cited alongside, same era.
A Reduction of Imitation Learning and Structured Prediction to No-Regret Online Learning
S. Ross, G. J. Gordon, and J. A. Bagnell · 2011
Cited alongside, same era.
Training Deterministic Parsers with Non-Deterministic Oracles
Y. Goldberg and J. Nivre · 2013
Cited alongside, same era.
Adam: A Method for Stochastic Optimization
D. Kingma and J. Ba · 2014
Cited alongside, same era.
Deep Learning for Real-Time Atari Game Play Using Offline Monte-Carlo Rree Search Planning
X. Guo, S. Singh, H. Lee, R. L. Lewis, and X. Wang · 2014
Human-Level Control through Deep Reinforcement Learning
V. Mnih et al · 2015
Later among the works it cites.
Learning to Search Better Than Your Teacher
K. Chang, A. Krishnamurthy, A. Agarwal, H. Daumé III, and J. Langford · 2015
Later among the works it cites.
Fast and Accurate Deep Network Learning by Exponential Linear Units(ELUs)
D.-A. Clevert, T. Unterthiner, and S. Hochreiter · 2015
Later among the works it cites.
Mastering the Game of Go with Deep Neural Networks and Tree Search
D. Silver et al · 2016
Later among the works it cites.
NeuroHex: A Deep Q-learning Hex Agent
K. Young, R. Hayward, and G. Vasan · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Reinforcement and Imitation Learning via Interactive No-Regret Learning
S. Ross and J. A. Bagnell · 2014
Cited alongside, same era.
D. Arpit, Y. Zhou, B. U. Kota, and V. Govindaraju · 2016
Later among the works it cites.
Mastering the Game of Go without Human Knowledge
D. Silver et al · 2017
Closest in time.