Fetching the paper…
Reading the bibliography…
Deep reinforcement learning (DRL) brings the power of deep neural networks to bear on the generic task of trial-and-error learning, and its effectiveness has been convincingly demonstrated on tasks such as Atari video games and the game of Go.
The second naive physics manifesto
Patrick J. Hayes · 1985
Earlier work this paper cites.
Generality in artificial intelligence
John McCarthy · 1987
Earlier work this paper cites.
The symbol grounding problem
Stevan Harnad · 1990
Earlier work this paper cites.
An analysis of first-order logics of probability
Joseph Y. Halpern · 1990
Earlier work this paper cites.
Representations of commonsense knowledge
Ernest Davis · 1990
Earlier work this paper cites.
Inductive logic programming
Stephen Muggleton · 1991
Earlier work this paper cites.
Default reasoning about spatial occupancy
Murray Shanahan · 1995
Earlier work this paper cites.
Reinforcement learning: An introduction
Richard S. Sutton and Andrew G. Barto · 1998
Earlier work this paper cites.
Relational reinforcement learning
Sašo Džeroski, Luc De Raedt, and Hendrik Blockeel · 1998
Earlier work this paper cites.
Perception as abduction: Turning sensor data into meaningful representation
Murray Shanahan · 2005
Earlier work this paper cites.
Universal intelligence: A definition of machine intelligence
Shane Legg and Marcus Hutter · 2007
Earlier work this paper cites.
Computational models of analogy
Dedre Gentner and Kenneth D. Forbus · 2011
Earlier work this paper cites.
Planning as satisfiability: heuristics
Jussi Rintanen · 2012
Earlier work this paper cites.
Auto-encoding variational bayes
Diederik P. Kingma and Max Welling · 2013
Cited alongside, same era.
Representation learning: A review and new perspectives
Yoshua Bengio, Aaron Courville, and Pascal Vincent · 2013
Cited alongside, same era.
Deep learning for real-time atari game play using offline monte-carlo tree search planning
Xiaoxiao Guo, Satinder Singh, Honglak Lee, Richard L. Lewis, and Xiaoshi Wang · 2014
Cited alongside, same era.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Cited alongside, same era.
Commonsense reasoning: An event calculus based approach
Erik T. Mueller · 2014
Cited alongside, same era.
Deep learning
Yann LeCun, Yoshua Bengio, and Geoffrey Hinton · 2015
Continuous deep q-learning with model-based acceleration
Shixiang Gu, Timothy P. Lillicrap, Ilya Sutskever, and Sergey Levine · 2016
Closest in time.
Classifying options for deep reinforcement learning
Kai Arulkumaran, Nat Dilokthanakul, Murray Shanahan, and Anil Anthony Bharath · 2016
Closest in time.
Successor features for transfer in reinforcement learning
André Barreto, Rémi Munos, Tom Schaul, and David Silver · 2016
Closest in time.
Andrei A. Rusu, Neil C. Rabinowitz, Guillaume Desjardins, Hubert Soyer, James Kirkpatrick, Koray Kavukcuoglu, Razvan Pascanu, and Raia Hadsell · 2016
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Deep learning in neural networks: An overview
Jürgen Schmidhuber · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A. Rusu, Joel Veness, Marc G. Bellemare, Alex Graves, Martin Riedmiller, Andreas K. Fidjeland, Georg Ostrovski, et al · 2015
Cited alongside, same era.
Data-efficient learning of feedback policies from image pixels using deep dynamical models
John-Alexander M. Assael, Niklas Wahlström, Thomas B. Schön, and Marc Peter Deisenroth · 2015
Cited alongside, same era.
Actor-mimic: Deep multitask and transfer reinforcement learning
Emilio Parisotto, Lei Jimmy Ba, and Ruslan Salakhutdinov · 2015
Cited alongside, same era.
End-to-end training of deep visuomotor policies
Sergey Levine, Chelsea Finn, Trevor Darrell, and Pieter Abbeel · 2016
Cited alongside, same era.
Mastering the game of go with deep neural networks and tree search
David Silver, Aja Huang, Chris J. Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al · 2016
Cited alongside, same era.
Alexander Vezhnevets, Volodymyr Mnih, John Agapiou, Simon Osindero, Alex Graves, Oriol Vinyals, and Koray Kavukcuoglu · 2016
Closest in time.
Graying the black box: Understanding dqns
Tom Zahavy, Nir Ben-Zrihem, and Shie Mannor · 2016
Closest in time.
Infogan: Interpretable representation learning by information maximizing generative adversarial nets
Xi Chen, Yan Duan, Rein Houthooft, John Schulman, Ilya Sutskever, and Pieter Abbeel · 2016
Closest in time.
Tagger: Deep unsupervised perceptual grouping
Klaus Greff, Antti Rasmus, Mathias Berglund, Tele Hotloo Hao, Jürgen Schmidhuber, and Harri Valpola · 2016
Closest in time.
Early visual concept learning with unsupervised deep learning
Irina Higgins, Loic Matthey, Xavier Glorot, Arka Pal, Benigno Uria, Charles Blundell, Shakir Mohamed, and Alexander Lerchner · 2016
Closest in time.
Building machines that learn and think like people
Brenden M. Lake, Tomer D. Ullman, Joshua B. Tenenbaum, and Samuel J. Gershman · 2016
Closest in time.
Guiying Li, Junlong Liu, Chunhui Jiang, and Ke Tang · 2016
Closest in time.
The frame problem
Murray Shanahan · 2016
Closest in time.