Fetching the paper…
Reading the bibliography…
Learning to navigate in complex environments with dynamic elements is an important milestone in developing AI agents.
Hippocampus, space, and memory
David S Olton, James T Becker, and Gail E Handelmann · 1979
Earlier work this paper cites.
Rule-injection hints as a means of improving network performance and learning time
Steven C Suddarth and YL Kergosien · 1990
Earlier work this paper cites.
Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning
Richard S Sutton, Doina Precup, and Satinder Singh · 1999
Earlier work this paper cites.
A solution to the simultaneous localization and map building (slam) problem
MWM Gamini Dissanayake, Paul Newman, Steve Clark, Hugh F. Durrant-Whyte, and Michael Csorba · 2001
Earlier work this paper cites.
Visualizing data using t-sne
Laurens van der Maaten and Geoffrey Hinton · 2008
Earlier work this paper cites.
Dynamic auto-encoders for semantic indexing
Piotr Mirowski, Marc’Aurelio Ranzato, and Yann LeCun · 2010
Earlier work this paper cites.
Lecture 6.5 – rmsprop: Divide the gradient by a running average of its recent magnitude
Tijmen Tieleman and Geoffrey Hinton · 2012
Earlier work this paper cites.
Speech recognition with deep recurrent neural networks
Alex Graves, Mohamed Abdelrahman, and Geoffrey Hinton · 2013
Earlier work this paper cites.
Evolving large-scale neural networks for vision-based reinforcement learning
Jan Koutnik, Giuseppe Cuccu, Jürgen Schmidhuber, and Faustino Gomez · 2013
Earlier work this paper cites.
How to construct deep recurrent neural networks
Razvan Pascanu, Caglar Gulcehre, Kyunghyun Cho, and Yoshua Bengio · 2013
Earlier work this paper cites.
Depth map prediction from a single image using a multi-scale deep network
David Eigen, Christian Puhrsch, and Rob Fergus · 2014
Earlier work this paper cites.
Jason Weston, Sumit Chopra, and Antoine Bordes · 2014
Cited alongside, same era.
Deep recurrent q-learning for partially observable mdps
Matthew J. Hausknecht and Peter Stone · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A. Rusu, Joel Veness, et al · 2015
Cited alongside, same era.
Massively parallel methods for deep reinforcement learning
Arun Nair, Praveen Srinivasan, Sam Blackwell, Cagdas Alcicek, Rory Fearon, et al · 2015
Cited alongside, same era.
Language understanding for text-based games using deep reinforcement learning
Karthik Narasimhan, Tejas D. Kulkarni, and Regina Barzilay · 2015
Cited alongside, same era.
Semi-supervised learning with ladder networks
Playing FPS games with deep reinforcement learning
Guillaume Lample and Devendra Singh Chaplot · 2016
Closest in time.
Recurrent reinforcement learning: A hybrid approach
Xiujun Li, Lihong Li, Jianfeng Gao, Xiaodong He, Jianshu Chen, Li Deng, and Ji He · 2016
Closest in time.
Asynchronous methods for deep reinforcement learning
Volodymyr Mnih, Adrià Puigdomènech Badia, Mehdi Mirza, Alex Graves, Timothy P. Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu · 2016
Closest in time.
Control of memory, active perception, and action in minecraft
Junhyuk Oh, Valliappa Chockalingam, Satinder P. Singh, and Honglak Lee · 2016
Closest in time.
Towards cognitive exploration through deep reinforcement learning for mobile robots
Lei Tai and Ming Liu · 2016
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Antti Rasmus, Mathias Berglund, Mikko Honkala, Harri Valpola, and Tapani Raiko · 2015
Cited alongside, same era.
Stacked what-where auto-encoders
Junbo Zhao, Michaël Mathieu, Ross Goroshin, and Yann LeCun · 2015
Cited alongside, same era.
Deep reinforcement learning in a 3-d blockworld environment
Trevor Barron, Matthew Whitehead, and Alan Yeung · 2016
Cited alongside, same era.
Charles Beattie, Joel Z. Leibo, Denis Teplyashin, Tom Ward, Marcus Wainwright, Heinrich Küttler, Andrew Lefrancq, Simon Green, Victor Valdes, Amir Sadik, Julian Schrittwieser, Keith Anderson, Sarah York, Max Cant, Adam Cain, Adrian Bolton, Stephen Gaffney, Helen King, Demis Hassabis, Shane Legg, and Stig Petersen · 2016
Cited alongside, same era.
Hybrid computing using a neural network with dynamic external memory
Alex Graves, Greg Wayne, Malcolm Reynolds, Tim Harley, Ivo Danihelka, Agnieszka Grabska-Barwińska, Sergio Gómez Colmenarejo, Edward Grefenstette, Tiago Ramalho, John Agapiou, et al · 2016
Cited alongside, same era.
Deep successor reinforcement learning
Tejas D. Kulkarni, Ardavan Saeedi, Simanta Gautam, and Samuel J. Gershman · 2016
Cited alongside, same era.
A deep hierarchical approach to lifelong learning in minecraft
Chen Tessler, Shahar Givony, Tom Zahavy, Daniel J. Mankowitz, and Shie Mannor · 2016
Closest in time.
Pixel recurrent neural networks
A. van den Oord, N. Kalchbrenner, and K. Kavukcuoglu · 2016
Closest in time.
Augmenting supervised neural networks with unsupervised objectives for large-scale image classification
Yuting Zhang, Kibok Lee, and Honglak Lee · 2016
Closest in time.
Target-driven visual navigation in indoor scenes using deep reinforcement learning
Yuke Zhu, Roozbeh Mottaghi, Eric Kolve, Joseph J. Lim, Abhinav Gupta, Li Fei-Fei, and Ali Farhadi · 2016
Closest in time.
Reinforcement learning with unsupervised auxiliary tasks
Max Jaderberg, Volodymir Mnih, Wojciech Czarnecki, Tom Schaul, Joel Z. Leibo, David Silver, and Koray Kavukcuoglu · 2017
Closest in time.