Fetching the paper…
Reading the bibliography…
We present the first massively distributed architecture for deep reinforcement learning.
Reinforcement learning for robots using neural networks
Lin, Long-Ji · 1993
Earlier work this paper cites.
Distributed reinforcement learning
Weiss, Gerhard · 1995
Earlier work this paper cites.
An analysis of temporal-difference learning with function approximation
Tsitsiklis, J. and Roy, B. Van · 1997
Earlier work this paper cites.
Reinforcement Learning: an Introduction
Sutton, R. and Barto, A · 1998
Earlier work this paper cites.
An algorithm for distributed reinforcement learning in cooperative multi-agent systems
Lauer, Martin and Riedmiller, Martin · 2000
Earlier work this paper cites.
Parallel reinforcement learning with linear function approximation
Grounds, Matthew and Kudenko, Daniel · 2008
Earlier work this paper cites.
Adaptive subgradient methods for online learning and stochastic optimization
Duchi, John, Hazan, Elad, and Singer, Yoram · 2011
Earlier work this paper cites.
Mapreduce for parallel reinforcement learning
Li, Yuxi and Schuurmans, Dale · 2011
Cited alongside, same era.
The arcade learning environment: An evaluation platform for general agents
Bellemare, Marc G, Naddaf, Yavar, Veness, Joel, and Bowling, Michael · 2012
Cited alongside, same era.
Context-dependent pre-trained deep neural networks for large-vocabulary speech recognition
Dahl, George E, Yu, Dong, Deng, Li, and Acero, Alex · 2012
Cited alongside, same era.
Large scale distributed deep networks
Dean, Jeffrey, Corrado, Greg, Monga, Rajat, Chen, Kai, Devin, Matthieu, Mao, Mark, Senior, Andrew, Tucker, Paul, Yang, Ke, Le, Quoc V, et al · 2012
Cited alongside, same era.
Imagenet classification with deep convolutional neural networks
Krizhevsky, Alex, Sutskever, Ilya, and Hinton, Geoff · 2012
Cited alongside, same era.
Deep learning with cots hpc systems
Speech recognition with deep recurrent neural networks
Graves, Alex, Mohamed, A-R, and Hinton, Geoffrey · 2013
Later among the works it cites.
Playing atari with deep reinforcement learning
Mnih, Volodymyr, Kavukcuoglu, Koray, Silver, David, Graves, Alex, Antonoglou, Ioannis, Wierstra, Daan, and Riedmiller, Martin · 2013
Later among the works it cites.
Concurrent reinforcement learning from customer interactions
Silver, David, Newnham, Leonard, Barker, David, Weller, Suzanne, and McFall, Jason · 2013
Later among the works it cites.
Very deep convolutional networks for large-scale image recognition
Simonyan, Karen and Zisserman, Andrew · 2014
Later among the works it cites.
Going deeper with convolutions
Szegedy, Christian, Liu, Wei, Jia, Yangqing, Sermanet, Pierre, Reed, Scott, Anguelov, Dragomir, Erhan, Dumitru, Vanhoucke, Vincent, and Rabinovich, Andrew · 2014
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Coates, Adam, Huval, Brody, Wang, Tao, Wu, David, Catanzaro, Bryan, and Andrew, Ng · 2013
Cited alongside, same era.
Human-level control through deep reinforcement learning
Mnih, Volodymyr, Kavukcuoglu, Koray, Silver, David, Rusu, Andrei A., Veness, Joel, Bellemare, Marc G., Graves, Alex, Riedmiller, Martin, Fidjeland, Andreas K., Ostrovski, Georg, Petersen, Stig, Beattie, Charles, Sadik, Amir, Antonoglou, Ioannis, King, Helen, Kumaran, Dharshan, Wierstra, Daan, Legg, Shane, and Hassabis, Demis · 2015
Closest in time.