Fetching the paper…
Reading the bibliography…
The deep reinforcement learning method usually requires a large number of training images and executing actions to obtain sufficient results.
K. Fukushima, “Neocognitron: A self-organizing neural network model for a mechanism of pattern recognition unaffected by shift in position,”
1980
Earlier work this paper cites.
A. G. Barto, R. S. Sutton, and C. W. Anderson, “Neuronlike adaptive elements that can solve difficult learning control problems,”
1983
Earlier work this paper cites.
C. J. Watkins and P. Dayan, “Q-learning,”
1992
Earlier work this paper cites.
L.-J. Lin, “Reinforcement learning for robots using neural networks,”
1993
Earlier work this paper cites.
G. Tesauro, “Temporal difference learning and td-gammon,”
1995
Earlier work this paper cites.
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner, “Gradient-based learning applied to document recognition,”
1998
Earlier work this paper cites.
R. S. Sutton and A. G. Barto,
1998
Earlier work this paper cites.
M. Riedmiller, “Neural fitted q iteration – first experiences with a data efficient neural reinforcement learning method,” in
2005
Earlier work this paper cites.
G. E. Hinton and R. R. Salakhutdinov, “Reducing the dimensionality of data with neural networks,”
2006
Earlier work this paper cites.
D. Erhan, Y. Bengio, A. Courville, P.-A. Manzagol, P. Vincent, and S. Bengio, “Why does unsupervised pre-training help deep learning?”
2010
Earlier work this paper cites.
V. Nair and G. E. Hinton, “Rectified linear units improve restricted boltzmann machines,” in
2010
Earlier work this paper cites.
S. Lange and M. Riedmiller, “Deep auto-encoder neural networks in reinforcement learning,” in
2010
Earlier work this paper cites.
J. Masci, U. Meier, D. Cireşan, and J. Schmidhuber, “Stacked convolutional auto-encoders for hierarchical feature extraction,” in
2011
Cited alongside, same era.
F. Abtahi and I. Fasel, “Deep belief nets as function approximators for reinforcement learning,” in
2011
Cited alongside, same era.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” in
2012
Cited alongside, same era.
Q. V. Le, “Building high-level features using large scale unsupervised learning,” in
2013
Cited alongside, same era.
M. G. Bellemare, Y. Naddaf, J. Veness, and M. Bowling, “The arcade learning environment: An evaluation platform for general agents,”
2013
Cited alongside, same era.
2015
Later among the works it cites.
2016
Later among the works it cites.
V. Mnih, A. Puigdomenech Badia, M. Mirza, A. Graves, T. P. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu, “Asynchronous Methods for Deep Reinforcement Learning,”
2016
Later among the works it cites.
T. Sandven, “Visual pretraining for deep q-learning,” Master’s thesis, Master’s thesis, NTNU, 2016
2016
Later among the works it cites.
G. Brockman, V. Cheung, L. Pettersson, J. Schneider, J. Schulman, J. Tang, and W. Zaremba, “Openai gym,” 2016. [Online]. Available: http://gym.openai.com/
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
X. Guo, S. Singh, H. Lee, R. L. Lewis, and X. Wang, “Deep learning for real-time atari game play using offline monte-carlo tree search planning,” in
2014
Cited alongside, same era.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, S. Petersen, C. Beattie, A. Sadik, I. Antonoglou, H. King, D. Kumaran, D. Wierstra, S. Legg, and D. Hassabis, “Human-level control through deep reinforcement learning,”
2015
Cited alongside, same era.
F. Chollet, “keras,” https://github.com/fchollet/keras, 2015
2015
Cited alongside, same era.
D. P. Kingma and J. Ba, “Adam: A method for stochastic optimization,” in
2015
Cited alongside, same era.
Y. LeCun, Y. Bengio, and G. Hinton, “Deep learning,”
2015
Cited alongside, same era.
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich, “Going deeper with convolutions,” in
2015
Cited alongside, same era.
V. Turchenko and A. Luczak, “Creation of a deep convolutional auto-encoder in caffe,”
2015
Cited alongside, same era.
2016
Later among the works it cites.
M. Plappert, “keras-rl,” https://github.com/matthiasplappert/keras-rl, 2016
2016
Later among the works it cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in
2016
Later among the works it cites.
X. Mao, C. Shen, and Y. Yang, “Image restoration using very deep fully convolutional encoder-decoder networks with symmetric skip connections,” in
2016
Later among the works it cites.
O. K. Oyedotun and K. Dimililer, “Pattern recognition: invariance learning in convolutional auto encoder network,”
2016
Later among the works it cites.
C. Finn, X. Y. Tan, Y. Duan, T. Darrell, S. Levine, and P. Abbeel, “Deep spatial autoencoders for visuomotor learning,” in
2016
Later among the works it cites.
K. Ito, T. Sueishi, Y. Yamakawa, and M. Ishikawa, “Tracking and recognition of a human hand in dynamic motion for janken (rock-paper-scissors) robot,” in
2016
Later among the works it cites.
D. Pathak, P. Agrawal, A. A. Efros, and T. Darrell, “Curiosity-driven exploration by self-supervised prediction,” in
2017
Later among the works it cites.