Fetching the paper…
Reading the bibliography…
Mapping states to actions in deep reinforcement learning is mainly based on visual information.
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu, “Asynchronous methods for deep reinforcement learning,” in
1937
Earlier work this paper cites.
C. J. Watkins and P. Dayan, “Q-learning,”
1992
Earlier work this paper cites.
id Software, “Doom,” 1993
1993
Earlier work this paper cites.
L. P. Kaelbling, M. L. Littman, and A. W. Moore, “Reinforcement learning: A survey,”
1996
Earlier work this paper cites.
R. S. Sutton, A. G. Barto
1998
Earlier work this paper cites.
B. Sallans and G. E. Hinton, “Reinforcement learning with factored states and actions,”
2004
Earlier work this paper cites.
D. Ververidis and C. Kotropoulos, “Emotional speech recognition: Resources, features, and methods,”
2006
Earlier work this paper cites.
A. Graves, A.-r. Mohamed, and G. Hinton, “Speech recognition with deep recurrent neural networks,” in
2013
Cited alongside, same era.
O. Abdel-Hamid, A.-r. Mohamed, H. Jiang, L. Deng, G. Penn, and D. Yu, “Convolutional neural networks for speech recognition,”
2014
Cited alongside, same era.
J. Ba, V. Mnih, and K. Kavukcuoglu, “Multiple object recognition with visual attention,”
2014
Cited alongside, same era.
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. Bernstein
2015
Cited alongside, same era.
K. J. Piczak, “Environmental sound classification with convolutional neural networks,” in
2015
Cited alongside, same era.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski
2015
Later among the works it cites.
S. Levine, C. Finn, T. Darrell, and P. Abbeel, “End-to-end training of deep visuomotor policies,”
2016
Later among the works it cites.
2016
Later among the works it cites.
2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
N. Heess, G. Wayne, D. Silver, T. Lillicrap, T. Erez, and Y. Tassa, “Learning continuous control policies by stochastic value gradients,” in
2015
Cited alongside, same era.
2017
Later among the works it cites.
2018
Later among the works it cites.