Fetching the paper…
Reading the bibliography…
Reinforcement learning is considered to be a strong AI paradigm which can be used to teach machines through interaction with the environment and learning from their mistakes, but it has not yet been successfully used for automotive applications.
Sutton, R. S. (1988). Learning to predict by the methods of temporal differences. Machine learning , 3 (1), 9-44
1988
Earlier work this paper cites.
Watkins, C. J. (1989). Learning from delayed rewards. Ph.D. dissertation, University of Cambridge England
1989
Earlier work this paper cites.
Watkins, C. J., & Dayan, P. (1992). Q-learning. Machine learning , 8 (3-4), 279-292
1992
Earlier work this paper cites.
Williams, R. J. (1992). Simple statistical gradient-following algorithms for connectionist reinforcement learning. Machine learning , 8 (3-4), 229-256
1992
Earlier work this paper cites.
Hochreiter, S., & Schmidhuber, J. (1997). Long short-term memory. Neural computation , 9 (8), 1735-1780
1997
Earlier work this paper cites.
Sutton, R. S., McAllester, D. A., Singh, S. P., Mansour, Y., & others. (1999). Policy Gradient Methods for Reinforcement Learning with Function Approximation. NIPS, 99, pp. 1057-1063
1999
Earlier work this paper cites.
Abbeel, P., & Ng, A. Y. (2004). Apprenticeship learning via inverse reinforcement learning. Proceedings of the twenty-first international conference on Machine learning, (p. 1)
2004
Earlier work this paper cites.
Krizhevsky, A., Sutskever, I., & Hinton, G. E. (2012). Imagenet classification with deep convolutional neural networks. Advances in neural information processing systems, (pp. 1097-1105)
2012
Earlier work this paper cites.
Karavolos, D. (2013). Q-learning with heuristic exploration in Simulated Car Racing
2013
Cited alongside, same era.
2013
Cited alongside, same era.
2013
Cited alongside, same era.
Mnih, V., Heess, N., Graves, A., & others. (2014). Recurrent models of visual attention. Advances in Neural Information Processing Systems, (pp. 2204-2212)
2014
Cited alongside, same era.
LeCun, Y., Bengio, Y., & Hinton, G. (2015). Deep learning. Nature , 521 (7553), 436-444
2015
Later among the works it cites.
Mnih, V., Kavukcuoglu, K., Silver, D., Rusu, A. A., Veness, J., Bellemare, M. G., et al. (2015). Human-level control through deep reinforcement learning. Nature , 518 (7540), 529-533
2015
Later among the works it cites.
2015
Later among the works it cites.
2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2015
Cited alongside, same era.
2015
Cited alongside, same era.
2015
Cited alongside, same era.
2015
Cited alongside, same era.
2016
Closest in time.
2016
Closest in time.
Sutton, R. S., & Barto, A. G. (2016). Reinforcement learning: An introduction. Online Draft
2016
Closest in time.