Fetching the paper…
Reading the bibliography…
In recent years, car makers and tech companies have been racing towards self driving cars.
Dynamic programming and lagrange multipliers
Richard Bellman · 1956
Earlier work this paper cites.
Introduction to the mathematical theory of control processes
Richard Bellman · 1971
Earlier work this paper cites.
A theory of the learnable
L. G. Valiant · 1984
Earlier work this paper cites.
Reinforcement learning in continuous time: Advantage updating
Leemon C Baird · 1994
Cited alongside, same era.
Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning
Richard S Sutton, Doina Precup, and Satinder Singh · 1999
Cited alongside, same era.
The instructor’s guide to real induction
Pete L Clark · 2012
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Later among the works it cites.
Discrete element crowd model for pedestrian evacuation through an exit
Peng Lin, Jian Ma, and Siuming Lo · 2016
Later among the works it cites.
Safe, multi-agent, reinforcement learning for autonomous driving
Shai Shalev-Shwartz, Shaked Shammah, and Amnon Shashua · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…