Fetching the paper…
Reading the bibliography…
The operational space of an autonomous vehicle (AV) can be diverse and vary significantly.
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu, “Asynchronous methods for deep reinforcement learning,” in International Conference on Machine Learning , 2016, pp. 1928–1937
1937
Earlier work this paper cites.
R. S. Sutton and A. G. Barto, Reinforcement learning: An introduction . MIT press Cambridge, 1998, vol. 1, no. 1
1998
Earlier work this paper cites.
M. Treiber, A. Hennecke, and D. Helbing, “Congested traffic states in empirical observations and microscopic simulations,” Physical review E , vol. 62, no. 2, p. 1805, 2000
2000
Earlier work this paper cites.
M. M. Minderhoud and P. H. Bovy, “Extended time-to-collision measures for road traffic safety assessment,” Accident Analysis & Prevention , vol. 33, no. 1, pp. 89–97, 2001
2001
Earlier work this paper cites.
A. Vahidi and A. Eskandarian, “Research advances in intelligent collision avoidance and adaptive cruise control,” IEEE transactions on intelligent transportation systems , vol. 4, no. 3, pp. 143–153, 2003
2003
Earlier work this paper cites.
T. Toledo and D. Zohar, “Modeling duration of lane changes,” Transportation Research Record: Journal of the Transportation Research Board , no. 1999, pp. 71–78, 2007
2007
Earlier work this paper cites.
A. Kesting, M. Treiber, and D. Helbing, “General lane-changing model mobil for car-following models,” Transportation Research Record , vol. 1999, no. 1, pp. 86–94, 2007
2007
Earlier work this paper cites.
A. L. Thomaz and C. Breazeal, “Teachable robots: Understanding human teaching behavior to build more effective robot learners,” Artificial Intelligence , vol. 172, no. 6-7, pp. 716–737, 2008
2008
Earlier work this paper cites.
J. Garcia and F. Fernández, “Safe exploration of state and action spaces in reinforcement learning,” Journal of Artificial Intelligence Research , vol. 45, pp. 515–564, 2012
2012
Earlier work this paper cites.
A. L. Maas, A. Y. Hannun, and A. Y. Ng, “Rectifier nonlinearities improve neural network acoustic models,” in Proc. icml , vol. 30, no. 1, 2013, p. 3
2013
Earlier work this paper cites.
2014
Earlier work this paper cites.
J. Erdmann, “Lane-changing model in sumo,” Proceedings of the SUMO2014 modeling mobility with open data , vol. 24, pp. 77–88, 2014
2014
Cited alongside, same era.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, et al. , “Human-level control through deep reinforcement learning,” Nature , vol. 518, no. 7540, p. 529, 2015
2015
Cited alongside, same era.
2015
Cited alongside, same era.
J. Garcıa and F. Fernández, “A comprehensive survey on safe reinforcement learning,” Journal of Machine Learning Research , vol. 16, no. 1, pp. 1437–1480, 2015
2015
Cited alongside, same era.
K. Arulkumaran, M. P. Deisenroth, M. Brundage, and A. A. Bharath, “Deep reinforcement learning: A brief survey,” IEEE Signal Processing Magazine , vol. 34, no. 6, pp. 26–38, 2017
2017
Later among the works it cites.
2017
Later among the works it cites.
H. Xu, Y. Gao, F. Yu, and T. Darrell, “End-to-end learning of driving models from large-scale video datasets,” in Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2017, pp. 2174–2182
2017
Later among the works it cites.
M. Zhang, N. Li, A. Girard, and I. Kolmanovsky, “A finite state machine based automated driving controller and its stochastic optimization,” in ASME 2017 Dynamic Systems and Control Conference . American Society of Mechanical Engineers, 2017, pp. V002T07A002–V002T07A002
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2015
Cited alongside, same era.
C. Chen, A. Seff, A. Kornhauser, and J. Xiao, “Deepdriving: Learning affordance for direct perception in autonomous driving,” in Computer Vision (ICCV), 2015 IEEE International Conference on . IEEE, 2015, pp. 2722–2730
2015
Cited alongside, same era.
B. Paden, M. Čáp, S. Z. Yong, D. Yershov, and E. Frazzoli, “A survey of motion planning and control techniques for self-driving urban vehicles,” IEEE Transactions on Intelligent Vehicles , vol. 1, no. 1, pp. 33–55, 2016
2016
Cited alongside, same era.
D. Silver, J. Schrittwieser, K. Simonyan, I. Antonoglou, A. Huang, A. Guez, T. Hubert, L. Baker, M. Lai, A. Bolton, et al. , “Mastering the game of go without human knowledge,” Nature , vol. 550, no. 7676, p. 354, 2017
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
N. Li, D. W. Oyler, M. Zhang, Y. Yildiz, I. Kolmanovsky, and A. R. Girard, “Game theoretic modeling of driver and vehicle interactions for verification and validation of autonomous vehicle control systems,” IEEE Transactions on control systems technology , 2017
2017
Later among the works it cites.
V. Ivanovic, E. Tseng, M. Hafner, and P. Madhavan, “Analytical and experimental study of automated lane change control based on lane centering algorithms,” Internal technical report, Ford Motor Company , 2017
2017
Later among the works it cites.
W. Schwarting, J. Alonso-Mora, and D. Rus, “Planning and decision-making for autonomous vehicles,” Annual Review of Control, Robotics, and Autonomous Systems , no. 0, 2018
2018
Later among the works it cites.
M. Hessel, J. Modayil, H. Van Hasselt, T. Schaul, G. Ostrovski, W. Dabney, D. Horgan, B. Piot, M. Azar, and D. Silver, “Rainbow: Combining improvements in deep reinforcement learning,” in Thirty-Second AAAI Conference on Artificial Intelligence , 2018
2018
Later among the works it cites.
S. Hecker, D. Dai, and L. Van Gool, “End-to-end learning of driving models with surround-view cameras and route planners,” in European Conference on Computer Vision (ECCV) , 2018
2018
Later among the works it cites.
A. Hill, A. Raffin, M. Ernestus, R. Traore, P. Dhariwal, C. Hesse, O. Klimov, A. Nichol, M. Plappert, A. Radford, J. Schulman, S. Sidor, and Y. Wu, “Stable baselines,” https://github.com/hill-a/stable-baselines , 2018
2018
Later among the works it cites.
H. Van Hasselt, A. Guez, and D. Silver, “Deep reinforcement learning with double q-learning.” in AAAI , vol. 16, 2016, pp. 2094–2100
2094
Closest in time.