Fetching the paper…
Reading the bibliography…
In order to drive safely and efficiently under merging scenarios, autonomous vehicles should be aware of their surroundings and make decisions by interacting with other road participants.
1904
Earlier work this paper cites.
R. J. Williams, “Simple statistical gradient-following algorithms for connectionist reinforcement learning,” Machine learning , vol. 8, no. 3-4, pp. 229–256, 1992
1992
Earlier work this paper cites.
V. R. Konda and J. N. Tsitsiklis, “Actor-critic algorithms,” in Advances in neural information processing systems , 2000, pp. 1008–1014
2000
Earlier work this paper cites.
Y.-H. Chang, T. Ho, and L. P. Kaelbling, “All learning is local: Multi-agent learning in global reward games,” in Advances in neural information processing systems , 2004, pp. 807–814
2004
Earlier work this paper cites.
M. Montemerlo, J. Becker, S. Bhat, H. Dahlkamp, D. Dolgov, S. Ettinger, D. Haehnel, T. Hilden, G. Hoffmann, B. Huhnke et al. , “Junior: The stanford entry in the urban challenge,” Journal of field Robotics , vol. 25, no. 9, pp. 569–597, 2008
2008
Earlier work this paper cites.
C. Urmson, J. Anhalt, D. Bagnell, C. Baker, R. Bittner, M. Clark, J. Dolan, D. Duggins, T. Galatali, C. Geyer et al. , “Autonomous driving in urban environments: Boss and the urban challenge,” Journal of Field Robotics , vol. 25, no. 8, pp. 425–466, 2008
2008
Earlier work this paper cites.
I. Miller, M. Campbell, D. Huttenlocher, F.-R. Kline, A. Nathan, S. Lupashin, J. Catlin, B. Schimpf, P. Moran, N. Zych et al. , “Team cornell’s skynet: Robust perception and planning in an urban environment,” Journal of Field Robotics , vol. 25, no. 8, pp. 493–527, 2008
2008
Cited alongside, same era.
C. R. Baker and J. M. Dolan, “Traffic interaction in the urban challenge: Putting boss on its best behavior,” in 2008 IEEE/RSJ International Conference on Intelligent Robots and Systems . IEEE, 2008, pp. 1752–1758
2008
Cited alongside, same era.
J. Nilsson and J. Sjöberg, “Strategic decision making for automated driving on two-lane, one way roads using model predictive control,” in Intelligent Vehicles Symposium (IV), 2013 IEEE . IEEE, 2013, pp. 1253–1258
2013
Cited alongside, same era.
D. Sadigh, S. Sastry, S. A. Seshia, and A. D. Dragan, “Planning for autonomous cars that leverage effects on human actions.” in Robotics: Science and Systems , 2016
2016
Cited alongside, same era.
2017
Later among the works it cites.
M. Mukadam, A. Cosgun, A. Nakhaei, and K. Fujimura, “Tactical decision making for lane changing with deep reinforcement learning,” in 31th Conference on Neural Information Processing Systems (NIPS) , 2017
2017
Later among the works it cites.
W. Schwarting, J. Alonso-Mora, and D. Rus, “Planning and decision-making for autonomous vehicles,” Annual Review of Control, Robotics, and Autonomous Systems , vol. 1, pp. 187–210, 2018
2018
Later among the works it cites.
J. N. Foerster, G. Farquhar, T. Afouras, N. Nardelli, and S. Whiteson, “Counterfactual multi-agent policy gradients,” in Thirty-Second AAAI Conference on Artificial Intelligence , 2018
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
C. Hubmann, M. Becker, D. Althoff, D. Lenz, and C. Stiller, “Decision making for autonomous driving considering interaction and uncertain prediction of surrounding vehicles,” in Intelligent Vehicles Symposium (IV), 2017 IEEE . IEEE, 2017, pp. 1671–1678
2017
Cited alongside, same era.
R. Lowe, Y. Wu, A. Tamar, J. Harb, O. P. Abbeel, and I. Mordatch, “Multi-agent actor-critic for mixed cooperative-competitive environments,” in Advances in Neural Information Processing Systems , 2017, pp. 6379–6390
2017
Cited alongside, same era.
2018
Later among the works it cites.