Fetching the paper…
Reading the bibliography…
On-ramp merging is a challenging task for autonomous vehicles (AVs), especially in mixed traffic where AVs coexist with human-driven vehicles (HDVs).
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu, “Asynchronous methods for deep reinforcement learning,” in International conference on machine learning , 2016, pp. 1928–1937
1937
Earlier work this paper cites.
L. N. Jacobson, K. C. Henry, and O. Mehyar, Real-time metering algorithm for centralized control , 1989, no. 1232
1989
Earlier work this paper cites.
R. J. Williams, “Simple statistical gradient-following algorithms for connectionist reinforcement learning,” Machine learning , vol. 8, no. 3-4, pp. 229–256, 1992
1992
Earlier work this paper cites.
M. Tan, “Multi-agent reinforcement learning: Independent vs. cooperative agents,” in Proceedings of the tenth international conference on machine learning , 1993, pp. 330–337
1993
Earlier work this paper cites.
M. Treiber, A. Hennecke, and D. Helbing, “Congested traffic states in empirical observations and microscopic simulations,” Physical review E , vol. 62, no. 2, p. 1805, 2000
2000
Earlier work this paper cites.
T. Ayres, L. Li, D. Schleuning, and D. Young, “Preferred time-headway of highway drivers,” in ITSC 2001. 2001 IEEE Intelligent Transportation Systems. Proceedings (Cat. No. 01TH8585) . IEEE, 2001, pp. 826–829
2001
Earlier work this paper cites.
J. Hourdakis and P. G. Michalopoulos, “Evaluation of ramp control effectiveness in two twin cities freeways,” Transportation Research Record , vol. 1811, no. 1, pp. 21–29, 2002
2002
Earlier work this paper cites.
M. Papageorgiou and A. Kotsialos, “Freeway ramp metering: An overview,” IEEE transactions on intelligent transportation systems , vol. 3, no. 4, pp. 271–281, 2002
2002
Earlier work this paper cites.
D. H. Wolpert and K. Tumer, “Optimal payoff functions for members of collectives,” in Modeling complexity in economic and social systems . World Scientific, 2002, pp. 355–369
2002
Earlier work this paper cites.
M. Papageorgiou, C. Diakaki, V. Dinopoulou, A. Kotsialos, and Y. Wang, “Review of road traffic control strategies,” Proceedings of the IEEE , vol. 91, no. 12, pp. 2043–2067, 2003
2003
Earlier work this paper cites.
D. Ni and J. D. Leonard II, “A simplified kinematic wave model at a merge bottleneck,” Applied mathematical modelling , vol. 29, no. 11, pp. 1054–1072, 2005
2005
Earlier work this paper cites.
D. Bagnell and A. Ng, “On local rewards and scaling distributed reinforcement learning,” Advances in Neural Information Processing Systems , vol. 18, pp. 91–98, 2005
2005
Earlier work this paper cites.
A. Kesting, M. Treiber, and D. Helbing, “General lane-changing model mobil for car-following models,” Transportation Research Record , vol. 1999, no. 1, pp. 86–94, 2007
2007
Earlier work this paper cites.
I. Papamichail and M. Papageorgiou, “Traffic-responsive linked ramp-metering control,” IEEE Transactions on Intelligent Transportation Systems , vol. 9, no. 1, pp. 111–121, 2008
2008
Earlier work this paper cites.
C. Thiemann, M. Treiber, and A. Kesting, “Estimating acceleration and lane-changing dynamics from next generation simulation trajectory data,” Transportation Research Record , vol. 2088, no. 1, pp. 90–101, 2008
2008
Earlier work this paper cites.
C.-M. Chou, C.-Y. Li, W.-M. Chien, and K.-c. Lan, “A feasibility study on vehicle-to-infrastructure communication: Wifi vs. wimax,” in 2009 tenth international conference on mobile data management: systems, services and middleware . IEEE, 2009, pp. 397–398
2009
Earlier work this paper cites.
L. Leclercq, J. A. Laval, and N. Chiabaut, “Capacity drops at merges: An endogenous model,” Procedia-Social and Behavioral Sciences , vol. 17, pp. 12–26, 2011
2011
Earlier work this paper cites.
T. Officials, A Policy on Geometric Design of Highways and Streets, 2011 . AASHTO, 2011
2011
Earlier work this paper cites.
2013
Earlier work this paper cites.
W. Cao, M. Mukai, and T. Kawabe, “Two-dimensional merging path generation using model predictive control,” Artificial Life and Robotics , vol. 17, no. 3-4, pp. 350–356, 2013
2013
Earlier work this paper cites.
W. Cao, M. Mukai, T. Kawabe, H. Nishira, and N. Fujiki, “Cooperative vehicle path generation during merging using model predictive control with real-time optimization,” Control Engineering Practice , vol. 34, pp. 98–105, 2015
2015
Earlier work this paper cites.
2015
Cited alongside, same era.
2015
Cited alongside, same era.
J. Garcıa and F. Fernández, “A comprehensive survey on safe reinforcement learning,” Journal of Machine Learning Research , vol. 16, no. 1, pp. 1437–1480, 2015
2015
Cited alongside, same era.
V. V. Dixit, S. Chand, and D. J. Nair, “Autonomous vehicles: disengagements, accidents and reaction times,” PLoS one , vol. 11, no. 12, p. e0168054, 2016
2016
Cited alongside, same era.
C. Yu, X. Wang, X. Xu, M. Zhang, H. Ge, J. Ren, L. Sun, B. Chen, and G. Tan, “Distributed multiagent coordinated learning for autonomous driving in highways based on dynamic coordination graphs,” IEEE Transactions on Intelligent Transportation Systems , vol. 21, no. 2, pp. 735–748, 2019
2019
Later among the works it cites.
T. Chu, J. Wang, L. Codecà, and Z. Li, “Multi-agent deep reinforcement learning for large-scale traffic signal control,” IEEE Transactions on Intelligent Transportation Systems , vol. 21, no. 3, pp. 1086–1095, 2019
2019
Later among the works it cites.
OpenAI, :, C. Berner, G. Brockman, B. Chan, V. Cheung, P. Dębiak, C. Dennison, D. Farhi, Q. Fischer, S. Hashme, C. Hesse, R. Józefowicz, S. Gray, C. Olsson, J. Pachocki, M. Petrov, H. P. d. O. Pinto, J. Raiman, T. Salimans, J. Schlatter, J. Schneider, S. Sidor, I. Sutskever, J. Tang, F. Wolski, and S. Zhang, “Dota 2 with large scale deep reinforcement learning,” 2019
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Rios-Torres and A. A. Malikopoulos, “A survey on the coordination of connected and automated vehicles at intersections and merging at highway on-ramps,” IEEE Transactions on Intelligent Transportation Systems , vol. 18, no. 5, pp. 1066–1077, 2016
2016
Cited alongside, same era.
F. M. Favarò, N. Nader, S. O. Eurich, M. Tripp, and N. Varadaraju, “Examining accident reports involving autonomous vehicles in california,” PLoS one , vol. 12, no. 9, p. e0184952, 2017
2017
Cited alongside, same era.
J. B. Rawlings, D. Q. Mayne, and M. Diehl, Model predictive control: theory, computation, and design . Nob Hill Publishing Madison, WI, 2017, vol. 2
2017
Cited alongside, same era.
R. Lowe, Y. I. Wu, A. Tamar, J. Harb, O. P. Abbeel, and I. Mordatch, “Multi-agent actor-critic for mixed cooperative-competitive environments,” in Advances in neural information processing systems , 2017, pp. 6379–6390
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
N. Li, H. Chen, I. Kolmanovsky, and A. Girard, “An explicit decision tree approach for automated driving,” in Dynamic Systems and Control Conference , vol. 58271. American Society of Mechanical Engineers, 2017, p. V001T45A003
2017
Cited alongside, same era.
P. Polack, F. Altché, B. d’Andréa Novel, and A. de La Fortelle, “The kinematic bicycle model: A consistent model for planning feasible trajectories for autonomous vehicles?” in 2017 IEEE intelligent vehicles symposium (IV) . IEEE, 2017, pp. 812–818
2017
Cited alongside, same era.
2020
Later among the works it cites.
P. Palanisamy, “Multi-agent connected autonomous driving using deep reinforcement learning,” in 2020 International Joint Conference on Neural Networks (IJCNN) . IEEE, 2020, pp. 1–7
2020
Later among the works it cites.
2020
Later among the works it cites.
S. Bhalla, S. G. Subramanian, and M. Crowley, “Deep multi agent reinforcement learning for autonomous driving,” in Canadian Conference on Artificial Intelligence . Springer, 2020, pp. 67–78
2020
Later among the works it cites.
2020
Later among the works it cites.
2020
Later among the works it cites.
D. Chen, L. Jiang, Y. Wang, and Z. Li, “Autonomous driving using safe reinforcement learning by incorporating a regret-based human lane-changing decision model,” in 2020 American Control Conference (ACC) . IEEE, 2020, pp. 4355–4361
2020
Later among the works it cites.
J. Wang, Y. Zhang, T.-K. Kim, and Y. Gu, “Shapley q-value: A local reward approach to solve global reward games,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 34, no. 05, 2020, pp. 7285–7292
2020
Later among the works it cites.
2020
Later among the works it cites.
D. M. Saxena, S. Bae, A. Nakhaei, K. Fujimura, and M. Likhachev, “Driving in dense traffic with model-free reinforcement learning,” in 2020 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2020, pp. 5385–5392
2020
Later among the works it cites.
W. Zhao, J. P. Queralta, and T. Westerlund, “Sim-to-real transfer in deep reinforcement learning for robotics: a survey,” in 2020 IEEE Symposium Series on Computational Intelligence (SSCI) . IEEE, 2020, pp. 737–744
2020
Later among the works it cites.
“Future of driving,” https://www.tesla.com/autopilot , accessed: 2021-03-31
2021
Closest in time.
“Apollo open platform,” https://apollo.auto/developer.html , accessed: 2021-03-31
2021
Closest in time.
N. Naderializadeh, J. Sydir, M. Simsek, and H. Nikopour, “Resource management in wireless networks via multi-agent deep reinforcement learning,” IEEE Transactions on Wireless Communications , 2021
2021
Closest in time.
J. K. Terry, N. Grammel, A. Hari, L. Santos, and B. Black, “Revisiting parameter sharing in multi-agent deep reinforcement learning,” 2021
2021
Closest in time.
2021
Closest in time.