Fetching the paper…
Reading the bibliography…
Existing inefficient traffic light control causes numerous problems, such as long delay and waste of energy.
S. Chiu and S. Chand, “Adaptive traffic signal control using fuzzy logic,” in The First IEEE Regional Conference on Aerospace Control Systems , April 1993, pp. 1371–1376
1993
Earlier work this paper cites.
S. Krauß, “Towards a unified view of microscopic traffic flow theories,” IFAC Transportation Systems , vol. 30, no. 8, pp. 901–905, June 1997
1997
Earlier work this paper cites.
R. S. Sutton and A. G. Barto, Reinforcement learning: An introduction . MIT press Cambridge, March 1998, vol. 1, no. 1
1998
Earlier work this paper cites.
B. De Schutter, “Optimal traffic light control for a single intersection,” in American Control Conference , vol. 3, June 1999, pp. 2195–2199
1999
Earlier work this paper cites.
B. Abdulhai, R. Pringle, and G. J. Karakoulas, “Reinforcement learning for true adaptive traffic signal control,” Journal of Transportation Engineering , vol. 129, no. 3, pp. 278–285, May 2003
2003
Earlier work this paper cites.
H. Hartenstein and L. Laberteaux, “A tutorial survey on vehicular ad hoc networks,” IEEE Communications magazine , vol. 46, no. 6, June 2008
2008
Earlier work this paper cites.
I. Arel, C. Liu, T. Urbanik, and A. Kohls, “Reinforcement learning-based multi-agent system for network traffic signal control,” IET Intelligent Transport Systems , vol. 4, no. 2, pp. 128–135, June 2010
2010
Earlier work this paper cites.
P. Balaji, X. German, and D. Srinivasan, “Urban traffic signal control using reinforcement learning agents,” IET Intelligent Transport Systems , vol. 4, no. 3, pp. 177–188, September 2010
2010
Earlier work this paper cites.
Y. K. Chin, N. Bolong, A. Kiring, S. S. Yang, and K. T. K. Teo, “Q-learning based traffic optimization in management of signal timing plan,” International Journal of Simulation, Systems, Science & Technology , vol. 12, no. 3, pp. 29–35, June 2011
2011
Earlier work this paper cites.
D. Krajzewicz, J. Erdmann, M. Behrisch, and L. Bieker, “Recent development and applications of sumo-simulation of urban mobility,” International Journal On Advances in Systems and Measurements , vol. 5, no. 3&4, pp. 128–138, December 2012
2012
Earlier work this paper cites.
M. Abdoos, N. Mozayani, and A. L. Bazzan, “Holonic multi-agent system for traffic signals control,” Engineering Applications of Artificial Intelligence , vol. 26, no. 5, pp. 1575–1587, May-Jun 2013
2013
Cited alongside, same era.
S. El-Tantawy, B. Abdulhai, and H. Abdelgawad, “Design of reinforcement learning parameters for seamless application of adaptive traffic signal control,” Journal of Intelligent Transportation Systems , vol. 18, no. 3, pp. 227–245, July 2014
2014
Cited alongside, same era.
2014
Cited alongside, same era.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski et al. , “Human-level control through deep reinforcement learning,” Nature , vol. 518, no. 7540, pp. 529–533, February 2015
2015
Cited alongside, same era.
E. van der Pol, “Deep reinforcement learning for coordination in traffic light control,” Master’s thesis, University of Amsterdam, August 2016
2016
Later among the works it cites.
2016
Later among the works it cites.
2016
Later among the works it cites.
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2015
Cited alongside, same era.
2015
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Delving deep into rectifiers: Surpassing human-level performance on imagenet classification,” in Proceedings of the IEEE international conference on computer vision , December 2015, pp. 1026–1034
2015
Cited alongside, same era.
2016
Cited alongside, same era.
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot et al. , “Mastering the game of go with deep neural networks and tree search,” Nature , vol. 529, no. 7587, pp. 484–489, January 2016
2016
Cited alongside, same era.
L. Li, Y. Lv, and F.-Y. Wang, “Traffic signal timing via deep reinforcement learning,” IEEE/CAA Journal of Automatica Sinica , vol. 3, no. 3, pp. 247–254, July 2016
2016
Cited alongside, same era.
2017
Later among the works it cites.
D. Silver, J. Schrittwieser, K. Simonyan, I. Antonoglou, A. Huang, A. Guez, T. Hubert, L. Baker, M. Lai, A. Bolton et al. , “Mastering the game of go without human knowledge,” Nature , vol. 550, no. 7676, p. 354, October 2017
2017
Later among the works it cites.
2017
Later among the works it cites.
X. Liang and G. Wang, “A convolutional neural network for transportation mode detection based on smartphone platform,” in 2017 IEEE 14th International Conference on Mobile Ad Hoc and Sensor Systems (MASS) . IEEE, October 2017, pp. 338–342
2017
Later among the works it cites.
H. Van Hasselt, A. Guez, and D. Silver, “Deep reinforcement learning with double q-learning,” in Proceedings of the Thirtieth AAAI Conference on Artificial Intelligence , February 2016, pp. 2094–2100
2094
Closest in time.