Fetching the paper…
Reading the bibliography…
Traffic light timing optimization is still an active line of research despite the wealth of scientific literature on the topic, and the problem remains unsolved for any non-toy scenario.
F. Webster, Traffic signal settings, road research technical paper no. 39, Road Research Laboratory
1958
Earlier work this paper cites.
J. D. Little, The synchronization of traffic signals by mixed-integer linear programming, Operations Research 14 (4) (1966) 568–594
1966
Earlier work this paper cites.
D. I. Robertson, ’TRANSYT’ method for area traffic control, Traffic Engineering & Control 10 (1969) 271–281
1969
Earlier work this paper cites.
P. Hunt, D. Robertson, R. Bretherton, M. C. Royle, The SCOOT on-line traffic signal optimisation technique, Traffic Engineering & Control 23 (4)
1982
Earlier work this paper cites.
N. H. Gartner, OPAC: A demand-responsive strategy for traffic signal control, 906, 1983
1983
Earlier work this paper cites.
J.-J. Henry, J. L. Farges, J. Tuffal, The PRODYN real time traffic algorithm, IFAC Proceedings Volumes 16 (4) (1983) 305–310
1983
Earlier work this paper cites.
F. Lin, Use of Binary Choice Decision Process for Adaptive Signal Control, Journal of Transportation Engineering 115 (3) (1989) 270–282, doi: 10.1061/(ASCE)0733-947X(1989)115:3(270)
1989
Earlier work this paper cites.
F. Boillot, Optimal signal control of urban traffic networks, in: International Conference on Road Traffic Monitoring and Control (6th: 1992: London, England). 6th International Conference on Road Traffic Monitoring and Control, 1992
1992
Earlier work this paper cites.
J. Favilla, A. Machion, F. Gomide, Fuzzy traffic control: adaptive strategies, in: Fuzzy Systems, 1993., Second IEEE International Conference on, IEEE, 506–511, 1993
1993
Earlier work this paper cites.
S. Chiu, S. Chand, Adaptive traffic signal control using fuzzy logic, in: Fuzzy Systems, 1993., Second IEEE International Conference on, IEEE, 1371–1376, 1993
1993
Earlier work this paper cites.
H.-T. Fritzsche, A model for traffic simulation, Traffic Engineering and Control 35 (5) (1994) 317–21
1994
Earlier work this paper cites.
G. Tesauro, Temporal difference learning and TD-Gammon, Communications of the ACM 38 (3) (1995) 58–68
1995
Earlier work this paper cites.
G. D. Cameron, G. I. Duncan, PARAMICS—Parallel microscopic simulation of road traffic, The Journal of Supercomputing 10 (1) (1996) 25–53
1996
Earlier work this paper cites.
S. Sen, K. L. Head, Controlled optimization of phases at an intersection, Transportation science 31 (1) (1997) 5–17
1997
Earlier work this paper cites.
T. L. Thorpe, Vehicle traffic light control using sarsa, in: Online]. Available: citeseer. ist. psu. edu/thorpe97vehicle. html, Citeseer, 1997
1997
Earlier work this paper cites.
J. B. Pollack, A. D. Blair, Why did TD-Gammon Work?, Advances in Neural Information Processing Systems (1997) 10–16
1997
Earlier work this paper cites.
R. S. Sutton, A. G. Barto, Reinforcement learning: An introduction, vol. 1, MIT press Cambridge, 1998
1998
Earlier work this paper cites.
R. S. Sutton, D. A. McAllester, S. P. Singh, Y. Mansour, et al., Policy Gradient Methods for Reinforcement Learning with Function Approximation., in: NIPS, vol. 99, 1057–1063, 1999
1999
Earlier work this paper cites.
N. M. Rouphail, B. B. Park, J. Sacks, Direct signal timing optimization: Strategy development and results, in: In XI Pan American Conference in Traffic and Transportation Engineering, Citeseer, 2000
2000
Earlier work this paper cites.
M. Wiering, et al., Multi-agent reinforcement learning for traffic light control, in: ICML, 1151–1158, 2000
2000
Earlier work this paper cites.
C. Guestrin, M. Lagoudakis, R. Parr, Coordinated reinforcement learning, in: ICML, vol. 2, 227–234, 2002
2002
Earlier work this paper cites.
E. Camponogara, W. Kraus Jr, Distributed learning agents in urban traffic control, in: Portuguese Conference on Artificial Intelligence, Springer, 324–335, 2003
2003
Cited alongside, same era.
M. Wiering, J. Vreeken, J. Van Veenen, A. Koopman, Simulation and optimization of traffic in a city, in: Intelligent Vehicles Symposium, 2004 IEEE, IEEE, 453–458, 2004
2004
Cited alongside, same era.
C.-Q. Cai, Z.-S. Yang, Study on urban traffic management based on multi-agent system, in: 2007 International Conference on Machine Learning and Cybernetics, vol. 1, IEEE, 25–29, 2007
2007
Cited alongside, same era.
P. Holm, D. Tomich, J. Sloboden, C. Lowrance, Traffic analysis toolbox volume iv: guidelines for applying corsim microsimulation modeling software, Tech. Rep., 2007
2007
Cited alongside, same era.
T. Tieleman, G. Hinton, Lecture 6.5-rmsprop: Divide the gradient by a running average of its recent magnitude, COURSERA: Neural Networks for Machine Learning 4 (2)
2012
Later among the works it cites.
S. El-Tantawy, B. Abdulhai, H. Abdelgawad, Multiagent reinforcement learning for integrated network of adaptive traffic signal controllers (MARLIN-ATSC): methodology and large-scale application on downtown Toronto, IEEE Transactions on Intelligent Transportation Systems 14 (3) (2013) 1140–1150
2013
Later among the works it cites.
2013
Later among the works it cites.
J. Garcia-Nieto, A. C. Olivera, E. Alba, Optimal cycle program of traffic lights with particle swarm optimization, IEEE Transactions on Evolutionary Computation 17 (6) (2013) 823–839
2013
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2007
Cited alongside, same era.
Y. Bengio, P. Lamblin, D. Popovici, H. Larochelle, et al., Greedy layer-wise training of deep networks, Advances in neural information processing systems 19 (2007) 153
2007
Cited alongside, same era.
L. Kuyer, S. Whiteson, B. Bakker, N. Vlassis, Multiagent reinforcement learning for urban traffic control using coordination graphs, in: Joint European Conference on Machine Learning and Knowledge Discovery in Databases, Springer, 656–671, 2008
2008
Cited alongside, same era.
C. Shao, Adaptive control strategy for isolated intersection and traffic network, Ph.D. thesis, The University of Akron, 2009
2009
Cited alongside, same era.
J. Casas, J. L. Ferrer, D. Garcia, J. Perarnau, A. Torday, Traffic simulation with aimsun, in: Fundamentals of traffic simulation, Springer, 173–232, 2010
2010
Cited alongside, same era.
I. Arel, C. Liu, T. Urbanik, A. Kohls, Reinforcement learning-based multi-agent system for network traffic signal control, IET Intelligent Transport Systems 4 (2) (2010) 128–135
2010
Cited alongside, same era.
B. Bakker, S. Whiteson, L. Kester, F. C. Groen, Traffic light control by multiagent reinforcement learning systems, in: Interactive Collaborative Information Systems, Springer, 475–510, 2010
2010
Cited alongside, same era.
P. Vincent, H. Larochelle, I. Lajoie, Y. Bengio, P.-A. Manzagol, Stacked denoising autoencoders: Learning useful representations in a deep network with a local denoising criterion, Journal of Machine Learning Research 11 (Dec) (2010) 3371–3408
2010
Cited alongside, same era.
A. L. Maas, A. Y. Hannun, A. Y. Ng, Rectifier nonlinearities improve neural network acoustic models, in: Proc. ICML, vol. 30, 2013
2013
Later among the works it cites.
K. I. Harrington, E. Awa, S. Cussat-Blanc, J. Pollack, Robot coverage control by evolved neuromodulation, in: Neural Networks (IJCNN), The 2013 International Joint Conference on, IEEE, 1–8, 2013
2013
Later among the works it cites.
S. El-Tantawy, B. Abdulhai, H. Abdelgawad, Design of reinforcement learning parameters for seamless application of adaptive traffic signal control, Journal of Intelligent Transportation Systems 18 (3) (2014) 227–245
2014
Later among the works it cites.
D. Kingma, J. Ba, Adam: A method for stochastic optimization, arXiv preprint arXiv:1412.6980
2014
Later among the works it cites.
D. Silver, G. Lever, N. Heess, T. Degris, D. Wierstra, M. Riedmiller, Deterministic Policy Gradient Algorithms, in: Proceedings of the 31st International Conference on Machine Learning (ICML-14), JMLR Workshop and Conference Proceedings, 387–395, URL http://jmlr.org/proceedings/papers/v32/silver14.pdf , 2014
2014
Later among the works it cites.
Y. Feng, K. L. Head, S. Khoshmagham, M. Zamanipour, A real-time adaptive signal control in a connected vehicle environment, Transportation Research Part C: Emerging Technologies 55 (2015) 460–473
2015
Later among the works it cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, et al., Human-level control through deep reinforcement learning, Nature 518 (7540) (2015) 529–533
2015
Later among the works it cites.
2015
Later among the works it cites.
2015
Later among the works it cites.
K. He, X. Zhang, S. Ren, J. Sun, Delving deep into rectifiers: Surpassing human-level performance on imagenet classification, in: Proceedings of the IEEE International Conference on Computer Vision, 1026–1034, 2015
2015
Later among the works it cites.
J. Schulman, S. Levine, P. Abbeel, M. I. Jordan, P. Moritz, Trust Region Policy Optimization., in: ICML, 1889–1897, 2015
2015
Later among the works it cites.
L. Li, Y. Lv, F. Y. Wang, Traffic signal timing via deep reinforcement learning, IEEE/CAA Journal of Automatica Sinica 3 (3) (2016) 247–254, ISSN 2329-9266, doi: 10.1109/JAS.2016.7508798
2016
Later among the works it cites.
E. van der Pol, Deep Reinforcement Learning for Coordination in Traffic Light Control
2016
Later among the works it cites.
2016
Later among the works it cites.
I. Goodfellow, Y. Bengio, A. Courville, Deep Learning, MIT Press, http://www.deeplearningbook.org , 2016
2016
Later among the works it cites.