Fetching the paper…
Reading the bibliography…
Sub-optimal control policies in transportation systems negatively impact mobility, the environment and human health.
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu, “Asynchronous methods for deep reinforcement learning,” in
1937
Earlier work this paper cites.
F. Webster, “Traffic signal settings, road research technical paper no. 39,”
1958
Earlier work this paper cites.
S. Linnainmaa, “Taylor expansion of the accumulated rounding error,”
1976
Earlier work this paper cites.
N. Gartner, “A demand-responsive strategy for traffic signal control,”
1983
Earlier work this paper cites.
D. E. Rumelhart, G. E. Hinton, and R. J. Williams, “Learning representations by back-propagating errors,”
1986
Earlier work this paper cites.
R. S. Sutton, “Learning to predict by the methods of temporal differences,”
1988
Earlier work this paper cites.
P. Lowrie, “Scats, sydney co-ordinated adaptive traffic system: A traffic responsive method of controlling urban traffic,” 1990
1990
Earlier work this paper cites.
C. J. Watkins and P. Dayan, “Q-learning,”
1992
Earlier work this paper cites.
L.-J. Lin, “Self-improving reactive agents based on reinforcement learning, planning and teaching,”
1992
Earlier work this paper cites.
S. Mikami and Y. Kakazu, “Genetic reinforcement learning for cooperative traffic signal control,” in
1994
Earlier work this paper cites.
T. L. Thorpe and C. W. Anderson, “Traffic light control using sarsa with three state representations,” Citeseer, Tech. Rep., 1996
1996
Earlier work this paper cites.
R. S. Sutton and A. G. Barto,
1998
Earlier work this paper cites.
R. S. Sutton, D. Precup, and S. Singh, “Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning,”
1999
Earlier work this paper cites.
P. Mirchandani and L. Head, “A real-time traffic signal control system: architecture, algorithms, and analysis,”
2001
Earlier work this paper cites.
E. Bingham, “Reinforcement learning in neurofuzzy traffic signal control,”
2001
Earlier work this paper cites.
E. Jones, T. Oliphant, P. Peterson
2001
Earlier work this paper cites.
F. Luyanda, D. Gettman, L. Head, S. Shelby, D. Bullock, and P. Mirchandani, “Acs-lite algorithmic architecture: applying adaptive control system technology to closed-loop traffic signal control systems,”
2003
Earlier work this paper cites.
B. Abdulhai, R. Pringle, and G. J. Karakoulas, “Reinforcement learning for true adaptive traffic signal control,”
2003
Earlier work this paper cites.
C. Gershenson, “Self-organizing traffic lights,”
2004
Earlier work this paper cites.
J. Lee, B. Abdulhai, A. Shalaby, and E.-H. Chung, “Real-time optimization for adaptive traffic signal control using genetic algorithms,”
2005
Earlier work this paper cites.
H. Prothmann, F. Rochner, S. Tomforde, J. Branke, C. Müller-Schloer, and H. Schmeck, “Organic control of traffic lights,” in
2008
Earlier work this paper cites.
L. Kuyer, S. Whiteson, B. Bakker, and N. Vlassis, “Multiagent reinforcement learning for urban traffic control using coordination graphs,” in
2008
Earlier work this paper cites.
L. Singh, S. Tripathi, and H. Arora, “Time optimization for traffic signal control using genetic algorithm,”
2009
Earlier work this paper cites.
A. Stevanovic,
2010
Earlier work this paper cites.
L. Prashanth and S. Bhatnagar, “Reinforcement learning with function approximation for traffic signal control,”
2011
Earlier work this paper cites.
D. Krajzewicz, J. Erdmann, M. Behrisch, and L. Bieker, “Recent development and applications of SUMO - Simulation of Urban MObility,”
2012
Cited alongside, same era.
T. Wongpiromsarn, T. Uthaicharoenpong, Y. Wang, E. Frazzoli, and D. Wang, “Distributed traffic signal control for maximum network throughput,” in
2012
Cited alongside, same era.
J. C. Medina and R. F. Benekohal, “Traffic signal control using reinforcement learning and the max-plus algorithm as a coordinating strategy,” in
2012
Cited alongside, same era.
P. Varaiya, “The max-pressure controller for arbitrary networks of signalized intersections,” in
2013
Cited alongside, same era.
S.-B. Cools, C. Gershenson, and B. D’Hooghe, “Self-organizing traffic lights: A realistic simulation,” in
2013
Cited alongside, same era.
E. Ricalde and W. Banzhaf, “Evolving adaptive traffic signal controllers for a real scenario using genetic programming with an epigenetic mechanism,” in
2017
Later among the works it cites.
S. Darmoul, S. Elkosantini, A. Louati, and L. B. Said, “Multi-agent immune networks to control interrupted flow at signalized intersections,”
2017
Later among the works it cites.
2017
Later among the works it cites.
M. Aslani, M. S. Mesgari, and M. Wiering, “Adaptive traffic signal control with actor-critic methods in a real-world traffic network with different traffic disruption events,”
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S. El-Tantawy, B. Abdulhai, and H. Abdelgawad, “Multiagent reinforcement learning for integrated network of adaptive traffic signal controllers (marlin-atsc): methodology and large-scale application on downtown toronto,”
2013
Cited alongside, same era.
M. Abdoos, N. Mozayani, and A. L. Bazzan, “Holonic multi-agent system for traffic signals control,”
2013
Cited alongside, same era.
S. El-Tantawy, B. Abdulhai, and H. Abdelgawad, “Design of reinforcement learning parameters for seamless application of adaptive traffic signal control,”
2014
Cited alongside, same era.
M. A. Khamis and W. Gomaa, “Adaptive multi-objective reinforcement learning with hybrid exploration for traffic signal control based on cooperative multi-agent framework,”
2014
Cited alongside, same era.
D. Kingma and J. Ba, “Adam: A method for stochastic optimization,”
2014
Cited alongside, same era.
L. Wu, Y. Ci, J. Chu, and H. Zhang, “The influence of intersections on fuel consumption in urban arterial road traffic: a single vehicle test in harbin, china,”
2015
Cited alongside, same era.
J. Gregoire, X. Qian, E. Frazzoli, A. De La Fortelle, and T. Wongpiromsarn, “Capacity-aware backpressure traffic signal control,”
2015
Cited alongside, same era.
2017
Later among the works it cites.
K.-L. A. Yau, J. Qadir, H. L. Khoo, M. H. Ling, and P. Komisarczuk, “A survey on reinforcement learning models and algorithms for traffic signal control,”
2017
Later among the works it cites.
N. Casas, “Deep deterministic policy gradient for urban traffic light control,”
2017
Later among the works it cites.
W. Liu, G. Qin, Y. He, and F. Jiang, “Distributed cooperative reinforcement learning-based traffic signal control that integrates v2x networks’ dynamic clustering,”
2017
Later among the works it cites.
M. G. Bellemare, W. Dabney, and R. Munos, “A distributional perspective on reinforcement learning,” in
2017
Later among the works it cites.
P.-L. Bacon, J. Harb, and D. Precup, “The option-critic architecture,” in
2017
Later among the works it cites.
2017
Later among the works it cites.
G. Cookson, “INRIX global traffic scorecard,” INRIX, Tech. Rep., 2018
2018
Later among the works it cites.
X. Li and J.-Q. Sun, “Signal multiobjective optimization for urban traffic network,”
2018
Later among the works it cites.
A. Louati, S. Darmoul, S. Elkosantini, and L. ben Said, “An artificial immune network to control interrupted flow at a signalized intersection,”
2018
Later among the works it cites.
2018
Later among the works it cites.
W. Genders, “Deep reinforcement learning adaptive traffic signal control,” Ph.D. dissertation, McMaster University, 2018
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.
W. Dabney, M. Rowland, M. G. Bellemare, and R. Munos, “Distributional reinforcement learning with quantile regression,” in
2018
Later among the works it cites.
2018
Later among the works it cites.
——, “A deep reinforcement learning network for traffic light cycle control,”
2019
Closest in time.
S. Wang, X. Xie, K. Huang, J. Zeng, and Z. Cai, “Deep reinforcement learning-based traffic signal control using high-resolution event-based data,”
2019
Closest in time.
W. Genders and S. Razavi, “Asynchronous n-step q-learning adaptive traffic signal control,”
2019
Closest in time.
T. Chu, J. Wang, L. Codecà, and Z. Li, “Multi-agent deep reinforcement learning for large-scale traffic signal control,”
2019
Closest in time.
P. Tabor, 2019. [Online]. Available:
2019
Closest in time.