V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu, “Asynchronous methods for deep reinforcement learning,” in ICML , 2016, pp. 1928–1937
1937
Earlier work this paper cites.
P. J. Dejax and T. G. Crainic, “Survey paper—a review of empty flows and fleet management models in freight transportation,” Transportation science , vol. 21, no. 4, pp. 227–248, 1987
1987
Earlier work this paper cites.
M. Tan, “Multi-agent reinforcement learning: Independent vs. cooperative agents,” in ICML , 1993, pp. 330–337
1993
Earlier work this paper cites.
R. S. Sutton and A. G. Barto, Reinforcement learning: An introduction . MIT press Cambridge, 1998, vol. 1, no. 1
1998
Earlier work this paper cites.
G. A. Godfrey and W. B. Powell, “An adaptive dynamic programming algorithm for dynamic fleet management, i: Single period travel times,” Transportation Science , vol. 36, no. 1, pp. 21–39, 2002
2002
Earlier work this paper cites.
——, “An adaptive dynamic programming algorithm for dynamic fleet management, ii: Multiperiod travel times,” Transportation Science , vol. 36, no. 1, pp. 40–54, 2002
2002
Earlier work this paper cites.
L. Busoniu, R. Babuska, and B. De Schutter, “A comprehensive survey of multiagent reinforcement learning,” IEEE Transactions on Systems, Man, And Cybernetics-Part C: Applications and Reviews, 38 (2), 2008 , 2008
2008
Earlier work this paper cites.
B. Bakker, S. Whiteson, L. Kester, and F. C. Groen, “Traffic light control by multiagent reinforcement learning systems,” in Interactive Collaborative Information Systems . Springer, 2010, pp. 475–510
2010
Earlier work this paper cites.
K. T. Seow, N. H. Dang, and D.-H. Lee, “A collaborative multiagent taxi-dispatch system,” IEEE T-ASE , vol. 7, no. 3, pp. 607–616, 2010
2010
Earlier work this paper cites.
M. Maciejewski and K. Nagel, “The influence of multi-agent cooperation on the efficiency of taxi dispatching,” in PPAM . Springer, 2013, pp. 751–760
2013
Earlier work this paper cites.
M. G. Bellemare, Y. Naddaf, J. Veness, and M. Bowling, “The arcade learning environment: An evaluation platform for general agents.” J. Artif. Intell. Res.(JAIR) , vol. 47, pp. 253–279, 2013
2013
Earlier work this paper cites.