Fetching the paper…
Reading the bibliography…
This paper aims to develop a paradigm that models the learning behavior of intelligent agents (including but not limited to autonomous vehicles, connected and automated vehicles, or human-driven vehicles with intelligent navigation systems where human drivers follow the navigation instructions completely) with a utility-optimizing goal and the system's equilibrating processes in a routing game among atomic selfish agents.
Multi-Agent Connected Autonomous Driving using Deep Reinforcement Learning
Palanisamy, P., 2019 · 1911
Earlier work this paper cites.
The cell transmission model: A dynamic representation of highway traffic consistent with the hydrodynamic theory
Daganzo, C.F., 1994 · 1994
Earlier work this paper cites.
Markov Games As a Framework for Multi-agent Reinforcement Learning, in: Proceedings of the Eleventh International Conference on International Conference on Machine Learning, Morgan Kaufmann Publishers Inc., San Francisco, CA, USA. pp. 157–163
Littman, M.L., 1994 · 1994
Earlier work this paper cites.
Markov Decision Processes: Discrete Stochastic Dynamic Programming
Puterman, M.L., 1994 · 1994
Earlier work this paper cites.
The cell transmission model, part II: Network traffic
Daganzo, C.F., 1995 · 1995
Earlier work this paper cites.
Dynamic user optimal traffic assignment model for many to one travel demand
Lam, W.H., Huang, H.J., 1995 · 1995
Earlier work this paper cites.
Applications and special classes of stochastic games, in: Competitive Markov Decision Processes. Springer, pp. 301–341
Filar, J., Vrieze, K., 1997 · 1997
Earlier work this paper cites.
Decomposition of the reactive dynamic assignments with queues for a many-to-many origin-destination pattern
Kuwahara, M., Akamatsu, T., 1997 · 1997
Earlier work this paper cites.
An Iterative Algorithm to Determine the Dynamic User Equilibrium in a Traffic Simulation Model
Gawron, C., 1998 · 1998
Earlier work this paper cites.
Multiagent Reinforcement Learning: Theoretical Framework and an Algorithm, in: Proceedings of the Fifteenth International Conference on Machine Learning, Morgan Kaufmann Publishers Inc., San Francisco, CA, USA. pp. 242–250
Hu, J., Wellman, M.P., 1998 · 1998
Earlier work this paper cites.
Introduction to Reinforcement Learning
Sutton, R.S., Barto, A.G., 1998 · 1998
Earlier work this paper cites.
A reactive dynamic user equilibrium model in network with queues
Li, J., Fujiwara, O., Kawakami, S., 2000 · 2000
Earlier work this paper cites.
A Linear Programming Model for the Single Destination System Optimum Dynamic Traffic Assignment Problem
Ziliaskopoulos, A.K., 2000 · 2000
Earlier work this paper cites.
Value-function reinforcement learning in markov games
Littman, M.L., 2001 · 2001
Earlier work this paper cites.
A comparative study of some macroscopic link models used in dynamic traffic assignment
Nie, X., Zhang, H.M., 2005 · 2005
Earlier work this paper cites.
Optimal vehicle routing with real-time traffic information
Seongmoon Kim, Lewis, M.E., White, C.C., 2005 · 2005
Earlier work this paper cites.
Routing games
Roughgarden, T., 2007 · 2007
Earlier work this paper cites.
The link transmission model for dynamic network loading URL: https://www.researchgate.net/publication/28360292_The_Link_Transmission_Model_for_dynamic_network_loading
Yperman, I., 2007 · 2007
Earlier work this paper cites.
Re-routing Agents in an Abstract Traffic Scenario, in: Zaverucha, G., da Costa, A.L. (Eds.), Advances in Artificial Intelligence - SBIA 2008, Springer, Berlin, Heidelberg. pp. 63–72
Bazzan, A.L.C., Klügl, F., 2008 · 2008
Earlier work this paper cites.
Traffic Light Control by Multiagent Reinforcement Learning Systems, in: Babuška, R., Groen, F.C.A. (Eds.), Interactive Collaborative Information Systems. Springer, Berlin, Heidelberg. Studies in Computational Intelligence, pp. 475–510
Bakker, B., Whiteson, S., Kester, L., Groen, F.C.A., 2010 · 2010
Earlier work this paper cites.
A cell-based merchant-nemhauser model for the system optimum dynamic traffic assignment problem
Nie, Y.M., 2011 · 2011
Earlier work this paper cites.
Modelling Transport
Ortuzar, J.d.D., Willumsen, L., 2011 · 2011
Cited alongside, same era.
Dynamic network loading: a stochastic differentiable model that derives link state distributions
Osorio, C., Flötteröd, G., Bierlaire, M., 2011 · 2011
Cited alongside, same era.
Continuous-time point-queue models in dynamic network loading
Ban, X.J., Pang, J.S., Liu, H.X., Ma, R., 2012 · 2012
Cited alongside, same era.
Independent reinforcement learners in cooperative Markov games: a survey regarding coordination problems
Matignon, L., Laurent, G.J., Fort-Piat, N.L., 2012 · 2012
Cited alongside, same era.
Dynamic user equilibrium based on a hydrodynamic model
Friesz, T.L., Han, K., Neto, P.A., Meimand, A., Yao, T., 2013 · 2013
Cited alongside, same era.
Modelling network flow with and without link interactions: The cases of point queue, spatial queue and cell transmission model
Efficient Large-Scale Fleet Management via Multi-Agent Deep Reinforcement Learning, in: Proceedings of the 24th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining, ACM, New York, NY, USA. pp. 1774–1783
Lin, K., Zhao, R., Xu, Z., Zhou, J., 2018 · 2018
Later among the works it cites.
A reinforcement learning framework for the adaptive routing problem in stochastic time-dependent network
Mao, C., Shen, Z., 2018 · 2018
Later among the works it cites.
Openai five
OpenAI, 2018 · 2018
Later among the works it cites.
Analysing the impact of travel information for minimising the regret of route choice
Ramos, G.d.O., Bazzan, A.L.C., da Silva, B.C., 2018 · 2018
Later among the works it cites.
Mean Field Multi-Agent Reinforcement Learning, in: International Conference on Machine Learning, pp. 5571–5580
Yang, Y., Luo, R., Li, M., Zhou, M., Zhang, W., Wang, J., 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Zhang, H., Nie, Y., Qian, S., 2013 · 2013
Cited alongside, same era.
Individual versus Difference Rewards on Reinforcement Learning for Route Choice, in: 2014 Brazilian Conference on Intelligent Systems, pp. 253–258
Grunitzki, R., Ramos, G.d.O., Bazzan, A.L.C., 2014 · 2014
Cited alongside, same era.
Continuous-time dynamic system optimum for single-destination traffic networks with queue spillbacks
Ma, R., Ban, X.J., Pang, J.S., 2014 · 2014
Cited alongside, same era.
Human-level control through deep reinforcement learning
Mnih, V., Kavukcuoglu, K., Silver, D., Rusu, A.A., Veness, J., Bellemare, M.G., Graves, A., Riedmiller, M., Fidjeland, A.K., Ostrovski, G., Petersen, S., Beattie, C., Sadik, A., Antonoglou, I., King, H., Kumaran, D., Wierstra, D., Legg, S., Hassabis, D., 2015 · 2015
Cited alongside, same era.
Stochastic games
Solan, E., Vieille, N., 2015 · 2015
Cited alongside, same era.
A multiagent reinforcement learning approach to en-route trip building, in: 2016 International Joint Conference on Neural Networks (IJCNN), pp. 5288–5295
Bazzan, A.L.C., Grunitzki, R., 2016 · 2016
Cited alongside, same era.
Solving the Dynamic Vehicle Routing Problem Under Traffic Congestion
Kim, G., Ong, Y.S., Cheong, T., Tan, P.S., 2016 · 2016
Cited alongside, same era.
Superhuman AI for multiplayer poker
Brown, N., Sandholm, T., 2019 · 2019
Later among the works it cites.
A unified equilibrium framework of new shared mobility systems
Di, X., Ban, X.J., 2019 · 2019
Later among the works it cites.
The mathematical foundations of dynamic user equilibrium
Friesz, T.L., Han, K., 2019 · 2019
Later among the works it cites.
Efficient Ridesharing Order Dispatching with Mean Field Multi-Agent Reinforcement Learning, in: The World Wide Web Conference, ACM, New York, NY, USA. pp. 983–994
Li, M., Qin, Z., Jiao, Y., Yang, Y., Wang, J., Wang, C., Wu, G., Ye, J., 2019 · 2019
Later among the works it cites.
Multi-agent Deep Reinforcement Learning for Zero Energy Communities, in: 2019 IEEE PES Innovative Smart Grid Technologies Europe (ISGT-Europe), pp. 1–5
Prasad, A., Dusparic, I., 2019 · 2019
Later among the works it cites.
Grandmaster level in StarCraft II using multi-agent reinforcement learning
Vinyals, O., Babuschkin, I., Czarnecki, W.M., Mathieu, M., Dudzik, A., Chung, J., Choi, D.H., Powell, R., Ewalds, T., Georgiev, P., Oh, J., Horgan, D., Kroiss, M., Danihelka, I., Huang, A., Sifre, L., Cai, T., Agapiou, J.P., Jaderberg, M., Vezhnevets, A.S., Leblond, R., Pohlen, T., Dalibard, V., Budden, D., Sulsky, Y., Molloy, J., Paine, T.L., Gulcehre, C., Wang, Z., Pfaff, T., Wu, Y., Ring, R., Yogatama, D., Wünsch, D., McKinney, K., Smith, O., Schaul, T., Lillicrap, T., Kavukcuoglu, K., Hassabis, D., Apps, C., Silver, D., 2019 · 2019
Later among the works it cites.
Deep Multi Agent Reinforcement Learning for Autonomous Driving, in: Goutte, C., Zhu, X. (Eds.), Advances in Artificial Intelligence, Springer International Publishing, Cham. pp. 67–78
Bhalla, S., Ganapathi Subramanian, S., Crowley, M., 2020 · 2020
Closest in time.
Toward A Thousand Lights: Decentralized Deep Reinforcement Learning for Large-Scale Traffic Signal Control
Chen, C., Wei, H., Xu, N., Zheng, G., Yang, M., Xiong, Y., Xu, K., Li, Z., 2020 · 2020
Closest in time.
Deep Reinforcement Learning for Multiagent Systems: A Review of Challenges, Solutions, and Applications
Nguyen, T.T., Nguyen, N.D., Nahavandi, S., 2020 · 2020
Closest in time.
Reward design for driver repositioning using multi-agent reinforcement learning
Shou, Z., Di, X., 2020 · 2020
Closest in time.
Optimal passenger-seeking policies on E-hailing platforms using Markov decision process and imitation learning
Shou, Z., Di, X., Ye, J., Zhu, H., Zhang, H., Hampshire, R., 2020 · 2020
Closest in time.
A reinforcement learning scheme for the equilibrium of the in-vehicle route choice problem based on congestion game
Zhou, B., Song, Q., Zhao, Z., Liu, T., 2020 · 2020
Closest in time.
Ridesharing user equilibrium with nodal matching cost and its implications for congestion tolling and platform pricing
Chen, X., Di, X., 2021 · 2021
Closest in time.
A survey on autonomous vehicle control in the era of mixed-autonomy: From physics-based to AI-guided driving policy learning
Di, X., Shi, R., 2021 · 2021
Closest in time.
Dynamic driving and routing games for autonomous vehicles on networks: A mean field game approach
Huang, K., Chen, X., Di, X., Du, Q., 2021 · 2021
Closest in time.