Fetching the paper…
Reading the bibliography…
Recent years have seen a great increase in the capacity and parallel processing power of data centers and cloud services.
Approximately solving mean field games via entropy-regularized deep reinforcement learning. In International Conference on Artificial Intelligence and Statistics . PMLR, 1909–1917
Kai Cui and Heinz Koeppl. 2021 · 1917
Earlier work this paper cites.
The journal of physical chemistry 81, 25 (1977), 2340–2361
Daniel T Gillespie. 1977 · 1977
Earlier work this paper cites.
Optimality of the shortest line discipline
Wayne Winston. 1977 · 1977
Earlier work this paper cites.
Deciding which queue to join: Some counterexamples
Ward Whitt. 1986 · 1986
Earlier work this paper cites.
Joining the right queue: A Markov decision-rule. In 26th IEEE Conference on Decision and Control , Vol. 26. IEEE, 1863–1868
KR Krishnan. 1987 · 1987
Earlier work this paper cites.
A survey of Markov decision models for control of networks of queues
Shaler Stidham and Richard Weber. 1993 · 1993
Earlier work this paper cites.
How useful is old information?
Michael Mitzenmacher. 2000 · 2000
Earlier work this paper cites.
The power of two choices in randomized load balancing
Michael Mitzenmacher. 2001 · 2001
Earlier work this paper cites.
Balancing queues by mean field interaction
Donald A Dawson, Jiashan Tang, and Yiqiang Q Zhao. 2005 · 2005
Earlier work this paper cites.
Conditional strong law of large number
Dariusz Majerek, Wioletta Nowak, and Wieslaw Zieba. 2005 · 2005
Earlier work this paper cites.
Large population stochastic dynamic games: closed-loop McKean-Vlasov systems and the Nash certainty equivalence principle
Minyi Huang, Roland P Malhamé, Peter E Caines, et al · 2006
Earlier work this paper cites.
Mean field games
Jean-Michel Lasry and Pierre-Louis Lions. 2007 · 2007
Earlier work this paper cites.
A maximum principle for SDEs of mean-field type
Daniel Andersson and Boualem Djehiche. 2011 · 2011
Earlier work this paper cites.
Discrete-time Markov control processes: basic optimality criteria . Vol. 30
Onésimo Hernández-Lerma and Jean B Lasserre. 2012 · 2012
Earlier work this paper cites.
Mean field games and mean field type control theory . Vol. 101
Alain Bensoussan, Jens Frehse, Phillip Yam, et al · 2013
Earlier work this paper cites.
Reinforcement learning in robotics: A survey
Jens Kober, J Andrew Bagnell, and Jan Peters. 2013 · 2013
Cited alongside, same era.
Team optimal control of coupled subsystems with mean-field sharing. In 53rd IEEE Conference on Decision and Control . IEEE, 1669–1674
Jalal Arabneydi and Aditya Mahajan. 2014 · 2014
Cited alongside, same era.
Markov decision processes: discrete stochastic dynamic programming
Martin L Puterman. 2014 · 2014
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Cited alongside, same era.
A concise introduction to decentralized POMDPs
Frans A Oliehoek and Christopher Amato. 2016 · 2016
Cited alongside, same era.
Steady-state analysis of shortest expected delay routing
Learning mean-field games. In Advances in Neural Information Processing Systems . 4966–4976
Xin Guo, Anran Hu, Renyuan Xu, and Junzi Zhang. 2019 · 2019
Later among the works it cites.
An overview for Markov decision processes in queues and networks. In International Conference of Celebrating Professor Jinhua Cao’s 80th Birthday . Springer, 44–71
Quan-Lin Li, Jing-Yu Ma, Rui-Na Fan, and Li Xia. 2019 · 2019
Later among the works it cites.
Open problem—load balancing using delayed information
David Lipshutz. 2019 · 2019
Later among the works it cites.
Applications of deep reinforcement learning in communications and networking: A survey
Nguyen Cong Luong, Dinh Thai Hoang, Shimin Gong, Dusit Niyato, Ping Wang, Ying-Chang Liang, and Dong In Kim. 2019 · 2019
Later among the works it cites.
Reinforcement learning in stationary mean-field games. In Proceedings of the 18th International Conference on Autonomous Agents and MultiAgent Systems . 251–259
Jayakumar Subramanian and Aditya Mahajan. 2019 · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jori Selen, Ivo Adan, Stella Kapodistria, and Johan van Leeuwaarden. 2016 · 2016
Cited alongside, same era.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov. 2017 · 2017
Cited alongside, same era.
RLlib: Abstractions for distributed reinforcement learning. In International Conference on Machine Learning . PMLR, 3053–3062
Eric Liang, Richard Liaw, Robert Nishihara, Philipp Moritz, Roy Fox, Ken Goldberg, Joseph Gonzalez, Michael Jordan, and Ion Stoica. 2018 · 2018
Cited alongside, same era.
Universality of power-of-d load balancing in many-server systems
Debankur Mukherjee, Sem C Borst, Johan SH Van Leeuwaarden, and Philip A Whiting. 2018 · 2018
Cited alongside, same era.
Markov–Nash Equilibria in Mean-Field Games with Discounted Cost
Naci Saldi, Tamer Basar, and Maxim Raginsky. 2018 · 2018
Cited alongside, same era.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto. 2018 · 2018
Cited alongside, same era.
Scalable load balancing in networked systems: A survey of recent advances
Mark van der Boor, Sem C Borst, Johan SH van Leeuwaarden, and Debankur Mukherjee. 2018 · 2018
Cited alongside, same era.
Later among the works it cites.
Hyper-scalable JSQ with sparse feedback
Mark van der Boor, Sem Borst, and Johan van Leeuwaarden. 2019 · 2019
Later among the works it cites.
Power-of-d-choices with memory: Fluid limit and optimality
Jonatha Anselmi and Francois Dufour. 2020 · 2020
Later among the works it cites.
Machine Learning for Communications
Vaneet Aggarwal. 2021 · 2021
Later among the works it cites.
Discrete-Time Mean Field Control with Environment States. In 2021 60th IEEE Conference on Decision and Control (CDC) . 5239–5246
Kai Cui, Anam Tahir, Mark Sinzger, and Heinz Koeppl. 2021 · 2021
Later among the works it cites.
Mean-field controls with Q-learning for cooperative MARL: convergence and complexity analysis
Haotian Gu, Xin Guo, Xiaoli Wei, and Renyuan Xu. 2021 · 2021
Later among the works it cites.
Washim Uddin Mondal, Mridul Agarwal, Vaneet Aggarwal, and Satish V Ukkusuri. 2021 · 2021
Later among the works it cites.
Multi-agent reinforcement learning: A selective overview of theories and algorithms
Kaiqing Zhang, Zhuoran Yang, and Tamer Başar. 2021 · 2021
Later among the works it cites.
Asymptotically optimal load balancing in large-scale heterogeneous systems with multiple dispatchers
Xingyu Zhou, Ness Shroff, and Adam Wierman. 2021 · 2021
Later among the works it cites.
McKean–Vlasov optimal control: the dynamic programming principle
Mao Fabrice Djete, Dylan Possamaï, and Xiaolu Tan. 2022 · 2022
Closest in time.