Mnih V, Kavukcuoglu K, Silver D, Rusu AA, Veness J, Bellemare MG, Graves A, Riedmiller M, Fidjeland AK, Ostrovski G, Petersen S, Beattie C, Sadik A, Antonoglou I, King H, Kumaran D, Wierstra D, Legg S, Hassabis D (2015) Human-level control through deep reinforcement learning. Nature 518(7540):529–533
2015
Later among the works it cites.
Pham H, Wei X (2016) Discrete time McKean–Vlasov control problem: a dynamic programming approach. Applied Mathematics & Optimization 74(3):487–506
2016
Later among the works it cites.
Shalev-Shwartz S, Shammah S, Shashua A (2016) Safe, multi-agent, reinforcement learning for autonomous driving. arXiv preprint arXiv:1610.03295
Original
2016
Later among the works it cites.
Silver D, Huang A, Maddison CJ, Guez A, Sifre L, Van Den Driessche G, Schrittwieser J, Antonoglou I, Panneershelvam V, Lanctot M (2016) Mastering the game of go with deep neural networks and tree search. Nature 529(7587):484
2016
Later among the works it cites.
Lacker D (2017) Limit theory for controlled McKean–Vlasov dynamics. SIAM Journal on Control and Optimization 55(3):1641–1672
2017
Later among the works it cites.
Nuño G (2017) Optimal social policies in mean field games. Applied Mathematics & Optimization 76(1):29–57
2017
Later among the works it cites.
Carmona R, Delarue F (2018) Probabilistic Theory of Mean Field Games with Applications I-II (Springer)
2018
Later among the works it cites.
Jin J, Song C, Li H, Gai K, Wang J, Zhang W (2018) Real-time bidding with multi-agent reinforcement learning in display advertising. Proceedings of the 27th ACM International Conference on Information and Knowledge Management , 2193–2201
2018
Later among the works it cites.
Sutton RS, Barto AG (2018) Reinforcement Learning: An Introduction (MIT press)
2018
Later among the works it cites.
Guo X, Hu A, Xu R, Zhang J (2019) Learning mean-field games. Advances in Neural Information Processing Systems , 4966–4976
2019
Closest in time.
Li M, Qin Z, Jiao Y, Yang Y, Wang J, Wang C, Wu G, Ye J (2019) Efficient ridesharing order dispatching with mean field multi-agent reinforcement learning. The World Wide Web Conference , 983–994
2019
Closest in time.
Vinyals O, Babuschkin I, Chung J, Mathieu M, Jaderberg M, Czarnecki WM, Dudzik A, Huang A, Georgiev P, Powell R (2019) Alphastar: Mastering the real-time strategy game starcraft II. DeepMind Blog 2
2019
Closest in time.
Aïd R, Basei M, Pham H (2020) A McKean–Vlasov approach to distributed electricity generation development. Mathematical Methods of Operations Research 91(2):269–310
2020
Closest in time.
Wang L, Yang Z, Wang Z (2020) Breaking the curse of many agents: Provable mean embedding Q-iteration for mean-field reinforcement learning. International Conference on Machine Learning , 10092–10103 (PMLR)
2020
Closest in time.
Gu H, Guo X, Wei X, Xu R (2021) Mean-field controls with Q-learning for cooperative MARL: convergence and complexity analysis. SIAM Journal on Mathematics of Data Science 3(4):1168–1196
2021
Closest in time.