Fetching the paper…
Reading the bibliography…
We study a multi-agent mean field type control problem in discrete time where the agents aim to find a socially optimal strategy and where the state and action spaces for the agents are assumed to be continuous.
Brown LD, Purves R (1973) Measurable selections of extrema. The annals of statistics pp 902–912
1973
Earlier work this paper cites.
Hinderer K (2005) Lipschitz continuity of value functions in markovian decision processes. Mathematical Methods of Operations Research 62(1):3–22
2005
Earlier work this paper cites.
Huang M, Malhamé RP, Caines PE (2006) Large population stochastic dynamic games: closed-loop mckean-vlasov systems and the nash certainty equivalence principle. Communications in Information & Systems 6(3):221–252
2006
Earlier work this paper cites.
Huang M, Caines PE, Malhamé RP (2007) Large-population cost-coupled LQG problems with nonuniform agents: individual-mass behavior and decentralized epsilon -Nash equilibria. IEEE transactions on automatic control 52(9):1560–1571
2007
Earlier work this paper cites.
Lasry JM, Lions PL (2007) Mean field games. Japanese journal of mathematics 2(1):229–260
2007
Earlier work this paper cites.
Bensoussan A, Frehse J, Yam P (2013) Mean field games and mean field type control theory, vol 101. Springer
2013
Earlier work this paper cites.
Carmona R, Delarue F (2013) Probabilistic analysis of mean-field games. SIAM Journal on Control and Optimization 51(4):2705–2734
2013
Earlier work this paper cites.
Tembine H, Zhu Q, Başar T (2013) Risk-sensitive mean-field games. IEEE Transactions on Automatic Control 59(4):835–850
2013
Earlier work this paper cites.
Gomes DA, Saúde J (2014) Mean field games models—a brief survey. Dynamic Games and Applications 4(2):110–154
2014
Earlier work this paper cites.
Laurière M, Pironneau O (2014) Dynamic programming for mean-field type control. Comptes Rendus Mathematique 352(9):707–713
2014
Earlier work this paper cites.
Lacker D (2017) Limit theory for controlled McKean–Vlasov dynamics. SIAM Journal on Control and Optimization 55(3):1641–1672
2017
Earlier work this paper cites.
Pham H, Wei X (2017) Dynamic programming for optimal control of stochastic McKean–Vlasov dynamics. SIAM Journal on Control and Optimization 55(2):1069–1101
2017
Earlier work this paper cites.
Bayraktar E, Cosso A, Pham H (2018) Randomized dynamic programming principle and Feynman-Kac representation for optimal control of Mckean-Vlasov dynamics. Transactions of the American Mathematical Society 370(3):2115–2160
2018
Cited alongside, same era.
Carmona R, Laurière M, Tan Z (2019) Model-free mean-field reinforcement learning: mean-field MDP and mean-field Q-learning. arXiv preprint arXiv:191012802
2019
Cited alongside, same era.
Fornasier M, Lisini S, Orrieri C, et al (2019) Mean-field optimal control as gamma-limit of finite agent controls. European Journal of Applied Mathematics 30(6):1153–1186
2019
Cited alongside, same era.
Guo X, Hu A, Xu R, et al (2019) Learning mean-field games. Advances in Neural Information Processing Systems 32
2019
Cited alongside, same era.
Saldi N, Başar T, Raginsky M (2019) Approximate nash equilibria in partially observed stochastic games with mean-field interactions. Mathematics of Operations Research 44(3):1006–1033
Sanjari S, Yuksel S (2021) Optimal policies for convex symmetric stochastic dynamic teams and their mean-field limit. SIAM Journal on Control and Optimization 59(2):777–804
2021
Later among the works it cites.
Anahtarci B, Kariksiz CD, Saldi N (2022) Q-learning in regularized mean-field games. Dynamic Games and Applications pp 1–29
2022
Closest in time.
Angiuli A, Fouque JP, Laurière M (2022) Unified reinforcement q-learning for mean field game and control problems. Mathematics of Control, Signals, and Systems 34(2):217–271
2022
Closest in time.
Germain M, Mikael J, Warin X (2022) Numerical resolution of Mckean-Vlasov FBSDEs using neural networks. Methodology and Computing in Applied Probability pp 1–30
2022
Closest in time.
Motte M, Pham H (2022) Mean-field Markov decision processes with common noise and open-loop controls. The Annals of Applied Probability 32(2):1421–1458
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
Subramanian J, Mahajan A (2019) Reinforcement learning in stationary mean-field games. In: Proceedings of the 18th International Conference on Autonomous Agents and MultiAgent Systems, pp 251–259
2019
Cited alongside, same era.
Elie R, Perolat J, Laurière M, et al (2020) On the convergence of model free learning in mean field games. In: Proceedings of the AAAI Conference on Artificial Intelligence, pp 7143–7150
2020
Cited alongside, same era.
Fu Z, Yang Z, Chen Y, et al (2020) Actor-critic provably finds Nash equilibria of linear-quadratic mean-field games URL https://openreview.net/forum?id=H1lhqpEYPr
2020
Cited alongside, same era.
Perrin S, Pérolat J, Laurière M, et al (2020) Fictitious play for mean field games: Continuous time analysis and applications. Advances in Neural Information Processing Systems 33:13199–13213
2020
Cited alongside, same era.
Wang L, Yang Z, Wang Z (2020) Breaking the curse of many agents: Provable mean embedding q-iteration for mean-field reinforcement learning. In: International conference on machine learning, PMLR, pp 10092–10103
2020
Cited alongside, same era.
Carmona R, Laurière M (2021) Convergence analysis of machine learning algorithms for the numerical solution of mean field control and games i: the ergodic case. SIAM Journal on Numerical Analysis 59(3):1455–1485
2021
Cited alongside, same era.
Gu H, Guo X, Wei X, et al (2021) Mean-field controls with q-learning for cooperative marl: convergence and complexity analysis. SIAM Journal on Mathematics of Data Science 3(4):1168–1196
2021
Cited alongside, same era.
2022
Closest in time.
Sanjari S, Saldi N, Yüksel S (2022) Optimality of independently randomized symmetric policies for exchangeable stochastic teams with infinitely many decision makers. Mathematics of Operations Research
2022
Closest in time.
Bäuerle N (2023) Mean field Markov decision processes. Applied Mathematics & Optimization 88(1):12
2023
Closest in time.
Bayraktar E, Zhang X (2023) Solvability of infinite horizon Mckean–Vlasov FBSDEs in mean field control problems and games. Applied Mathematics & Optimization 87(1):13
2023
Closest in time.
Bayraktar E, Cecchin A, Chakraborty P (2023) Mean field control and finite agent approximation for regime-switching jump diffusions. Applied Mathematics & Optimization 88(2):36
2023
Closest in time.
Gu H, Guo X, Wei X, et al (2023) Dynamic programming principles for mean-field controls with learning. Operations Research
2023
Closest in time.
Motte M, Pham H (2023) Quantitative propagation of chaos for mean field Markov decision process with common noise. Electronic Journal of Probability 28:1–24
2023
Closest in time.
Pásztor B, Krause A, Bogunovic I (2023) Efficient model-based multi-agent mean-field reinforcement learning. Transactions on Machine Learning Research URL https://openreview.net/forum?id=gvcDSDYUZx
2023
Closest in time.