Fetching the paper…
Reading the bibliography…
Multi-agent reinforcement learning (MARL) remains difficult to scale to many agents.
Topics in propagation of chaos
Alain-Sol Sznitman · 1991
Earlier work this paper cites.
Discrete-time Markov control processes with discounted unbounded costs: optimality criteria
Onésimo Hernández-Lerma and Myriam Muñoz de Ozak · 1992
Earlier work this paper cites.
Constructive approximation , volume 303
Ronald A DeVore and George G Lorentz · 1993
Earlier work this paper cites.
Multi-agent reinforcement learning: Independent vs. cooperative agents
Ming Tan · 1993
Earlier work this paper cites.
Inductive reasoning and bounded rationality
W Brian Arthur · 1994
Earlier work this paper cites.
Policy gradient methods for reinforcement learning with function approximation
Richard S Sutton, David McAllester, Satinder Singh, and Yishay Mansour · 1999
Earlier work this paper cites.
The complexity of decentralized control of Markov decision processes
Daniel S Bernstein, Robert Givan, Neil Immerman, and Shlomo Zilberstein · 2002
Earlier work this paper cites.
Probability measures on metric spaces , volume 352
Kalyanapuram Rangachari Parthasarathy · 2005
Earlier work this paper cites.
Large population stochastic dynamic games: closed-loop McKean-Vlasov systems and the Nash certainty equivalence principle
Minyi Huang, Roland P Malhamé, and Peter E Caines · 2006
Earlier work this paper cites.
Mean field games
Jean-Michel Lasry and Pierre-Louis Lions · 2007
Earlier work this paper cites.
Multi-agent reinforcement learning: A selective overview of theories and algorithms
Kaiqing Zhang, Zhuoran Yang, and Tamer Başar · 2007
Earlier work this paper cites.
The complexity of computing a Nash equilibrium
Constantinos Daskalakis, Paul W Goldberg, and Christos H Papadimitriou · 2009
Earlier work this paper cites.
Optimal transport: old and new , volume 338
Cédric Villani · 2009
Earlier work this paper cites.
Continuous time Markov chain models for chemical reaction networks
David F Anderson and Thomas G Kurtz · 2011
Earlier work this paper cites.
A mean field approach for optimization in discrete time
Nicolas Gast and Bruno Gaujal · 2011
Earlier work this paper cites.
Discrete-time Markov control processes: basic optimality criteria , volume 30
Onésimo Hernández-Lerma and Jean B Lasserre · 2012
Earlier work this paper cites.
Nash, social and centralized solutions to consensus problems via mean field control theory
Mojtaba Nourian, Peter E Caines, Roland P Malhame, and Minyi Huang · 2012
Earlier work this paper cites.
Mean field games and mean field type control theory , volume 101
Alain Bensoussan, Jens Frehse, and Phillip Yam · 2013
Earlier work this paper cites.
Convergence of probability measures
Patrick Billingsley · 2013
Earlier work this paper cites.
Swarm robotics: A review from the swarm engineering perspective
Manuele Brambilla, Eliseo Ferrante, Mauro Birattari, and Marco Dorigo · 2013
Earlier work this paper cites.
ϵ \epsilon -Nash mean field game theory for nonlinear stochastic dynamical systems with major and minor agents
Mojtaba Nourian and Peter E Caines · 2013
Earlier work this paper cites.
Mean field games with partially observed major player and stochastic mean field
Nevroz Şen and Peter E Caines · 2014
Earlier work this paper cites.
Deterministic policy gradient algorithms
David Silver, Guy Lever, Nicolas Heess, Thomas Degris, Daan Wierstra, and Martin Riedmiller · 2014
Earlier work this paper cites.
Greg Brockman, Vicki Cheung, Ludwig Pettersson, Jonas Schneider, John Schulman, Jie Tang, and Wojciech Zaremba · 2016
Earlier work this paper cites.
ϵ \epsilon -Nash equilibria for partially observed LQG mean field games with a major player
Peter E Caines and Arman C Kizilkale · 2016
Earlier work this paper cites.
Mean field games with common noise
René Carmona, François Delarue, and Daniel Lacker · 2016
Earlier work this paper cites.
Efficient deployment of multiple unmanned aerial vehicles for optimal wireless coverage
Mohammad Mozaffari, Walid Saad, Mehdi Bennis, and Mérouane Debbah · 2016
Earlier work this paper cites.
Mean field game theory with a partially observed major agent
Nevroz Şen and Peter E Caines · 2016
Earlier work this paper cites.
Mean-field-type games in engineering
Boualem Djehiche, Alain Tcheukam, and Hamidou Tembine · 2017
Earlier work this paper cites.
Cooperative multi-agent control using deep reinforcement learning
Jayesh K Gupta, Maxim Egorov, and Mykel Kochenderfer · 2017
Earlier work this paper cites.
Random measures, theory and applications , volume 1
Olav Kallenberg · 2017
Cited alongside, same era.
Mathematics of Epidemics on Networks: From Exact to Approximate Models , volume 46
István Z Kiss, Joel C Miller, and Péter L Simon · 2017
Cited alongside, same era.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Cited alongside, same era.
Probabilistic Theory of Mean Field Games with Applications I-II
René Carmona and François Delarue · 2018
Cited alongside, same era.
A refined mean field approximation
Nicolas Gast and Benny Van Houdt · 2018
Cited alongside, same era.
RLlib: Abstractions for distributed reinforcement learning
Eric Liang, Richard Liaw, Robert Nishihara, Philipp Moritz, Roy Fox, Ken Goldberg, Joseph Gonzalez, Michael Jordan, and Ion Stoica · 2018
Benchmarking multi-agent deep reinforcement learning algorithms in cooperative tasks
Georgios Papoudakis, Filippos Christianos, Lukas Schäfer, and Stefano V Albrecht · 2021
Later among the works it cites.
Critical nodes in graphon mean field games
Rinel Foguen Tchuendom, Peter E Caines, and Minyi Huang · 2021
Later among the works it cites.
Attacks on formation control for multiagent systems
Yue Yang, Yang Xiao, and Tieshan Li · 2021
Later among the works it cites.
Correlated equilibria for mean field games with progressive strategies
Ofelia Bonesini, Luciano Campi, and Markus Fischer · 2022
Later among the works it cites.
Solving n-player dynamic routing games with congestion: A mean-field approach
Theophile Cabannes, Mathieu Laurière, Julien Perolat, Raphael Marinier, Sertan Girgin, Sarah Perrin, Olivier Pietquin, Alexandre M Bayen, Eric Goubault, and Romuald Elie · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
En route truck–drone parcel delivery for optimal vehicle routing strategies
Mario Marinelli, Leonardo Caggiani, Michele Ottomanelli, and Mauro Dell’Orco · 2018
Cited alongside, same era.
Bellman equation and viscosity solutions for mean-field stochastic control problem
Huyên Pham and Xiaoli Wei · 2018
Cited alongside, same era.
Markov–Nash equilibria in mean-field games with discounted cost
Naci Saldi, Tamer Başar, and Maxim Raginsky · 2018
Cited alongside, same era.
Mean field multi-agent reinforcement learning
Yaodong Yang, Rui Luo, Minne Li, Ming Zhou, Weinan Zhang, and Jun Wang · 2018
Cited alongside, same era.
Graphon mean field games and the GMFG equations: ε \varepsilon -Nash equilibria
Peter E Caines and Minyi Huang · 2019
Cited alongside, same era.
Learning mean-field games
Xin Guo, Anran Hu, Renyuan Xu, and Junzi Zhang · 2019
Cited alongside, same era.
Correlated equilibria and mean field games: a simple model
Luciano Campi and Markus Fischer · 2022
Later among the works it cites.
Learning graphon mean field games and approximate Nash equilibria
Kai Cui and Heinz Koeppl · 2022
Later among the works it cites.
A general framework for learning mean-field games
Xin Guo, Anran Hu, Renyuan Xu, and Junzi Zhang · 2022
Later among the works it cites.
Learning mean field games: A survey
Mathieu Laurière, Sarah Perrin, Matthieu Geist, and Olivier Pietquin · 2022
Later among the works it cites.
Xin Liu, Honghao Wei, and Lei Ying · 2022
Later among the works it cites.
On the approximation of cooperative heterogeneous multi-agent reinforcement learning (MARL) using mean field control (MFC)
Washim Uddin Mondal, Mridul Agarwal, Vaneet Aggarwal, and Satish V Ukkusuri · 2022
Later among the works it cites.
Mean-field Markov decision processes with common noise and open-loop controls
Médéric Motte and Huyên Pham · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
Long Ouyang, Jeff Wu, Xu Jiang, Diogo Almeida, Carroll L Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al · 2022
Later among the works it cites.
Scaling mean field games by online mirror descent
Julien Pérolat, Sarah Perrin, Romuald Elie, Mathieu Laurière, Georgios Piliouras, Matthieu Geist, Karl Tuyls, and Olivier Pietquin · 2022
Later among the works it cites.
Generalization in mean field games by learning master policies
Sarah Perrin, Mathieu Laurière, Julien Pérolat, Romuald Élie, Matthieu Geist, and Olivier Pietquin · 2022
Later among the works it cites.
Decentralized mean field games
Sriram Ganapathi Subramanian, Matthew E Taylor, Mark Crowley, and Pascal Poupart · 2022
Later among the works it cites.
The surprising effectiveness of PPO in cooperative multi-agent games
Chao Yu, Akash Velu, Eugene Vinitsky, Jiaxuan Gao, Yu Wang, Alexandre Bayen, and Yi Wu · 2022
Later among the works it cites.
A unified algebraic perspective on Lipschitz neural networks
Alexandre Araujo, Aaron J Havens, Blaise Delattre, Alexandre Allauzen, and Bin Hu · 2023
Closest in time.
Mean field Markov decision processes
Nicole Bäuerle · 2023
Closest in time.
Dynamic programming principles for mean-field controls with learning
Haotian Gu, Xin Guo, Xiaoli Wei, and Renyuan Xu · 2023
Closest in time.
Local lipschitz bounds of deep neural networks
Calypso Herrera, Florian Krach, and Josef Teichmann · 2023
Closest in time.
Mean-field control based approximation of multi-agent reinforcement learning in presence of a non-decomposable shared global state
Washim Uddin Mondal, Vaneet Aggarwal, and Satish Ukkusuri · 2023
Closest in time.
Quantitative propagation of chaos for mean field Markov decision process with common noise
Médéric Motte and Huyên Pham · 2023
Closest in time.
Efficient model-based multi-agent mean-field reinforcement learning
Barna Pásztor, Andreas Krause, and Ilija Bogunovic · 2023
Closest in time.
Leveraging connected and automated vehicles for participatory traffic control
Minghui Wu, Xingmin Wang, Yafeng Yin, Henry Liu, Ben Wang, Jerome P Lynch, et al · 2023
Closest in time.
Policy mirror ascent for efficient and independent learning in mean field games
Batuhan Yardim, Semih Cayci, Matthieu Geist, and Niao He · 2023
Closest in time.
Learning decentralized partially observable mean field control for artificial collective behavior
Kai Cui, Sascha H. Hauck, Christian Fabian, and Heinz Koeppl · 2024
Closest in time.
Zero-sum games between mean-field teams: Reachability-based analysis under mean-field sharing
Yue Guan, Mohammad Afshari, and Panagiotis Tsiotras · 2024
Closest in time.
Scalable multi-agent reinforcement learning for networked systems with average reward
Guannan Qu, Yiheng Lin, Adam Wierman, and Na Li · 2086
Closest in time.