Fetching the paper…
Reading the bibliography…
We address in this paper Reinforcement Learning (RL) among agents that are grouped into teams such that there is cooperation within each team but general-sum (non-zero sum) competition across different teams.
Markov games as a framework for multi-agent reinforcement learning
Michael L Littman · 1994
Earlier work this paper cites.
Intelligent agents for negotiations in market games. i. model
V Krishna and VC Ramesh · 1998
Earlier work this paper cites.
Dynamic Noncooperative Game Theory
Tamer Başar and Geert Jan Olsder · 1998
Earlier work this paper cites.
Recursive Macroeconomic Theory
Thomas J Sargent and Lars Ljungqvist · 2000
Earlier work this paper cites.
Individual and mass behaviour in large population stochastic wireless power control problems: Centralized and Nash equilibrium solutions
Minyi Huang, Peter E Caines, and Roland P Malhamé · 2003
Earlier work this paper cites.
Large population stochastic dynamic games: Closed-loop McKean-Vlasov systems and the Nash certainty equivalence principle
Minyi Huang, Roland P Malhamé, and Peter E Caines · 2006
Earlier work this paper cites.
Jeux à champ moyen. i–le cas stationnaire
Jean-Michel Lasry and Pierre-Louis Lions · 2006
Earlier work this paper cites.
Nash equilibrium seeking in noncooperative games
Paul Frihauf, Miroslav Krstic, and Tamer Başar · 2011
Earlier work this paper cites.
Numerical properties of stochastic linear quadratic model with applications in finance
Ivan Ivanov and Boyan Lomev · 2012
Earlier work this paper cites.
George AF Seber and Alan J Lee · 2012
Earlier work this paper cites.
Mean field games and systemic risk
René Carmona, Jean-Pierre Fouque, and Li-Hsien Sun · 2015
Earlier work this paper cites.
Multi-agent actor-critic for mixed cooperative-competitive environments
Ryan Lowe, Yi I Wu, Aviv Tamar, Jean Harb, OpenAI Pieter Abbeel, and Igor Mordatch · 2017
Earlier work this paper cites.
Global convergence of policy gradient methods for the linear quadratic regulator
Maryam Fazel, Rong Ge, Sham Kakade, and Mehran Mesbahi · 2018
Earlier work this paper cites.
Mean field game of controls and an application to trade crowding
Pierre Cardaliaguet and Charles-Albert Lehalle · 2018
Earlier work this paper cites.
Mean field multi-agent reinforcement learning
Yaodong Yang, Rui Luo, Minne Li, Ming Zhou, Weinan Zhang, and Jun Wang · 2018
Earlier work this paper cites.
Risk-sensitive zero-sum differential games
Jun Moon, Tyrone E Duncan, and Tamer Başar · 2018
Earlier work this paper cites.
Derivative-free methods for policy optimization: Guarantees for linear quadratic systems
Dhruv Malik, Ashwin Pananjady, Kush Bhatia, Koulik Khamaru, Peter Bartlett, and Martin Wainwright · 2019
Earlier work this paper cites.
Provably global convergence of actor-critic: A case for linear quadratic regulator with ergodic cost
Zhuoran Yang, Yongxin Chen, Mingyi Hong, and Zhaoran Wang · 2019
Cited alongside, same era.
Learning mean-field games
Xin Guo, Anran Hu, Renyuan Xu, and Junzi Zhang · 2019
Cited alongside, same era.
Reinforcement learning in stationary mean-field games
Jayakumar Subramanian and Aditya Mahajan · 2019
Cited alongside, same era.
Policy optimization provably converges to Nash equilibria in zero-sum linear quadratic games
Kaiqing Zhang, Zhuoran Yang, and Tamer Başar · 2019
Cited alongside, same era.
Global convergence of policy gradient for sequential zero-sum linear quadratic dynamic games
Jingjing Bu, Lillian J Ratliff, and Mehran Mesbahi · 2019
Cited alongside, same era.
Linear-quadratic mean-field reinforcement learning: convergence of policy gradient methods
Adversarial linear-quadratic mean-field games over multigraphs
Muhammad Aneeq Uz Zaman, Sujay Bhatt, and Tamer Başar · 2021
Later among the works it cites.
Linear-quadratic zero-sum mean-field type games: Optimality conditions and policy optimization
René Carmona, Kenza Hamidouche, Mathieu Laurière, and Zongjun Tan · 2021
Later among the works it cites.
Mean-field controls with Q-learning for cooperative MARL: convergence and complexity analysis
Haotian Gu, Xin Guo, Xiaoli Wei, and Renyuan Xu · 2021
Later among the works it cites.
Nash equilibria for exchangeable team against team games and their mean field limit
Sina Sanjari, Naci Saldi, and Serdar Yüksel · 2022
Later among the works it cites.
On improving model-free algorithms for decentralized multi-agent reinforcement learning
Weichao Mao, Lin Yang, Kaiqing Zhang, and Tamer Başar · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
René Carmona, Mathieu Laurière, and Zongjun Tan · 2019
Cited alongside, same era.
Eric Mazumdar, Lillian J Ratliff, Michael I Jordan, and S Shankar Sastry · 2019
Cited alongside, same era.
A tractable mean field game model for the analysis of crowd evacuation dynamics
Noureddine Toumi, Roland Malhamé, and Jerome Le Ny · 2020
Cited alongside, same era.
Policy optimization for linear-quadratic zero-sum mean-field type games
René Carmona, Kenza Hamidouche, Mathieu Laurière, and Zongjun Tan · 2020
Cited alongside, same era.
On the convergence of model free learning in mean field games
Romuald Elie, Julien Perolat, Mathieu Laurière, Matthieu Geist, and Olivier Pietquin · 2020
Cited alongside, same era.
Robust multi-agent reinforcement learning with model uncertainty
Kaiqing Zhang, Tao Sun, Yunzhe Tao, Sahika Genc, Sunil Mallya, and Tamer Başar · 2020
Cited alongside, same era.
Multi-agent reinforcement learning: A selective overview of theories and algorithms
Kaiqing Zhang, Zhuoran Yang, and Tamer Başar · 2021
Cited alongside, same era.
Unified reinforcement Q-learning for mean field game and control problems
Andrea Angiuli, Jean-Pierre Fouque, and Mathieu Laurière · 2022
Later among the works it cites.
Scalable deep reinforcement learning algorithms for mean field games
Mathieu Lauriere, Sarah Perrin, Sertan Girgin, Paul Muller, Ayush Jain, Theophile Cabannes, Georgios Piliouras, Julien Pérolat, Romuald Elie, Olivier Pietquin, et al · 2022
Later among the works it cites.
Learning mean field games: A survey
Mathieu Laurière, Sarah Perrin, Matthieu Geist, and Olivier Pietquin · 2022
Later among the works it cites.
Independent policy gradient for large-scale markov potential games: Sharper rates, function approximation, and game-agnostic convergence
Dongsheng Ding, Chen-Yu Wei, Kaiqing Zhang, and Mihailo Jovanovic · 2022
Later among the works it cites.
Revisiting LQR control from the perspective of receding-horizon policy gradient
Xiangyuan Zhang and Tamer Başar · 2023
Later among the works it cites.
Policy gradient methods find the Nash equilibrium in N-player general-sum linear-quadratic games
Ben Hambly, Renyuan Xu, and Huining Yang · 2023
Later among the works it cites.
Reinforcement learning for non-stationary discrete-time linear–quadratic mean-field games in multiple populations
Muhammad Aneeq Uz Zaman, Erik Miehling, and Tamer Başar · 2023
Later among the works it cites.
Oracle-free reinforcement learning in mean-field games along a single sample path
Muhammad Aneeq Uz Zaman, Alec Koppel, Sujay Bhatt, and Tamer Başar · 2023
Later among the works it cites.
Model-free mean-field reinforcement learning: mean-field MDP and mean-field Q-learning
René Carmona, Mathieu Laurière, and Zongjun Tan · 2023
Later among the works it cites.
Learning the Kalman filter with fine-grained sample complexity
Xiangyuan Zhang, Bin Hu, and Tamer Başar · 2023
Later among the works it cites.
Toward a theoretical foundation of policy optimization for learning control policies
Bin Hu, Kaiqing Zhang, Na Li, Mehran Mesbahi, Maryam Fazel, and Tamer Başar · 2023
Later among the works it cites.