Fetching the paper…
Reading the bibliography…
Multi-agent reinforcement learning (MARL) becomes more challenging in the presence of more agents, as the capacity of the joint state and action spaces grows exponentially in the number of agents.
Arora, S · 1901
Earlier work this paper cites.
Probabilistic symmetry and invariant neural networks
Bloem-Reddy, B · 1901
Earlier work this paper cites.
Neural proximal/trust region policy optimization attains globally optimal policy
Liu, B · 1906
Earlier work this paper cites.
Model-free mean-field reinforcement learning: mean-field mdp and mean-field q-learning
Carmona, R · 1910
Earlier work this paper cites.
Pic: Permutation invariant critic for multi-agent deep reinforcement learning
Liu, I.-J · 1911
Earlier work this paper cites.
Multi-agent reinforcement learning: A selective overview of theories and algorithms
Zhang, K · 1911
Earlier work this paper cites.
Convergence analysis of a proximal-like minimization algorithm using bregman functions
Chen, G · 1993
Earlier work this paper cites.
Markov games as a framework for multi-agent reinforcement learning
Littman, M. L · 1994
Earlier work this paper cites.
Reinforcement learning in the multi-robot domain
Matarić, M. J · 1997
Earlier work this paper cites.
Approximately optimal approximate reinforcement learning
Kakade, S · 2002
Earlier work this paper cites.
The ai economist: Improving equality and productivity with ai-driven tax policies
Zheng, S · 2004
Earlier work this paper cites.
Cooperative multi-agent learning: The state of the art
Panait, L · 2005
Earlier work this paper cites.
Half field offense in robocup soccer: A multiagent reinforcement learning case study
Kalyanakrishnan, S · 2006
Cited alongside, same era.
A hilbert space embedding for distributions
Smola, A · 2007
Cited alongside, same era.
Hilbert space embeddings of conditional distributions with applications to dynamical systems
Song, L · 2009
Cited alongside, same era.
Approximate policy iteration: A survey and some new methods
Bertsekas, D. P · 2011
Cited alongside, same era.
An overview of recent progress in the study of distributed multi-agent coordination
Cao, Y · 2013
Cited alongside, same era.
Reinforcement learning in robotics: A survey
Kober, J · 2013
Cited alongside, same era.
Value-decomposition networks for cooperative multi-agent learning
Sunehag, P · 2017
Later among the works it cites.
Learning deep mean field games for modeling large population behavior
Yang, J · 2017
Later among the works it cites.
Deep sets
Zaheer, M · 2017
Later among the works it cites.
A note on lazy training in supervised differentiable programming
Chizat, L · 2018
Later among the works it cites.
Neural tangent kernel: Convergence and generalization in neural networks
Jacot, A · 2018
Later among the works it cites.
Deep reinforcement learning for event-driven multi-agent decision processes
Menda, K · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Muandet, K · 2016
Cited alongside, same era.
Safe, multi-agent, reinforcement learning for autonomous driving
Shalev-Shwartz, S · 2016
Cited alongside, same era.
Sbeed: Convergent reinforcement learning with nonlinear function approximation
Dai, B · 2017
Cited alongside, same era.
Semi-supervised classification with graph convolutional networks
Kipf, T. N · 2017
Cited alongside, same era.
Multi-agent actor-critic for mixed cooperative-competitive environments
Lowe, R · 2017
Cited alongside, same era.
Proximal policy optimization algorithms
Schulman, J · 2017
Cited alongside, same era.
Later among the works it cites.
Emergence of grounded compositional language in multi-agent populations
Mordatch, I · 2018
Later among the works it cites.
Reinforcement learning: An introduction
Sutton, R. S · 2018
Later among the works it cites.
Mean field multi-agent reinforcement learning
Yang, Y · 2018
Later among the works it cites.
Neural temporal-difference learning converges to global optima
Cai, Q · 2019
Later among the works it cites.
Grandmaster level in starcraft ii using multi-agent reinforcement learning
Vinyals, O · 2019
Later among the works it cites.
Breaking the curse of many agents: Provable mean embedding q-iteration for mean-field reinforcement learning
Wang, L · 2020
Later among the works it cites.