Fetching the paper…
Reading the bibliography…
Traditional multi-agent reinforcement learning algorithms are not scalable to environments with more than a few agents, since these algorithms are exponential in the number of agents.
On the likelihood that one unknown probability exceeds another in view of the evidence of two samples
William R Thompson. 1933 · 1933
Earlier work this paper cites.
Introduction to phase transitions and critical phenomena
H Eugene Stanley and Guenter Ahlers. 1973 · 1973
Earlier work this paper cites.
Q-learning
Christopher JCH Watkins and Peter Dayan. 1992 · 1992
Earlier work this paper cites.
Multi-agent reinforcement learning: Independent vs. cooperative agents. In Proceedings of the tenth international conference on machine learning . 330–337
Ming Tan. 1993 · 1993
Earlier work this paper cites.
Markov games as a framework for multi-agent reinforcement learning
Michael L Littman. 1994 · 1994
Earlier work this paper cites.
Asynchronous stochastic approximation and Q-learning
John N Tsitsiklis. 1994 · 1994
Earlier work this paper cites.
Relativistic mean field theory in finite nuclei
P Ring. 1996 · 1996
Earlier work this paper cites.
Convergence results for single-step on-policy reinforcement-learning algorithms
Satinder Singh, Tommi Jaakkola, Michael L Littman, and Csaba Szepesvári. 2000 · 2000
Earlier work this paper cites.
Nash Q-learning for general-sum stochastic games
Junling Hu and Michael P Wellman. 2003 · 2003
Earlier work this paper cites.
Brown’s original fictitious play
Ulrich Berger. 2007 · 2007
Earlier work this paper cites.
Mean field games
Jean-Michel Lasry and Pierre-Louis Lions. 2007 · 2007
Earlier work this paper cites.
Mean field stochastic adaptive control
Arman C Kizilkale and Peter E Caines. 2012 · 2012
Cited alongside, same era.
Partially Observable Mean Field Reinforcement Learning
Sriram Ganapathi Subramanian, Matthew E. Taylor, Mark Crowley, and Pascal Poupart. 2020 · 2012
Cited alongside, same era.
Deep recurrent Q-learning for partially observable MDPs. In 2015 AAAI Fall Symposium Series
Matthew Hausknecht and Peter Stone. 2015 · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Cited alongside, same era.
Deep exploration via bootstrapped DQN. In Advances in neural information processing systems . 4026–4034
Ian Osband, Charles Blundell, Alexander Pritzel, and Benjamin Van Roy. 2016 · 2016
Cited alongside, same era.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto. 2018 · 2018
Later among the works it cites.
Magent: A many-agent reinforcement learning platform for artificial collective intelligence. In Thirty-Second AAAI Conference on Artificial Intelligence
Lianmin Zheng, Jiacheng Yang, Han Cai, Ming Zhou, Weinan Zhang, Jun Wang, and Yong Yu. 2018 · 2018
Later among the works it cites.
Approximate fictitious play for mean field games
Romuald Elie, Julien Pérolat, Mathieu Laurière, Matthieu Geist, and Olivier Pietquin. 2019 · 2019
Later among the works it cites.
Learning mean-field games. In Advances in Neural Information Processing Systems . 4967–4977
Xin Guo, Anran Hu, Renyuan Xu, and Junzi Zhang. 2019 · 2019
Later among the works it cites.
A survey and critique of multiagent deep reinforcement learning
Pablo Hernandez-Leal, Bilal Kartal, and Matthew E Taylor. 2019 · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
QMDP-net: Deep learning for planning under partial observability. In Advances in Neural Information Processing Systems . 4694–4704
Peter Karkus, David Hsu, and Wee Sun Lee. 2017 · 2017
Cited alongside, same era.
A unified game-theoretic approach to multiagent reinforcement learning. In Advances in Neural Information Processing Systems . 4190–4203
Marc Lanctot, Vinicius Zambaldi, Audrunas Gruslys, Angeliki Lazaridou, Karl Tuyls, Julien Pérolat, David Silver, and Thore Graepel. 2017 · 2017
Cited alongside, same era.
On improving deep reinforcement learning for POMDPs
Pengfei Zhu, Xin Li, Pascal Poupart, and Guanghui Miao. 2017 · 2017
Cited alongside, same era.
Learning with opponent-learning awareness. In Proceedings of the 17th International Conference on Autonomous Agents and MultiAgent Systems . International Foundation for Autonomous Agents and Multiagent Systems, 122–130
Jakob Foerster, Richard Y Chen, Maruan Al-Shedivat, Shimon Whiteson, Pieter Abbeel, and Igor Mordatch. 2018 · 2018
Cited alongside, same era.
Decentralised learning in systems with many, many strategic agents. In Thirty-Second AAAI Conference on Artificial Intelligence
David Mguni, Joel Jennings, and Enrique Munoz de Cote. 2018 · 2018
Cited alongside, same era.
Partially Observable Mean Field Reinforcement Learning
Sriram Ganapathi Subramanian. [n.d.]
Cited in the paper.
Learning Deep Mean Field Games for Modeling Large Population Behavior. In International Conference on Learning Representations (ICLR)
Jiachen Yang, Xiaojing Ye, Rakshit Trivedi, Huan Xu, and Hongyuan Zha. 2018b
Cited in the paper.
Coordinating the Crowd: Inducing Desirable Equilibria in Non-Cooperative Systems. In Proceedings of the 18th International Conference on Autonomous Agents and MultiAgent Systems . International Foundation for Autonomous Agents and Multiagent Systems, 386–394
David Mguni, Joel Jennings, Emilio Sison, Sergio Valcarcel Macua, Sofia Ceppi, and Enrique Munoz de Cote. 2019 · 2019
Later among the works it cites.
Reinforcement learning in stationary mean-field games. In Proceedings of the 18th International Conference on Autonomous Agents and MultiAgent Systems . International Foundation for Autonomous Agents and Multiagent Systems, 251–259
Jayakumar Subramanian and Aditya Mahajan. 2019 · 2019
Later among the works it cites.
On the Convergence of Model Free Learning in Mean Field Games. In AAAI Conference on Artificial Intelligence (AAAI 2020)
Romuald Elie, Julien Perolat, Mathieu Laurière, Matthieu Geist, and Olivier Pietquin. 2020 · 2020
Closest in time.
Multi Type Mean Field Reinforcement Learning. In Proceedings of the Autonomous Agents and Multi Agent Systems (AAMAS 2020) . IFAAMAS, Auckland, New Zealand
Sriram Ganapathi Subramanian, Pascal Poupart, Matthew E. Taylor, and Nidhi Hegde. 2020 · 2020
Closest in time.
A unified analysis of value-function-based reinforcement-learning algorithms
Csaba Szepesvári and Michael L Littman. 1999 · 2060
Closest in time.