Fetching the paper…
Reading the bibliography…
Multi-agent settings are quickly gathering importance in machine learning.
Iterative solution of games by fictitious play
George W Brown. 1951 · 1951
Earlier work this paper cites.
Games and Decisions: Introduction and Critical Survey
R Duncan Luce and Howard Raiffa. 1957 · 1957
Earlier work this paper cites.
The Application of Decision Theory and Dynamic Programming to Adaptive Control Systems
King Lee and K Louis. 1967 · 1967
Earlier work this paper cites.
Game theory, 1991
Drew Fudenberg and Jean Tirole. 1991 · 1991
Earlier work this paper cites.
Game theory: analysis of conflict
B Myerson Roger. 1991 · 1991
Earlier work this paper cites.
Game theory for applied economists
Robert Gibbons. 1992 · 1992
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald J Williams. 1992 · 1992
Earlier work this paper cites.
Multiagent reinforcement learning in the iterated prisoner’s dilemma
Tuomas W Sandholm and Robert H Crites. 1996 · 1996
Earlier work this paper cites.
Adversarial reinforcement learning
William Uther and Manuela Veloso. 1997 · 1997
Earlier work this paper cites.
The dynamics of reinforcement learning in cooperative multiagent systems
Caroline Claus and Craig Boutilier. 1998 · 1998
Earlier work this paper cites.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto. 1998 · 1998
Earlier work this paper cites.
Policy gradient methods for reinforcement learning with function approximation.. In NIPS
Richard S Sutton, David A McAllester, Satinder P Singh, Yishay Mansour, et al · 1999
Earlier work this paper cites.
Friend-or-foe Q-learning in general-sum games. In ICML
Michael L Littman. 2001 · 2001
Earlier work this paper cites.
Multiagent learning using a variable learning rate
Michael Bowling and Manuela Veloso. 2002 · 2002
Earlier work this paper cites.
Efficient Learning Equilibrium. In Advances in Neural Information Processing Systems
Ronen I. Brafman and Moshe Tennenholtz. 2003 · 2003
Earlier work this paper cites.
The evolution of cooperation: revised edition
Robert M Axelrod. 2006 · 2006
Earlier work this paper cites.
Cyclic equilibria in Markov games. In Advances in Neural Information Processing Systems
Martin Zinkevich, Amy Greenwald, and Michael L Littman. 2006 · 2006
Cited alongside, same era.
AWESOME: A general multiagent learning algorithm that converges in self-play and learns a best response against stationary opponents
Vincent Conitzer and Tuomas Sandholm. 2007 · 2007
Cited alongside, same era.
A comprehensive survey of multiagent reinforcement learning
Lucian Busoniu, Robert Babuska, and Bart De Schutter. 2008 · 2008
Cited alongside, same era.
A Polynomial-time Nash Equilibrium Algorithm for Repeated Stochastic Games. In 24th Conference on Uncertainty in Artificial Intelligence (UAI’08)
Enrique Munoz de Cote and Michael L. Littman. 2008 · 2008
Cited alongside, same era.
Classes of multiagent q-learning dynamics with epsilon-greedy exploration. In Proceedings of the 27th International Conference on Machine Learning (ICML-10)
Michael Wunder, Michael L Littman, and Monica Babes. 2010 · 2010
Multi-agent cooperation and the emergence of (natural) language
Angeliki Lazaridou, Alexander Peysakhovich, and Marco Baroni. 2016 · 2016
Later among the works it cites.
Unrolled generative adversarial networks
Luke Metz, Ben Poole, David Pfau, and Jascha Sohl-Dickstein. 2016 · 2016
Later among the works it cites.
Learning multiagent communication with backpropagation. In Advances in Neural Information Processing Systems
Sainbayar Sukhbaatar, Rob Fergus, et al · 2016
Later among the works it cites.
Learning Cooperative Visual Dialog Agents with Deep Reinforcement Learning
Abhishek Das, Satwik Kottur, José MF Moura, Stefan Lee, and Dhruv Batra. 2017 · 2017
Closest in time.
Stabilising experience replay for deep multi-agent reinforcement learning. In 34th International Conference of Machine Learning
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Multi-Agent Learning with Policy Prediction.. In AAAI
Chongjie Zhang and Victor R Lesser. 2010 · 2010
Cited alongside, same era.
Learning to compete, coordinate, and cooperate in repeated games using reinforcement learning
Jacob W Crandall and Michael A Goodrich. 2011 · 2011
Cited alongside, same era.
No-regret reductions for imitation learning and structured prediction. In In AISTATS
Stéphane Ross, Geoffrey J Gordon, and J Andrew Bagnell. 2011 · 2011
Cited alongside, same era.
Iterated Prisoner’s Dilemma contains strategies that dominate any evolutionary opponent
William H Press and Freeman J Dyson. 2012 · 2012
Cited alongside, same era.
Opponent Modelling by Sequence Prediction and Lookahead in Two-Player Games.. In ICAISC (2)
Richard Mealing and Jonathan L Shapiro. 2013 · 2013
Cited alongside, same era.
Multiagent learning in the presence of memory-bounded agents
Doran Chakraborty and Peter Stone. 2014 · 2014
Cited alongside, same era.
Opponent Modelling by Expectation-Maximisation and Sequence Prediction in Simplified Poker
Richard Mealing and Jonathan Shapiro. 2015 · 2015
Cited alongside, same era.
Jakob Foerster, Nantas Nardelli, Gregory Farquhar, Philip Torr, Pushmeet Kohli, Shimon Whiteson, et al · 2017
Closest in time.
Learning against sequential opponents in repeated stochastic games
Pablo Hernandez-Leal and Michael Kaisers. 2017 · 2017
Closest in time.
A Survey of Learning in Multiagent Environments: Dealing with Non-Stationarity
Pablo Hernandez-Leal, Michael Kaisers, Tim Baarslag, and Enrique Munoz de Cote. 2017 · 2017
Closest in time.
A Unified Game-Theoretic Approach to Multiagent Reinforcement Learning. In Advances in Neural Information Processing Systems (NIPS)
Marc Lanctot, Vinicius Zambaldi, Audrunas Gruslys, Angeliki Lazaridou, Karl Tuyls, Julien Perolat, David Silver, and Thore Graepel. 2017 · 2017
Closest in time.
Multi-agent Reinforcement Learning in Sequential Social Dilemmas. In Proceedings of the 16th Conference on Autonomous Agents and MultiAgent Systems
Joel Z Leibo, Vinicius Zambaldi, Marc Lanctot, Janusz Marecki, and Thore Graepel. 2017 · 2017
Closest in time.
Maintaining cooperation in complex social dilemmas using deep reinforcement learning
Adam Lerer and Alexander Peysakhovich. 2017 · 2017
Closest in time.
Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments
Ryan Lowe, Yi Wu, Aviv Tamar, Jean Harb, Pieter Abbeel, and Igor Mordatch. 2017 · 2017
Closest in time.
Emergence of Grounded Compositional Language in Multi-Agent Populations
Igor Mordatch and Pieter Abbeel. 2017 · 2017
Closest in time.
Deep Decentralized Multi-task Multi-Agent RL under Partial Observability
Shayegan Omidshafiei, Jason Pazis, Christopher Amato, Jonathan P How, and John Vian. 2017 · 2017
Closest in time.
Counterfactual Multi-Agent Policy Gradients. In AAAI
Jakob Foerster, Gregory Farquhar, Triantafyllos Afouras, Nantas Nardelli, and Shimon Whiteson. 2018 · 2018
Closest in time.
Neil C Rabinowitz, Frank Perbet, H Francis Song, Chiyuan Zhang, SM Eslami, and Matthew Botvinick. 2018 · 2018
Closest in time.