Fetching the paper…
Reading the bibliography…
Matrix games like Prisoner's Dilemma have guided research on social dilemmas for decades.
The evolution of reciprocal altruism
Robert L. Trivers · 1971
Earlier work this paper cites.
Prisoner’s dilemma–recollections and observations
Anatol Rapoport · 1974
Earlier work this paper cites.
The Evolution of Cooperation
Robert Axelrod · 1984
Earlier work this paper cites.
An evolutionary approach to norms
Robert Axelrod · 1986
Earlier work this paper cites.
Tit for tat in heterogeneous populations
Martin A Nowak and Karl Sigmund · 1992
Earlier work this paper cites.
Evolutionary games and spatial chaos
Martin A Nowak and Robert M May · 1992
Earlier work this paper cites.
A strategy of win-stay, lose-shift that outperforms tit-for-tat in the prisoner’s dilemma game
Martin Nowak, Karl Sigmund, et al · 1993
Earlier work this paper cites.
Markov games as a framework for multi-agent reinforcement learning
M. L. Littman · 1994
Earlier work this paper cites.
High and low trusters’ responses to fear in a payoff matrix
Craig D Parks and Lorne G Hulbert · 1995
Earlier work this paper cites.
Multiagent reinforcement learning in the iterated prisoner’s dilemma
T.W. Sandholm and R.H. Crites · 1996
Earlier work this paper cites.
A neural substrate of prediction and reward
W. Schultz, P. Dayan, and P.R. Montague · 1997
Earlier work this paper cites.
Evolution of indirect reciprocity by image scoring
Martin A Nowak and Karl Sigmund · 1998
Earlier work this paper cites.
Multiagent reinforcement learning: Theoretical framework and an algorithm
J. Hu and M. P. Wellman · 1998
Earlier work this paper cites.
Introduction to Reinforcement Learning
Richard S. Sutton and Andrew G. Barto · 1998
Earlier work this paper cites.
Friend-or-foe Q-learning in general-sum games
Michael Littman · 2001
Earlier work this paper cites.
Learning dynamics in social dilemmas
Michael W Macy and Andreas Flache · 2002
Earlier work this paper cites.
Analyzing complex strategic interactions in multi-agent systems
William E Walsh, Rajarshi Das, Gerald Tesauro, and Jeffrey O Kephart · 2002
Earlier work this paper cites.
Value function approximation in zero-sum Markov games
M. G. Lagoudakis and R. Parr · 2002
Cited alongside, same era.
Correlated-Q learning
A. Greenwald and K. Hall · 2003
Cited alongside, same era.
Solving transition independent decentralized Markov decision processes
Raphen Becker, Shlomo Zilberstein, Victor Lesser, and Claudia V Goldman · 2004
Cited alongside, same era.
A framework for sequential planning in multi-agent settings
Piotr J Gmytrasiewicz and Prashant Doshi · 2005
Cited alongside, same era.
Uncertainty-based competition between prefrontal and dorsolateral striatal systems for behavioral control
Nathaniel D Daw, Yael Niv, and Peter Dayan · 2005
Cited alongside, same era.
Learning to cooperate in multi-agent social dilemmas
Enrique Munoz de Cote, Alessandro Lazaric, and Marcello Restelli · 2006
Cited alongside, same era.
The world of independent learners is not Markovian
Guillaume J. Laurent, Laëtitia Matignon, and N. Le Fort-Piat · 2011
Later among the works it cites.
Game theory and multiagent reinforcement learning
Ann Nowé, Peter Vrancx, and Yann-Michaël De Hauwere · 2012
Later among the works it cites.
Batch reinforcement learning
Sascha Lange, Thomas Gabel, and Martin Riedmiller · 2012
Later among the works it cites.
The psychology of social dilemmas: A review
Paul AM Van Lange, Jeff Joireman, Craig D Parks, and Eric Van Dijk · 2013
Later among the works it cites.
Empirically evaluating multiagent learning algorithms
Erik Zawadzki, Asher Lipson, and Kevin Leyton-Brown · 2014
Later among the works it cites.
Evolutionary dynamics of multi-agent learning: A survey
Daan Bloembergen, Karl Tuyls, Daniel Hennes, and Michael Kaisers · 2015
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A simple rule for the evolution of cooperation on graphs and social networks
Hisashi Ohtsuki, Christoph Hauert, Erez Lieberman, and Martin A Nowak · 2006
Cited alongside, same era.
A new route to the evolution of cooperation
Francisco C Santos and Jorge M Pacheco · 2006
Cited alongside, same era.
Methods for empirical game-theoretic analysis (extended abstract)
Michael Wellman · 2006
Cited alongside, same era.
Cyclic equilibria in Markov games
M. Zinkevich, A. Greenwald, and M. Littman · 2006
Cited alongside, same era.
Time, uncertainty, and individual differences in decisions to cooperate in resource dilemmas
Katherine V Kortenkamp and Colleen F Moore · 2006
Cited alongside, same era.
Micromotives and macrobehavior
Thomas C. Schelling · 2006
Cited alongside, same era.
Later among the works it cites.
Emotional multiagent reinforcement learning in spatial social dilemmas
Chao Yu, Minjie Zhang, Fenghui Ren, and Guozhen Tan · 2015
Later among the works it cites.
Human-level control through deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, S. Petersen, C. Beattie, A. Sadik, I. Antonoglou, H. King, D. Kumaran, D. Wierstra, S. Legg, and D. Hassabis · 2015
Later among the works it cites.
Approximate dynamic programming for two-player zero-sum Markov games
J. Pérolat, B. Scherrer, B. Piot, and O. Pietquin · 2015
Later among the works it cites.
Reinforcement learning improves behaviour from evaluative feedback
Michael L Littman · 2015
Later among the works it cites.
Cooperation emergence under resource-constrained peer punishment
Samhar Mahmoud, Simon Miles, and Michael Luck · 2016
Later among the works it cites.
Coordinate to cooperate or compete: abstract goals and joint intentions in social interaction
Max Kleiman-Weiner, M K Ho, J L Austerweil, Michael L Littman, and Josh B Tenenbaum · 2016
Later among the works it cites.
Softened approximate policy iteration for Markov games
J. Pérolat, B. Piot, M. Geist, B. Scherrer, and O. Pietquin · 2016
Later among the works it cites.
Algorithms for computing strategies in two-player simultaneous move games
Branislav Bošanský, Viliam Lisý, Marc Lanctot, Jiří Čermák, and Mark H.M. Winands · 2016
Later among the works it cites.
On the use of non-stationary strategies for solving two-player zero-sum Markov games
J. Pérolat, B. Piot, B. Scherrer, and O. Pietquin · 2016
Later among the works it cites.
Mastering the game of Go with deep neural networks and tree search
David Silver, Aja Huang, Chris J. Maddison, Arthur Guez, Laurent Sifre, George van den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, Sander Dieleman, Dominik Grewe, John Nham, Nal Kalchbrenner, Ilya Sutskever, Timothy Lillicrap, Madeleine Leach, Koray Kavukcuoglu, Thore Graepel, and Demis Hassabis · 2016
Later among the works it cites.
How other-regarding preferences can promote cooperation in non-zero-sum grid games
Joseph L. Austerweil, Stephen Brawner, Amy Greenwald, Elizabeth Hilliard, Mark Ho, Michael L. Littman, James MacGlashan, and Carl Trimbach · 2016
Later among the works it cites.