Fetching the paper…
Reading the bibliography…
Social dilemmas, where mutual cooperation can lead to high payoffs but participants face incentives to cheat, are ubiquitous in multi-agent interaction.
Noncooperative collusion under imperfect price information
Edward J Green and Robert H Porter · 1984
Earlier work this paper cites.
The folk theorem in repeated games with discounting or with incomplete information
Drew Fudenberg and Eric Maskin · 1986
Earlier work this paper cites.
Toward a theory of discounted repeated games with imperfect monitoring
Dilip Abreu, David Pearce, and Ennio Stacchetti · 1990
Earlier work this paper cites.
Bargaining and market behavior in jerusalem, ljubljana, pittsburgh, and tokyo: An experimental study
Alvin E Roth, Vesna Prasnikar, Masahiro Okuno-Fujiwara, and Shmuel Zamir · 1991
Earlier work this paper cites.
A reinforcement learning method for maximizing undiscounted rewards
Anton Schwartz · 1993
Earlier work this paper cites.
The folk theorem with imperfect public information
Drew Fudenberg, David Levine, and Eric Maskin · 1994
Earlier work this paper cites.
Reinforcement learning algorithm for partially observable markov decision problems
Tommi Jaakkola, Satinder P Singh, and Michael I Jordan · 1995
Earlier work this paper cites.
Temporal difference learning and td-gammon
Gerald Tesauro · 1995
Earlier work this paper cites.
Multiagent reinforcement learning in the iterated prisoner’s dilemma
Tuomas W Sandholm and Robert H Crites · 1996
Earlier work this paper cites.
The theory of learning in games , volume 2
Drew Fudenberg and David K Levine · 1998
Earlier work this paper cites.
A theory of fairness, competition, and cooperation
Ernst Fehr and Klaus M Schmidt · 1999
Earlier work this paper cites.
Belief-free equilibria in repeated games
Jeffrey C Ely, Johannes Hörner, and Wojciech Olszewski · 2005
Earlier work this paper cites.
A polynomial-time nash equilibrium algorithm for repeated games
Michael L Littman and Peter Stone · 2005
Earlier work this paper cites.
The evolution of cooperation: revised edition
Robert M Axelrod · 2006
Earlier work this paper cites.
Evolutionary dynamics
Martin A Nowak · 2006
Earlier work this paper cites.
Awesome: A general multiagent learning algorithm that converges in self-play and learns a best response against stationary opponents
Vincent Conitzer and Tuomas Sandholm · 2007
Earlier work this paper cites.
If multi-agent learning is the answer, what is the question?
Yoav Shoham, Rob Powers, and Trond Grenager · 2007
Earlier work this paper cites.
Social reward shaping in the prisoner’s dilemma
Monica Babes, Enrique Munoz De Cote, and Michael L Littman · 2008
Earlier work this paper cites.
A polynomial-time nash equilibrium algorithm for repeated stochastic games
Enrique Munoz de Cote and Michael L Littman · 2008
Earlier work this paper cites.
Accidental outcomes guide punishment in a ‘trembling hand’ game
Fiery Cushman, Anna Dreber, Ying Wang, and Jay Costa · 2009
Cited alongside, same era.
Reinforcement learning for robot soccer
Martin Riedmiller, Thomas Gabel, Roland Hafner, and Sascha Lange · 2009
Cited alongside, same era.
Information can wreck cooperation: A counterpoint to kandori (1992)
Yuichiro Kamada and Scott Duke Kominers · 2010
Cited alongside, same era.
Markov chains and stochastic stability
Sean P Meyn and Richard L Tweedie · 2012
Cited alongside, same era.
A survey of real-time strategy game ai research and competition in starcraft
Santiago Ontanón, Gabriel Synnaeve, Alberto Uriarte, Florian Richoux, David Churchill, and Mike Preuss · 2013
Cited alongside, same era.
Moral tribes: Emotion, reason, and the gap between us and them
Joshua Greene · 2014
Cited alongside, same era.
Asynchronous methods for deep reinforcement learning
Volodymyr Mnih, Adria Puigdomenech Badia, Mehdi Mirza, Alex Graves, Timothy Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu · 2016
Later among the works it cites.
Nudging cooperation in a crowd experiment
Tamara Niella, Nicolás Stier-Moses, and Mariano Sigman · 2016
Later among the works it cites.
Mastering the game of go with deep neural networks and tree search
David Silver, Aja Huang, Chris J Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al · 2016
Later among the works it cites.
Nicolas Usunier, Gabriel Synnaeve, Zeming Lin, and Soumith Chintala · 2016
Later among the works it cites.
Training agent for first-person shooter game with actor-critic curriculum learning
Yuxin Wu and Yuandong Tian · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cooperating with the future
Oliver P Hauser, David G Rand, Alexander Peysakhovich, and Martin A Nowak · 2014
Cited alongside, same era.
Markov decision processes: discrete stochastic dynamic programming
Martin L Puterman · 2014
Cited alongside, same era.
Social heuristics shape intuitive cooperation
David G Rand, Alexander Peysakhovich, Gordon T Kraft-Todd, George E Newman, Owen Wurzbacher, Martin A Nowak, and Joshua D Greene · 2014
Cited alongside, same era.
Memory-based control with recurrent neural networks
Nicolas Heess, Jonathan J. Hunt, Timothy P. Lillicrap, and David Silver · 2015
Cited alongside, same era.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Cited alongside, same era.
When punishment doesn’t pay: ’cold glow’ and decisions to punish
Aurélie Ouss and Alexander Peysakhovich · 2015
Cited alongside, same era.
Later among the works it cites.
Jacob W Crandall, Mayada Oudah, Fatimah Ishowo-Oloko, Sherief Abdallah, Jean-François Bonnefon, Manuel Cebrian, Azim Shariff, Michael A Goodrich, Iyad Rahwan, et al · 2017
Closest in time.
Learning cooperative visual dialog agents with deep reinforcement learning
Abhishek Das, Satwik Kottur, José MF Moura, Stefan Lee, and Dhruv Batra · 2017
Closest in time.
Emergent language in a multi-modal, multi-step referential game
Katrina Evtimova, Andrew Drozdov, Douwe Kiela, and Kyunghyun Cho · 2017
Closest in time.
Emergence of language with multi-agent games: Learning to communicate with sequences of symbols
Serhii Havrylov and Ivan Titov · 2017
Closest in time.
Multi-agent cooperation and the emergence of (natural) language
Angeliki Lazaridou, Alexander Peysakhovich, and Marco Baroni · 2017
Closest in time.
Multi-agent reinforcement learning in sequential social dilemmas
Joel Z Leibo, Vinicius Zambaldi, Marc Lanctot, Janusz Marecki, and Thore Graepel · 2017
Closest in time.
Maintaining cooperation in complex social dilemmas using deep reinforcement learning
Adam Lerer and Alexander Peysakhovich · 2017
Closest in time.
Multi-agent actor-critic for mixed cooperative-competitive environments
Ryan Lowe, Yi Wu, Aviv Tamar, Jean Harb, Pieter Abbeel, and Igor Mordatch · 2017
Closest in time.
Resilient cooperators stabilize long-run cooperation in the finitely repeated prisoner’s dilemma
Andrew Mao, Lili Dworkin, Siddharth Suri, and Duncan J Watts · 2017
Closest in time.
A multi-agent reinforcement learning model of common-pool resource appropriation
Julien Perolat, Joel Z Leibo, Vinicius Zambaldi, Charles Beattie, Karl Tuyls, and Thore Graepel · 2017
Closest in time.
Prosocial learning agents solve generalized stag hunts better than selfish ones
Alexander Peysakhovich and Adam Lerer · 2017
Closest in time.
Locally noisy autonomous agents improve global human coordination in network experiments
Hirokazu Shirado and Nicholas A Christakis · 2017
Closest in time.
Multiagent cooperation and competition with deep reinforcement learning
Ardi Tampuu, Tambet Matiisen, Dorian Kodelja, Ilya Kuzovkin, Kristjan Korjus, Juhan Aru, Jaan Aru, and Raul Vicente · 2017
Closest in time.