Fetching the paper…
Reading the bibliography…
Social dilemmas are situations where individuals face a temptation to increase their payoffs at a cost to total welfare.
The strategy of conflict
Schelling, Thomas C · 1980
Earlier work this paper cites.
The evolution of cooperation
Axelrod, Robert and Hamilton, William Donald · 1981
Earlier work this paper cites.
The evolution of cooperation
Axelrod, Robert M · 1984
Earlier work this paper cites.
The folk theorem in repeated games with discounting or with incomplete information
Fudenberg, Drew and Maskin, Eric · 1986
Earlier work this paper cites.
Impure altruism and donations to public goods: A theory of warm-glow giving
Andreoni, James · 1990
Earlier work this paper cites.
Bargaining and market behavior in jerusalem, ljubljana, pittsburgh, and tokyo: An experimental study
Roth, Alvin E, Prasnikar, Vesna, Okuno-Fujiwara, Masahiro, and Zamir, Shmuel · 1991
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Williams, Ronald J · 1992
Earlier work this paper cites.
A strategy of win-stay, lose-shift that outperforms tit-for-tat in the prisoner’s dilemma game
Nowak, Martin and Sigmund, Karl · 1993
Earlier work this paper cites.
A folk theorem for stochastic games
Dutta, Prajit K · 1995
Earlier work this paper cites.
Temporal difference learning and td-gammon
Tesauro, Gerald · 1995
Earlier work this paper cites.
Multiagent reinforcement learning in the iterated prisoner’s dilemma
Sandholm, Tuomas W and Crites, Robert H · 1996
Earlier work this paper cites.
The theory of learning in games , volume 2
Fudenberg, Drew and Levine, David K · 1998
Earlier work this paper cites.
A theory of fairness, competition, and cooperation
Fehr, Ernst and Schmidt, Klaus M · 1999
Earlier work this paper cites.
Algorithms for inverse reinforcement learning
Ng, Andrew Y, Russell, Stuart J, et al · 2000
Earlier work this paper cites.
Friend-or-foe q-learning in general-sum games
Littman, Michael L · 2001
Earlier work this paper cites.
Apprenticeship learning via inverse reinforcement learning
Abbeel, Pieter and Ng, Andrew Y · 2004
Earlier work this paper cites.
A polynomial-time nash equilibrium algorithm for repeated games
Littman, Michael L and Stone, Peter · 2005
Earlier work this paper cites.
Evolutionary dynamics
Nowak, Martin A · 2006
Earlier work this paper cites.
Awesome: A general multiagent learning algorithm that converges in self-play and learns a best response against stationary opponents
Conitzer, Vincent and Sandholm, Tuomas · 2007
Earlier work this paper cites.
Tit-for-tat or win-stay, lose-shift?
Imhof, Lorens A, Fudenberg, Drew, and Nowak, Martin A · 2007
Cited alongside, same era.
The complexity of finding nash equilibria
Papadimitriou, Christos H · 2007
Cited alongside, same era.
If multi-agent learning is the answer, what is the question?
Shoham, Yoav, Powers, Rob, and Grenager, Trond · 2007
Cited alongside, same era.
Social reward shaping in the prisoner’s dilemma
Babes, Monica, De Cote, Enrique Munoz, and Littman, Michael L · 2008
Cited alongside, same era.
A polynomial-time nash equilibrium algorithm for repeated stochastic games
de Cote, Enrique Munoz and Littman, Michael L · 2008
Cited alongside, same era.
Action understanding as inverse planning
Baker, Chris L, Saxe, Rebecca, and Tenenbaum, Joshua B · 2009
Cited alongside, same era.
Vizdoom: A doom-based ai research platform for visual reinforcement learning
Kempka, Michał, Wydmuch, Marek, Runc, Grzegorz, Toczek, Jakub, and Jaśkowski, Wojciech · 2016
Later among the works it cites.
Coordinate to cooperate or compete: abstract goals and joint intentions in social interaction
Kleiman-Weiner, Max, Ho, Mark K, Austerweil, Joe L, Michael L, Littman, and Tenenbaum, Joshua B · 2016
Later among the works it cites.
Asynchronous methods for deep reinforcement learning
Mnih, Volodymyr, Badia, Adria Puigdomenech, Mirza, Mehdi, Graves, Alex, Lillicrap, Timothy, Harley, Tim, Silver, David, and Kavukcuoglu, Koray · 2016
Later among the works it cites.
Mastering the game of go with deep neural networks and tree search
Silver, David, Huang, Aja, Maddison, Chris J, Guez, Arthur, Sifre, Laurent, Van Den Driessche, George, Schrittwieser, Julian, Antonoglou, Ioannis, Panneershelvam, Veda, Lanctot, Marc, et al · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Reinforcement learning for robot soccer
Riedmiller, Martin, Gabel, Thomas, Hafner, Roland, and Lange, Sascha · 2009
Cited alongside, same era.
The evolution of cooperation in infinitely repeated games: Experimental evidence
Bó, Pedro Dal and Fréchette, Guillaume R · 2011
Cited alongside, same era.
Slow to anger and fast to forgive: Cooperation in an uncertain world
Fudenberg, Drew, Rand, David G, and Dreber, Anna · 2012
Cited alongside, same era.
Powering up with indirect reciprocity in a large-scale field experiment
Yoeli, Erez, Hoffman, Moshe, Rand, David G, and Nowak, Martin A · 2013
Cited alongside, same era.
Cooperating with the future
Hauser, Oliver P, Rand, David G, Peysakhovich, Alexander, and Nowak, Martin A · 2014
Cited alongside, same era.
What makes a price fair? an experimental study of market experience and endogenous fairness norms
Herz, Holger and Taubinsky, Dmitry · 2014
Cited alongside, same era.
Usunier, Nicolas, Synnaeve, Gabriel, Lin, Zeming, and Chintala, Soumith · 2016
Later among the works it cites.
Training agent for first-person shooter game with actor-critic curriculum learning
Wu, Yuxin and Tian, Yuandong · 2016
Later among the works it cites.
Crandall, Jacob W, Oudah, Mayada, Ishowo-Oloko, Fatimah, Abdallah, Sherief, Bonnefon, Jean-François, Cebrian, Manuel, Shariff, Azim, Goodrich, Michael A, Rahwan, Iyad, et al · 2017
Closest in time.
Learning cooperative visual dialog agents with deep reinforcement learning
Das, Abhishek, Kottur, Satwik, Moura, José MF, Lee, Stefan, and Batra, Dhruv · 2017
Closest in time.
Emergent language in a multi-modal, multi-step referential game
Evtimova, Katrina, Drozdov, Andrew, Kiela, Douwe, and Cho, Kyunghyun · 2017
Closest in time.
Emergence of language with multi-agent games: Learning to communicate with sequences of symbols
Havrylov, Serhii and Titov, Ivan · 2017
Closest in time.
Multi-agent cooperation and the emergence of (natural) language
Lazaridou, Angeliki, Peysakhovich, Alexander, and Baroni, Marco · 2017
Closest in time.
Multi-agent reinforcement learning in sequential social dilemmas
Leibo, Joel Z, Zambaldi, Vinicius, Lanctot, Marc, Marecki, Janusz, and Graepel, Thore · 2017
Closest in time.
Deal or no deal? end-to-end learning for negotiation dialogues
Lewis, Mike, Yarats, Denis, Dauphin, Yann N, Parikh, Devi, and Batra, Dhruv · 2017
Closest in time.
Multi-agent actor-critic for mixed cooperative-competitive environments
Lowe, Ryan, Wu, Yi, Tamar, Aviv, Harb, Jean, Abbeel, Pieter, and Mordatch, Igor · 2017
Closest in time.
A multi-agent reinforcement learning model of common-pool resource appropriation
Perolat, Julien, Leibo, Joel Z, Zambaldi, Vinicius, Beattie, Charles, Tuyls, Karl, and Graepel, Thore · 2017
Closest in time.
Prosocial learning agents solve generalized stag hunts better than selfish ones
Peysakhovich, Alexander and Lerer, Adam · 2017
Closest in time.
Locally noisy autonomous agents improve global human coordination in network experiments
Shirado, Hirokazu and Christakis, Nicholas A · 2017
Closest in time.
Mastering the game of go without human knowledge
Silver, David, Schrittwieser, Julian, Simonyan, Karen, Antonoglou, Ioannis, Huang, Aja, Guez, Arthur, Hubert, Thomas, Baker, Lucas, Lai, Matthew, Bolton, Adrian, et al · 2017
Closest in time.
Multiagent cooperation and competition with deep reinforcement learning
Tampuu, Ardi, Matiisen, Tambet, Kodelja, Dorian, Kuzovkin, Ilya, Korjus, Kristjan, Aru, Juhan, Aru, Jaan, and Vicente, Raul · 2017
Closest in time.