Fetching the paper…
Reading the bibliography…
Deep reinforcement learning has become an important paradigm for constructing agents that can enter complex multi-agent situations and improve their policies through experience.
Zur theorie der gesellschaftsspiele
J v Neumann. 1928 · 1928
Earlier work this paper cites.
Asynchronous methods for deep reinforcement learning. In International Conference on Machine Learning
Volodymyr Mnih, Adria Puigdomenech Badia, Mehdi Mirza, Alex Graves, Timothy Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu. 2016 · 1937
Earlier work this paper cites.
The folk theorem in repeated games with discounting or with incomplete information
Drew Fudenberg and Eric Maskin. 1986 · 1986
Earlier work this paper cites.
A general theory of equilibrium selection in games
John C Harsanyi, Reinhard Selten, et al · 1988
Earlier work this paper cites.
Tacit coordination games, strategic uncertainty, and coordination failure
John B Van Huyck, Raymond C Battalio, and Richard O Beil. 1990 · 1990
Earlier work this paper cites.
Strategic uncertainty, equilibrium selection, and coordination failure in average opinion games
John B Van Huyck, Raymond C Battalio, and Richard O Beil. 1991 · 1991
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald J Williams. 1992 · 1992
Earlier work this paper cites.
Global games and equilibrium selection
Hans Carlsson and Eric Van Damme. 1993 · 1993
Earlier work this paper cites.
Learning, mutation, and long run equilibria in games
Michihiro Kandori, George J Mailath, and Rafael Rob. 1993 · 1993
Earlier work this paper cites.
Reinforcement learning: A survey
Leslie Pack Kaelbling, Michael L Littman, and Andrew W Moore. 1996 · 1996
Earlier work this paper cites.
Multiagent reinforcement learning in the iterated prisoner’s dilemma
Tuomas W Sandholm and Robert H Crites. 1996 · 1996
Earlier work this paper cites.
Predicting how people play games: Reinforcement learning in experimental games with unique, mixed strategy equilibria
Ido Erev and Alvin E Roth. 1998 · 1998
Earlier work this paper cites.
The theory of learning in games
Drew Fudenberg and David K Levine. 1998 · 1998
Earlier work this paper cites.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto. 1998 · 1998
Earlier work this paper cites.
Contagion
Stephen Morris. 2000 · 2000
Earlier work this paper cites.
Friend-or-foe Q-learning in general-sum games. In ICML
Michael L Littman. 2001 · 2001
Earlier work this paper cites.
Reinforcement learning of coordination in cooperative multi-agent systems
Spiros Kapetanakis and Daniel Kudenko. 2002 · 2002
Earlier work this paper cites.
Behavioral game theory: Experiments in strategic interaction
Colin Camerer. 2003 · 2003
Earlier work this paper cites.
Reinforcement learning to play an optimal Nash equilibrium in team Markov games. In Advances in neural information processing systems
Xiaofeng Wang and Tuomas Sandholm. 2003 · 2003
Earlier work this paper cites.
A polynomial-time Nash equilibrium algorithm for repeated games
Michael L Littman and Peter Stone. 2005 · 2005
Earlier work this paper cites.
Evolutionary dynamics
Martin A Nowak. 2006 · 2006
Earlier work this paper cites.
Social reward shaping in the prisoner’s dilemma. In Proceedings of the 7th international joint conference on Autonomous agents and multiagent systems-Volume 3
Monica Babes, Enrique Munoz De Cote, and Michael L Littman. 2008 · 2008
Cited alongside, same era.
Theoretical advantages of lenient learners: An evolutionary game theoretic perspective
Liviu Panait, Karl Tuyls, and Sean Luke. 2008 · 2008
Cited alongside, same era.
Game theory of mind
Wako Yoshida, Ray J Dolan, and Karl J Friston. 2008 · 2008
Cited alongside, same era.
Efficient influence maximization in social networks. In Proceedings of the 15th ACM SIGKDD international conference on Knowledge discovery and data mining
Wei Chen, Yajun Wang, and Siyu Yang. 2009 · 2009
Cited alongside, same era.
The complexity of computing a Nash equilibrium
Constantinos Daskalakis, Paul W Goldberg, and Christos H Papadimitriou. 2009 · 2009
Cited alongside, same era.
Mastering the game of Go with deep neural networks and tree search
David Silver, Aja Huang, Chris J Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al · 2016
Later among the works it cites.
Value iteration networks. In Advances in Neural Information Processing Systems
Aviv Tamar, Yi Wu, Garrett Thomas, Sergey Levine, and Pieter Abbeel. 2016 · 2016
Later among the works it cites.
Training agent for first-person shooter game with actor-critic curriculum learning
Yuxin Wu and Yuandong Tian. 2016 · 2016
Later among the works it cites.
Jacob W Crandall, Mayada Oudah, Fatimah Ishowo-Oloko, Sherief Abdallah, Jean-François Bonnefon, Manuel Cebrian, Azim Shariff, Michael A Goodrich, Iyad Rahwan, et al · 2017
Closest in time.
Learning Cooperative Visual Dialog Agents with Deep Reinforcement Learning
Abhishek Das, Satwik Kottur, José MF Moura, Stefan Lee, and Dhruv Batra. 2017 · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Social and economic networks
Matthew O Jackson. 2010 · 2010
Cited alongside, same era.
Theoretical considerations of potential-based reward shaping for multi-agent systems. In The 10th International Conference on Autonomous Agents and Multiagent Systems-Volume 1
Sam Devlin and Daniel Kudenko. 2011 · 2011
Cited alongside, same era.
An empirical study of potential-based reward shaping and advice in complex, multi-agent systems
Sam Devlin, Daniel Kudenko, and Marek Grześ. 2011 · 2011
Cited alongside, same era.
A polynomial-time Nash equilibrium algorithm for repeated stochastic games
Enrique Munoz De Cote and Michael L Littman. 2012 · 2012
Cited alongside, same era.
Independent reinforcement learners in cooperative Markov games: a survey regarding coordination problems
Laetitia Matignon, Guillaume J Laurent, and Nadine Le Fort-Piat. 2012 · 2012
Cited alongside, same era.
Cooperating with the future
Oliver P Hauser, David G Rand, Alexander Peysakhovich, and Martin A Nowak. 2014 · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik Kingma and Jimmy Ba. 2014 · 2014
Cited alongside, same era.
Closest in time.
Emergent Language in a Multi-Modal, Multi-Step Referential Game
Katrina Evtimova, Andrew Drozdov, Douwe Kiela, and Kyunghyun Cho. 2017 · 2017
Closest in time.
Counterfactual Multi-Agent Policy Gradients
Jakob Foerster, Gregory Farquhar, Triantafyllos Afouras, Nantas Nardelli, and Shimon Whiteson. 2017b · 2017
Closest in time.
Learning with Opponent-Learning Awareness
Jakob N Foerster, Richard Y Chen, Maruan Al-Shedivat, Shimon Whiteson, Pieter Abbeel, and Igor Mordatch. 2017a · 2017
Closest in time.
Emergence of Language with Multi-agent Games: Learning to Communicate with Sequences of Symbols
Serhii Havrylov and Ivan Titov. 2017 · 2017
Closest in time.
A Unified Game-Theoretic Approach to Multiagent Reinforcement Learning
Marc Lanctot, Vinicius Zambaldi, Audrunas Gruslys, Angeliki Lazaridou, Karl Tuyls, Julien Perolat, David Silver, and Thore Graepel. 2017 · 2017
Closest in time.
Multi-agent cooperation and the emergence of (natural) language. In International Conference on Learning Representations
Angeliki Lazaridou, Alexander Peysakhovich, and Marco Baroni. 2017 · 2017
Closest in time.
Multi-agent Reinforcement Learning in Sequential Social Dilemmas. In Proceedings of the 16th Conference on Autonomous Agents and MultiAgent Systems
Joel Z Leibo, Vinicius Zambaldi, Marc Lanctot, Janusz Marecki, and Thore Graepel. 2017 · 2017
Closest in time.
Maintaining cooperation in complex social dilemmas using deep reinforcement learning
Adam Lerer and Alexander Peysakhovich. 2017 · 2017
Closest in time.
Deal or No Deal? End-to-End Learning for Negotiation Dialogues
Mike Lewis, Denis Yarats, Yann N Dauphin, Devi Parikh, and Dhruv Batra. 2017 · 2017
Closest in time.
Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments
Ryan Lowe, Yi Wu, Aviv Tamar, Jean Harb, Pieter Abbeel, and Igor Mordatch. 2017 · 2017
Closest in time.
Resilient cooperators stabilize long-run cooperation in the finitely repeated Prisoner?s Dilemma
Andrew Mao, Lili Dworkin, Siddharth Suri, and Duncan J Watts. 2017 · 2017
Closest in time.
Multiagent Bidirectionally-Coordinated Nets for Learning to Play StarCraft Combat Games
Peng Peng, Quan Yuan, Ying Wen, Yaodong Yang, Zhenkun Tang, Haitao Long, and Jun Wang. 2017 · 2017
Closest in time.
Consequentialist conditional cooperation in social dilemmas with imperfect information
Alexander Peysakhovich and Adam Lerer. 2017 · 2017
Closest in time.
Locally noisy autonomous agents improve global human coordination in network experiments
Hirokazu Shirado and Nicholas A Christakis. 2017 · 2017
Closest in time.
Multiagent cooperation and competition with deep reinforcement learning
Ardi Tampuu, Tambet Matiisen, Dorian Kodelja, Ilya Kuzovkin, Kristjan Korjus, Juhan Aru, Jaan Aru, and Raul Vicente. 2017 · 2017
Closest in time.