Fetching the paper…
Reading the bibliography…
Multi-agent reinforcement learning offers a way to study how communication could emerge in communities of agents needing to solve specific problems.
Theory of Games and Economic Behavior
John Von Neumann and Oskar Morgenstern · 1944
Earlier work this paper cites.
Design for Bidding
S. J. Simon · 1949
Earlier work this paper cites.
Non-cooperative games
J.F. Nash · 1951
Earlier work this paper cites.
The Strategy of Conflict
T.C. Schelling · 1960
Earlier work this paper cites.
Convention: A philosophical study
David Lewis · 1969
Earlier work this paper cites.
Strategic information transmission
V.P. Crawford and J. Sobel · 1982
Earlier work this paper cites.
An experimental analysis of ultimatum bargaining
Werner Güth, Rolf Schmittberger, and Bernd Schwarze · 1982
Earlier work this paper cites.
Perfect equilibrium in a bargaining model
Ariel Rubinstein · 1982
Earlier work this paper cites.
The nash bargaining solution in economic modelling
Ken Binmore, Ariel Rubinstein, and Asher Wolinsky · 1986
Earlier work this paper cites.
Efficiency in evolutionary games: Darwin, nash, and the secret handshake
A. J. Robson · 1990
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald J. Williams · 1992
Earlier work this paper cites.
Cheap talk
J. Farrell and M. Rabin · 1996
Earlier work this paper cites.
Multiagent reinforcement learning in the iterated prisoner’s dilemma
Tuomas W. Sandholm and Robert H. Crites · 1996
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
The Theory of Learning in Games
D. Fudenberg and D. Levine · 1998
Earlier work this paper cites.
The evolution of language
Martin A. Nowak and David C. Krakauer · 1999
Cited alongside, same era.
Learning to trade via direct reinforcement
John Moody and Matthew Saffell · 2001
Cited alongside, same era.
Signals, evolution and the explanatory power of transient information
B. Skyrms · 2002
Cited alongside, same era.
Progress in the simulation of emergent communication and language
Kyle Wagner, James A. Reggia, Juan Uriagereka, and Gerald S. Wilkinson · 2003
Cited alongside, same era.
Cooperative multi-agent learning: The state of the art
Liviu Panait and Sean Luke · 2005
Cited alongside, same era.
Learning to communicate in a decentralized environment
Claudia V. Goldman, Martin Allen, and Shlomo Zilberstein · 2007
Cited alongside, same era.
Learning to communicate with deep multi-agent reinforcement learning
Jakob N. Foerster, Yannis M. Assael, Nando de Freitas, and Shimon Whiteson · 2016
Later among the works it cites.
Multi-agent cooperation and the emergence of (natural) language
Angeliki Lazaridou, Alexander Peysakhovich, and Marco Baroni · 2016
Later among the works it cites.
Learning multiagent communication with backpropagation
S. Sukhbaatar, A. Szlam, and R. Fergus · 2016
Later among the works it cites.
Emergent complexity via multi-agent competition
Trapit Bansal, Jakub Pachocki, Szymon Sidor, Ilya Sutskever, and Igor Mordatch · 2017
Later among the works it cites.
Emergent language in a multi-modal, multi-step referential game
Katrina Evtimova, Andrew Drozdov, Douwe Kiela, and Kyunghyun Cho · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A comprehensive survey of multiagent reinforcement learning
L. Busoniu, R. Babuska, and B. De Schutter · 2008
Cited alongside, same era.
Game Theory: A Multi-Leveled Approach
H. Peters · 2008
Cited alongside, same era.
Game theory of mind
Wako Yoshida, Ray J. Dolan, and Karl J. Friston · 2008
Cited alongside, same era.
Rectified linear units improve restricted boltzmann machines
Vinod Nair and Geoffrey E Hinton · 2010
Cited alongside, same era.
Classes of multiagent q-learning dynamics with ϵ \epsilon -greedy exploration
Michael Wunder, Michael Littman, and Monica Babes · 2010
Cited alongside, same era.
Multiagent learning: Basics, challenges, and prospects
Karl Tuyls and Gerhard Weiss · 2012
Cited alongside, same era.
Later among the works it cites.
Learning with opponent-learning awareness
Jakob N. Foerster, Richard Y. Chen, Maruan Al-Shedivat, Shimon Whiteson, Pieter Abbeel, and Igor Mordatch · 2017
Later among the works it cites.
Emergence of language with multi-agent games: Learning to communicate with sequences of symbols
Serhii Havrylov and Ivan Titov · 2017
Later among the works it cites.
A unified game-theoretic approach to multiagent reinforcement learning
M. Lanctot, Vinicius Zambaldi, Audrunas Gruslys, Angeliki Lazaridou, Karl Tuyls, Julien Perolat, David Silver, and Thore Graepel · 2017
Later among the works it cites.
Multi-agent Reinforcement Learning in Sequential Social Dilemmas
Joel Z. Leibo, Vinicius Zambaldi, Marc Lanctot, Janusz Marecki, and Thore Graepel · 2017
Later among the works it cites.
Deal or no deal? end-to-end learning of negotiation dialogues
Mike Lewis, Denis Yarats, Yann Dauphin, Devi Parikh, and Dhruv Batra · 2017
Later among the works it cites.
Prosocial learning agents solve generalized Stag Hunts better than selfish ones
A. Peysakhovich and A. Lerer · 2017
Later among the works it cites.
Mastering the game of go without human knowledge
David Silver, Julian Schrittwieser, Karen Simonyan, Ioannis Antonoglou, Aja Huang, Arthur Guez, Thomas Hubert, Lucas Baker, Matthew Lai, Adrian Bolton, Yutian Chen, Timothy Lillicrap, Fan Hui, Laurent Sifre, George van den Driessche, Thore Graepel, and Demis Hassabis · 2017
Later among the works it cites.
Cooperating with machines
Jacob W Crandall, Mayada Oudah, Fatimah Ishowo-Oloko, Sherief Abdallah, Jean-François Bonnefon, Manuel Cebrian, Azim Shariff, Michael A Goodrich, Iyad Rahwan, et al · 2018
Closest in time.
Machine Theory of Mind
N. C. Rabinowitz, F. Perbet, H. F. Song, C. Zhang, S. M. A. Eslami, and M. Botvinick · 2018
Closest in time.