Fetching the paper…
Reading the bibliography…
Humans use language to collectively execute abstract strategies besides using it as a referential tool for identifying physical entities.
Convention
David Lewis · 1969
Earlier work this paper cites.
Strategic information transmission
V.P. Crawford and J. Sobel · 1982
Earlier work this paper cites.
Simple statistical gradient following algorithms for connectionist reinforcement learning
Ronald J Williams · 1992
Earlier work this paper cites.
Multi-agent reinforcement learning: Independent vs. cooperative agents
Ming Tan · 1993
Earlier work this paper cites.
Markov games as a framework for multi-agent reinforcement learning
Michael L. Littman · 1994
Earlier work this paper cites.
Cheap talk
J. Farrell and M. Rabin · 1996
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber · 1997
Earlier work this paper cites.
Random Geometric Graphs
Mathew Penrose · 2003
Earlier work this paper cites.
Books about us politics
Valdis Krebs · 2004
Earlier work this paper cites.
How did language go discrete
Michael Studdert-Kennedy · 2005
Earlier work this paper cites.
Visualizing data using t-sne
Laurens van der Maaten and Geoffrey Hinton · 2008
Earlier work this paper cites.
Speech and Language Processing : An Introduction to Natural Language Processing, Computational Linguistics, and Speech Recognition - Chapter 23
Dan Jurafsky and James H. Martin · 2009
Cited alongside, same era.
Multi-agent reinforcement learning: An overview
Lucian Busoniu, Robert Babuska, and Bart De Schutter · 2010
Cited alongside, same era.
Auto-encoding variational bayes
Diederik P. Kingma and Max Welling · 2013
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba · 2014
Cited alongside, same era.
Sapiens: A Brief History of Humankind
Yuval Noah Harari · 2015
Cited alongside, same era.
Fast and accurate deep network learning by exponential linear units (elus)
Djork-Arne Clevert, Thomas Unterthiner, and Sepp Hochreiter · 2016
Cited alongside, same era.
Categorical reparameterization with gumbel-softmax
Eric Jang, Shixiang Gu, and Ben Poole · 2017
Later among the works it cites.
The concrete distribution: A continuous relaxation of discrete random variables
Chris J Maddison, Andriy Mnih, and Yee Whye Teh · 2017
Later among the works it cites.
Semi-supervised classification with graph convolutional networks
Thomas N. Kipf and Max Welling · 2017
Later among the works it cites.
Learning cooperative visual dialog agents with deep reinforcement learning
A. Das, S. Kottur, J. M. F. Moura, S. Lee, and D. Batra · 2017
Later among the works it cites.
Emergence of grounded compositional language in multi-agent populations
Igor Mordatch and Pieter Abbeel · 2018
Later among the works it cites.
Emergent communication through negotiation
Kris Cao, Angeliki Lazaridou, Marc Lanctot, Joel Z Leibo, Karl Tuyls, and Stephen Clark · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Learning multiagent communication with backpropagation
Sainbayar Sukhbaatar, Arthur Szlam, and Rob Fergus · 2016
Cited alongside, same era.
Learning to communicate with deep multi-agent reinforcement learning
Jakob Foerster, Ioannis Alexandros Assael, Nando de Freitas, and Shimon Whiteson · 2016
Cited alongside, same era.
A paradigm for situated and goal driven language learning
Jon Gauthier and Igor Mordatch · 2016
Cited alongside, same era.
Multi-agent cooperation and the emergence of (natural) language
Angeliki Lazaridou, Alexander Peysakhovich, and Marco Baroni · 2017
Cited alongside, same era.
Emergence of language with multi-agent games: Learning to communicate with sequences of symbols
Serhii Havrylov and Ivan Titov · 2017
Cited alongside, same era.
Later among the works it cites.
Emergence of communication in an interactive world with consistent speakers
Ben Bogin, Mor Geva, and Jonathan Berant · 2018
Later among the works it cites.
Autonomously reusing knowledge in multiagent reinforcement learning
Felipe Leno Da Silva, Matthew E. Taylor, and Anna Helena Reali Costa · 2018
Later among the works it cites.
Emergence of linguistic communication from referential games with symbolic and pixel input
Angeliki Lazaridou, Karl Moritz Hermann, Karl Tuyls, and Stephen Clark · 2018
Later among the works it cites.
Fully decentralized multi-agent reinforcement learning with networked agents
Kaiqing Zhang, Zhuoran Yang, Han Liu, Tong Zhang, and Tamer Basar · 2018
Later among the works it cites.
Tarmac: Targeted multi-agent communication
Abhishek Das, Théophile Gervet, Joshua Romoff, Dhruv Batra, Devi Parikh, Mike Rabbat, and Joelle Pineau · 2019
Closest in time.