Fetching the paper…
Reading the bibliography…
How do we know if communication is emerging in a multi-agent system? The vast majority of recent papers on emergent communication show that adding a communication channel leads to an increase in reward or task success.
Principal components analysis
Karl Pearson. 1901 · 1901
Earlier work this paper cites.
Asynchronous methods for deep reinforcement learning. In International Conference on Machine Learning
Volodymyr Mnih, Adria Puigdomenech Badia, Mehdi Mirza, Alex Graves, Timothy Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu. 2016 · 1937
Earlier work this paper cites.
Philosophical investigations
Ludwig Wittgenstein. 1953 · 1953
Earlier work this paper cites.
Subjectivity and correlation in randomized strategies
Robert J Aumann. 1974 · 1974
Earlier work this paper cites.
Mate selection—a selection for a handicap
Amotz Zahavi. 1975 · 1975
Earlier work this paper cites.
Honest signalling: The Philip Sidney game
John M Smith. 1991 · 1991
Earlier work this paper cites.
Function optimization using connectionist reinforcement learning algorithms
Ronald J Williams and Jing Peng. 1991 · 1991
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald J Williams. 1992 · 1992
Earlier work this paper cites.
The mathematics of statistical machine translation: Parameter estimation
Peter F Brown, Vincent J Della Pietra, Stephen A Della Pietra, and Robert L Mercer. 1993 · 1993
Earlier work this paper cites.
Markov games as a framework for multi-agent reinforcement learning. In Proceedings of the eleventh international conference on machine learning
Michael L Littman. 1994 · 1994
Earlier work this paper cites.
Conversation and cooperation in social dilemmas: A meta-analysis of experiments from 1958 to 1992
David Sally. 1995 · 1995
Earlier work this paper cites.
Cheap talk
Joseph Farrell and Matthew Rabin. 1996 · 1996
Earlier work this paper cites.
The evolution of language
Martin A Nowak and David C Krakauer. 1999 · 1999
Earlier work this paper cites.
Policy gradient methods for reinforcement learning with function approximation. In Advances in neural information processing systems
Richard S Sutton, David A McAllester, Satinder P Singh, and Yishay Mansour. 2000 · 2000
Earlier work this paper cites.
Costly signaling and cooperation
Herbert Gintis, Eric Alden Smith, and Samuel Bowles. 2001 · 2001
Earlier work this paper cites.
Progress in the simulation of emergent communication and language
Kyle Wagner, James A Reggia, Juan Uriagereka, and Gerald S Wilkinson. 2003 · 2003
Earlier work this paper cites.
Evolutionary dynamics of Lewis signaling games: signaling systems vs. partial pooling
Simon M Huttegger, Brian Skyrms, Rory Smead, and Kevin JS Zollman. 2010 · 2010
Cited alongside, same era.
Intriguing properties of neural networks
Christian Szegedy, Wojciech Zaremba, Ilya Sutskever, Joan Bruna, Dumitru Erhan, Ian Goodfellow, and Rob Fergus. 2013 · 2013
Cited alongside, same era.
Generative adversarial nets. In Advances in neural information processing systems
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio. 2014 · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba. 2014 · 2014
Cited alongside, same era.
Deep learning
Yann LeCun, Yoshua Bengio, and Geoffrey Hinton. 2015 · 2015
Cited alongside, same era.
Learning cooperative visual dialog agents with deep reinforcement learning
Abhishek Das, Satwik Kottur, José MF Moura, Stefan Lee, and Dhruv Batra. 2017 · 2017
Later among the works it cites.
Emergent Communication in a Multi-Modal, Multi-Step Referential Game
Katrina Evtimova, Andrew Drozdov, Douwe Kiela, and Kyunghyun Cho. 2017 · 2017
Later among the works it cites.
Emergence of language with multi-agent games: learning to communicate with sequences of symbols. In Advances in Neural Information Processing Systems
Serhii Havrylov and Ivan Titov. 2017 · 2017
Later among the works it cites.
Natural Language Does Not Emerge’Naturally’in Multi-Agent Dialog
Satwik Kottur, José MF Moura, Stefan Lee, and Dhruv Batra. 2017 · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Trust region policy optimization. In International Conference on Machine Learning
John Schulman, Sergey Levine, Pieter Abbeel, Michael Jordan, and Philipp Moritz. 2015 · 2015
Cited alongside, same era.
Understanding intermediate layers using linear classifier probes
Guillaume Alain and Yoshua Bengio. 2016 · 2016
Cited alongside, same era.
Learning to Communicate to Solve Riddles with Deep Distributed Recurrent Q-Networks
Jakob N. Foerster, Yannis M. Assael1, Nando de Freitas, and Shimon Whiteson. 2016 · 2016
Cited alongside, same era.
A paradigm for situated and goal-driven language learning
Jon Gauthier and Igor Mordatch. 2016 · 2016
Cited alongside, same era.
Reinforcement learning with unsupervised auxiliary tasks
Max Jaderberg, Volodymyr Mnih, Wojciech Marian Czarnecki, Tom Schaul, Joel Z Leibo, David Silver, and Koray Kavukcuoglu. 2016 · 2016
Cited alongside, same era.
Multi-agent cooperation and the emergence of (natural) language
Angeliki Lazaridou, Alexander Peysakhovich, and Marco Baroni. 2016 · 2016
Cited alongside, same era.
Causal inference in statistics: a primer
Judea Pearl, Madelyn Glymour, and Nicholas P Jewell. 2016 · 2016
Cited alongside, same era.
Jason Lee, Kyunghyun Cho, Jason Weston, and Douwe Kiela. 2017 · 2017
Later among the works it cites.
Emergence of Grounded Compositional Language in Multi-Agent Populations
Igor Mordatch and Pieter Abbeel. 2017 · 2017
Later among the works it cites.
Automatic differentiation in pytorch
Adam Paszke, Sam Gross, Soumith Chintala, Gregory Chanan, Edward Yang, Zachary DeVito, Zeming Lin, Alban Desmaison, Luca Antiga, and Adam Lerer. 2017 · 2017
Later among the works it cites.
Emergence of communication in an interactive world with consistent speakers
Ben Bogin, Mor Geva, and Jonathan Berant. 2018 · 2018
Later among the works it cites.
Emergent Communication through Negotiation
Kris Cao, Angeliki Lazaridou, Marc Lanctot, Joel Z Leibo, Karl Tuyls, and Stephen Clark. 2018 · 2018
Later among the works it cites.
Compositional Obverter Communication Learning From Raw Visual Input
Edward Choi, Angeliki Lazaridou, and Nando de Freitas. 2018 · 2018
Later among the works it cites.
Bayesian action decoding for deep multi-agent reinforcement learning
Jakob Foerster, Francis Song, Edward Hughes, Neil Burch, Iain Dunning, Shimon Whiteson, Matthew Botvinick, and Michael Bowling. 2018 · 2018
Later among the works it cites.
Intrinsic Social Motivation via Causal Influence in Multi-Agent RL
Natasha Jaques, Angeliki Lazaridou, Edward Hughes, Caglar Gulcehre, Pedro A Ortega, DJ Strouse, Joel Z Leibo, and Nando de Freitas. 2018 · 2018
Later among the works it cites.
Emergence of linguistic communication from referential games with symbolic and pixel input
Angeliki Lazaridou, Karl Moritz Hermann, Karl Tuyls, and Stephen Clark. 2018 · 2018
Later among the works it cites.
Learning Social Conventions in Markov Games
Adam Lerer and Alexander Peysakhovich. 2018 · 2018
Later among the works it cites.
Understanding Agent Incentives using Causal Influence Diagrams, Part I: Single Action Settings
Tom Everitt, Pedro A Ortega, Elizabeth Barnes, and Shane Legg. 2019 · 2019
Closest in time.