Fetching the paper…
Reading the bibliography…
The literature in modern machine learning has only negative results for learning to communicate between competitive agents using standard RL.
Convention: A Philosophical Study
David Lewis. 1969 · 1969
Earlier work this paper cites.
Animal signals: Information or Manipulation?
Richard Dawkins and John R. Krebs. 1978 · 1978
Earlier work this paper cites.
Animal Signals: Ethological and Games-Theory Approaches are Not Incompatible
Robert A. Hinde. 1981 · 1981
Earlier work this paper cites.
Strategic Information Transmission
Vincent P. Crawford and Joel Sobel. 1982 · 1982
Earlier work this paper cites.
An experimental analysis of ultimatum bargaining
Werner Güth, Rolf Schmittberger, and Bernd Schwarze. 1982 · 1982
Earlier work this paper cites.
Animal Signals: Mind Reading and Manipulation
John R. Krebs and Richard Dawkins. 1984 · 1984
Earlier work this paper cites.
Simple Statistical Gradient-Following Algorithms for Connectionist Reinforcement Learning
Ronald J. Williams. 1992 · 1992
Earlier work this paper cites.
Cheap Talk
Joseph Farrell and Matthew Rabin. 1996 · 1996
Earlier work this paper cites.
"Other-Play" for Zero-Shot Coordination
Hengyuan Hu, Adam Lerer, Alex Peysakhovich, and Jakob Foerster. 2020 · 2003
Earlier work this paper cites.
Multi-Agent Reinforcement Learning:a critical survey
Yoav Shoham, Rob Powers, and Trond Grenager. 2003 · 2003
Earlier work this paper cites.
Progress in the simulation of emergent communication and language
Kyle Wagner, James A Reggia, Juan Uriagereka, and Gerald S Wilkinson. 2003 · 2003
Earlier work this paper cites.
Chapter 19 Gradient Estimation
Michael Fu. 2006 · 2006
Earlier work this paper cites.
Rectified Linear Units Improve Restricted Boltzmann Machines Vinod Nair. In ICML , Vol. 27. 807–814
Vinod Nair and Geoffrey Hinton. 2010 · 2010
Earlier work this paper cites.
Signals: Evolution, Learning, & Information
Brian Skyrms. 2010 · 2010
Earlier work this paper cites.
Bayesian persuasion
Emir Kamenica and Matthew Gentzkow. 2011 · 2011
Earlier work this paper cites.
Deterministic Chaos and the Evolution of Meaning
Elliott O. Wagner. 2012 · 2012
Earlier work this paper cites.
Adam: A Method for Stochastic Optimization
Diederik P. Kingma and Jimmy Ba. 2014 · 2014
Earlier work this paper cites.
Gradient Estimation Using Stochastic Computation Graphs. In NIPS
John Schulman, Nicolas Manfred Otto Heess, Theophane Weber, and Pieter Abbeel. 2015 · 2015
Cited alongside, same era.
Learning to communicate with deep multi-agent reinforcement learning. In Advances in Neural Information Processing Systems . 2137–2145
Jakob Foerster, Ioannis Alexandros Assael, Nando de Freitas, and Shimon Whiteson. 2016 · 2016
Cited alongside, same era.
Deep Learning
Ian Goodfellow, Yoshua Bengio, and Aaron Courville. 2016 · 2016
Cited alongside, same era.
Multi-agent cooperation and the emergence of (natural) language
Angeliki Lazaridou, Alexander Peysakhovich, and Marco Baroni. 2016 · 2016
Cited alongside, same era.
Self-Assembling Games
Jeffrey A. Barrett and Brian Skyrms. 2017 · 2017
Cited alongside, same era.
Emergent communication in a multi-modal, multi-step referential game
Learning when to Communicate at Scale in Multiagent Cooperative and Competitive Tasks. In ICLR
Amanpreet Singh, Tushar Jain, and Sainbayar Sukhbaatar. 2018 · 2018
Later among the works it cites.
Propositional Content in Signals
Brian Skyrms and Jeffrey A. Barrett. 2018 · 2018
Later among the works it cites.
Reinforcement Learning: An Introduction (second ed.)
Richard S. Sutton and Andrew G. Barto. 2018 · 2018
Later among the works it cites.
The Hanabi Challenge: A New Frontier for AI Research
Nolan Bard, Jakob N. Foerster, A. P. Sarath Chandar, Neil Burch, Marc Lanctot, H. Francis Song, Emilio Parisotto, Vincent Dumoulin, Subhodeep Moitra, Edward Hughes, Iain Dunning, Shibl Mourad, Hugo Larochelle, Marc G. Bellemare, and Michael H. Bowling. 2019 · 2019
Later among the works it cites.
Oríon - Asynchronous Distributed Hyperparameter Optimization
Xavier Bouthillier, Christos Tsirigotis, François Corneau-Tremblay, Pierre Delaunay, Michael Noukhovitch, Reyhane Askari, Peter Henderson, Dendi Suhubdy, Frédéric Bastien, and Pascal Lamblin. 2019 · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Katrina Evtimova, Andrew Drozdov, Douwe Kiela, and Kyunghyun Cho. 2017 · 2017
Cited alongside, same era.
Emergence of language with multi-agent games: Learning to communicate with sequences of symbols. In Advances in neural information processing systems . 2149–2159
Serhii Havrylov and Ivan Titov. 2017 · 2017
Cited alongside, same era.
Categorical Reparameterization with Gumbel-Softmax. In ICLR 2017 : International Conference on Learning Representations 2017
Eric Jang, Shixiang Gu, and Ben Poole. 2017 · 2017
Cited alongside, same era.
A Unified Game-Theoretic Approach to Multiagent Reinforcement Learning. In NIPS
Marc Lanctot, Vinicius Zambaldi, Audrunas Gruslys, Angeliki Lazaridou, Karl Tuyls, Julien Perolat, David Silver, and Thore Graepel. 2017 · 2017
Cited alongside, same era.
Multi-agent Reinforcement Learning in Sequential Social Dilemmas
Joel Z. Leibo, Vinícius Flores Zambaldi, Marc Lanctot, Janusz Marecki, and Thore Graepel. 2017 · 2017
Cited alongside, same era.
The Concrete Distribution: A Continuous Relaxation of Discrete Random Variables. In ICLR 2017 : International Conference on Learning Representations 2017
Chris J. Maddison, Andriy Mnih, and Yee Whye Teh. 2017 · 2017
Cited alongside, same era.
Emergence of Grounded Compositional Language in Multi-Agent Populations
Igor Mordatch and Pieter Abbeel. 2017 · 2017
Cited alongside, same era.
Anti-efficient encoding in emergent communication. In NeurIPS 2019 : Thirty-third Conference on Neural Information Processing Systems . 6293–6303
Rahma Chaabouni, Eugene Kharitonov, Emmanuel Dupoux, and Marco Baroni. 2019 · 2019
Later among the works it cites.
TarMAC: Targeted Multi-Agent Communication. In ICML
Abhishek Das, Théophile Gervet, Joshua Romoff, Dhruv Batra, Devi Parikh, Mike Rabbat, and Joelle Pineau. 2019 · 2019
Later among the works it cites.
Biases for Emergent Communication in Multi-agent Reinforcement Learning. In NeurIPS 2019 : Thirty-third Conference on Neural Information Processing Systems . 13111–13121
Tom Eccles, Yoram Bachrach, Guy Lever, Angeliki Lazaridou, and Thore Graepel. 2019 · 2019
Later among the works it cites.
On the Pitfalls of Measuring Emergent Communication. In AAMAS
Ryan Lowe, Jakob N. Foerster, Y-Lan Boureau, Joelle Pineau, and Yann Dauphin. 2019 · 2019
Later among the works it cites.
PyTorch: An Imperative Style, High-Performance Deep Learning Library
Adam Paszke, Sam Gross, Francisco Massa, Adam Lerer, James Bradbury, Gregory Chanan, Trevor Killeen, Zeming Lin, Natalia Gimelshein, Luca Antiga, Alban Desmaison, Andreas Kopf, Edward Yang, Zachary DeVito, Martin Raison, Alykhan Tejani, Sasank Chilamkurthy, Benoit Steiner, Lu Fang, Junjie Bai, and Soumith Chintala. 2019 · 2019
Later among the works it cites.
Experiment Tracking with Weights and Biases
Lukas Biewald. 2020 · 2020
Later among the works it cites.
Compositionality and Generalization In Emergent Languages. In ACL 2020: 58th annual meeting of the Association for Computational Linguistics . Association for Computational Linguistics, 4427–4442
Rahma Chaabouni, Eugene Kharitonov, Diane Bouchacourt, Emmanuel Dupoux, and Marco Baroni. 2020 · 2020
Later among the works it cites.
Entropy Minimization In Emergent Languages. In ICML 2020: 37th International Conference on Machine Learning
Diane Bouchacourt Evgeny Kharitonov, Rahma Chaabouni and Marco Baroni. 2020 · 2020
Later among the works it cites.
Communicative Bottlenecks Lead to Maximal Information Transfer
Travis LaCroix. 2020 · 2020
Later among the works it cites.
Emergent Multi-Agent Communication in the Deep Learning Era
Angeliki Lazaridou and Marco Baroni. 2020 · 2020
Later among the works it cites.
On Emergent Communication in Competitive Multi-Agent Teams
Paul Pu Liang, Jeffrey Chen, Ruslan Salakhutdinov, Louis-Philippe Morency, and Satwik Kottur. 2020 · 2020
Later among the works it cites.