Fetching the paper…
Reading the bibliography…
This paper considers cooperative Multi-Agent Reinforcement Learning, focusing on emergent communication in settings where multiple pairs of independent learners interact at varying frequencies.
“Stochastic Games*”
L. Shapley · 1953
Earlier work this paper cites.
“Multi-agent reinforcement learning: Independent vs. cooperative agents”
Ming Tan · 1993
Earlier work this paper cites.
“Catastrophic forgetting in connectionist networks”
Robert French · 1999
Earlier work this paper cites.
“MNIST handwritten digit database”, http://yann.lecun.com/exdb/mnist/, 2010
Yann LeCun and Corinna Cortes · 2010
Earlier work this paper cites.
“DeCAF: A Deep Convolutional Activation Feature for Generic Visual Recognition”
Jeff Donahue et al · 2014
Earlier work this paper cites.
“Learning to Communicate with Deep Multi-Agent Reinforcement Learning”
Jakob. Foerster, Yannis. Assael, Nando de Freitas and Shimon Whiteson · 2016
Earlier work this paper cites.
“Learning Multiagent Communication with Backpropagation”
Sainbayar Sukhbaatar, Arthur Szlam and Rob Fergus · 2016
Cited alongside, same era.
“Social influence as intrinsic motivation for multi-agent deep reinforcement learning”
Natasha Jaques et al · 2019
Cited alongside, same era.
“Biases for Emergent Communication in Multi-agent Reinforcement Learning”
Tom Eccles et al · 2019
Cited alongside, same era.
“Tarmac: Targeted multi-agent communication”
Abhishek Das et al · 2019
Cited alongside, same era.
“On the pitfalls of measuring emergent communication”
Ryan Lowe et al · 2019
Cited alongside, same era.
“PyTorch: An Imperative Style, High-Performance Deep Learning Library”
Adam Paszke et al · 2019
Cited alongside, same era.
“Graph Convolutional Reinforcement Learning”
Jiechuan Jiang, Chen Dun, Tiejun Huang and Zongqing Lu · 2020
Later among the works it cites.
““Other-Play” for Zero-Shot Coordination”
Hengyuan Hu, Adam Lerer, Alex Peysakhovich and Jakob Foerster · 2020
Later among the works it cites.
“A New Formalism, Method and Open Issues for Zero-Shot Coordination”
Johannes Treutlein, Michael Dennis, Caspar Oesterheld and Jakob Foerster · 2021
Closest in time.
“Quasi-equivalence discovery for zero-shot emergent communication”
Kalesha Bullard et al · 2021
Closest in time.
“A continual learning survey: Defying forgetting in classification tasks”
Matthias Delange et al · 2021
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…