Fetching the paper…
Reading the bibliography…
We consider the problem of the limited-bandwidth communication for multi-agent reinforcement learning, where agents cooperate with the assistance of a communication protocol and a scheduler.
A mathematical theory of communication
Shannon, C. E · 1948
Earlier work this paper cites.
Information theory and statistical mechanics
Jaynes, E. T · 1957
Earlier work this paper cites.
The information bottleneck method
Tishby, N., Pereira, F. C., and Bialek, W · 2000
Earlier work this paper cites.
Telecommunication System Engineering , pp. 398–399
Freeman, R · 2004
Earlier work this paper cites.
Elements of Information Theory
Cover, T. M. and Thomas, J. A · 2012
Earlier work this paper cites.
Coordinating multi-agent reinforcement learning with limited communication
Zhang, C. and Lesser, V · 2013
Earlier work this paper cites.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Ioffe, S. and Szegedy, C · 2015
Earlier work this paper cites.
Deep variational information bottleneck
Alemi, A. A., Fischer, I., Dillon, J. V., and Murphy, K · 2016
Earlier work this paper cites.
Learning to communicate with deep multi-agent reinforcement learning
Foerster, J., Assael, I. A., de Freitas, N., and Whiteson, S · 2016
Earlier work this paper cites.
Multi-agent cooperation and the emergence of (natural) language
Lazaridou, A., Peysakhovich, A., and Baroni, M · 2016
Cited alongside, same era.
Learning multiagent communication with backpropagation
Sukhbaatar, S., Fergus, R., et al · 2016
Cited alongside, same era.
Multi-agent actor-critic for mixed cooperative-competitive environments
Lowe, R., Wu, Y., Tamar, A., Harb, J., Abbeel, O. P., and Mordatch, I · 2017
Cited alongside, same era.
Multiagent bidirectionally-coordinated nets for learning to play starcraft combat games
Peng, P., Yuan, Q., Wen, Y., Yang, Y., Tang, Z., Long, H., and Wang, J · 2017
Cited alongside, same era.
Learning attentional communication for multi-agent cooperation
Jiang, J. and Lu, Z · 2018
Cited alongside, same era.
Learning when to communicate at scale in multiagent cooperative and competitive tasks
Singh, A., Jain, T., and Sukhbaatar, S · 2018
Later among the works it cites.
TarMAC: Targeted multi-agent communication
Das, A., Gervet, T., Romoff, J., Batra, D., Parikh, D., Rabbat, M., and Pineau, J · 2019
Closest in time.
Learning to schedule communication in multi-agent reinforcement learning
Kim, D., Moon, S., Hostallero, D., Kang, W. J., Lee, T., Son, K., and Yi, Y · 2019
Closest in time.
On the pitfalls of measuring emergent communication
Lowe, R., Foerster, J., Boureau, Y.-L., Pineau, J., and Dauphin, Y · 2019
Closest in time.
Learning multi-agent communication under limited-bandwidth restriction for internet packet routing
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kilinc, O. and Montana, G · 2018
Cited alongside, same era.
Emergence of grounded compositional language in multi-agent populations
Mordatch, I. and Abbeel, P · 2018
Cited alongside, same era.
Qmix: monotonic value function factorisation for deep multi-agent reinforcement learning
Rashid, T., Samvelyan, M., De Witt, C. S., Farquhar, G., Foerster, J., and Whiteson, S · 2018
Cited alongside, same era.
Mao, H., Gong, Z., Zhang, Z., Xiao, Z., and Ni, Y · 2019
Closest in time.
OpenAI Five
OpenAI · 2019
Closest in time.
Robocup Federation Official Website
RoboCup · 2019
Closest in time.
The StarCraft Multi-Agent Challenge
Samvelyan, M., Rashid, T., de Witt, C. S., Farquhar, G., Nardelli, N., Rudner, T. G. J., Hung, C.-M., Torr, P. H. S., Foerster, J., and Whiteson, S · 2019
Closest in time.