Fetching the paper…
Reading the bibliography…
Communicating with each other in a distributed manner and behaving as a group are essential in multi-agent reinforcement learning.
Learning to schedule communication in multi-agent reinforcement learning
Kim, D.; Moon, S.; Hostallero, D.; Kang, W. J.; Lee, T.; Son, K.; and Yi, Y. 2019 · 1902
Earlier work this paper cites.
On the pitfalls of measuring emergent communication
Lowe, R.; Foerster, J.; Boureau, Y.-L.; Pineau, J.; and Dauphin, Y. 2019 · 1903
Earlier work this paper cites.
Learning efficient multi-agent communication: an information bottleneck approach
Wang, R.; He, X.; Yu, R.; Qiu, W.; An, B.; and Rabinovich, Z. 2019 · 1911
Earlier work this paper cites.
Learning agent communication under limited bandwidth by message pruning
Mao, H.; Zhang, Z.; Xiao, Z.; Gong, Z.; and Ni, Y. 2019 · 1912
Earlier work this paper cites.
A mathematical theory of communication
Shannon, C. E. 1948 · 1948
Earlier work this paper cites.
The principle of maximum entropy
Guiasu, S.; and Shenitzer, A. 1985 · 1985
Earlier work this paper cites.
Telecommunication System Engineering
Freeman, R. L. 2004 · 2004
Earlier work this paper cites.
Improving coordination with communication in multi-agent reinforcement learning
Szer, D.; and Charpillet, F. 2004 · 2004
Earlier work this paper cites.
Distributed Event-Triggered Control for Multi-Agent Systems
Dimarogonas, D. V.; Frazzoli, E.; and Johansson, K. H. 2012 · 2012
Earlier work this paper cites.
Coordinating multi-agent reinforcement learning with limited communication
Zhang, C.; and Lesser, V. 2013 · 2013
Earlier work this paper cites.
Learning to communicate with deep multi-agent reinforcement learning
Foerster, J.; Assael, I. A.; De Freitas, N.; and Whiteson, S. 2016 · 2016
Earlier work this paper cites.
End-to-end training of deep visuomotor policies
Levine, S.; Finn, C.; Darrell, T.; and Abbeel, P. 2016 · 2016
Earlier work this paper cites.
Learning multiagent communication with backpropagation
Sukhbaatar, S.; Fergus, R.; et al. 2016 · 2016
Cited alongside, same era.
Deep reinforcement learning with visual attention for vehicle classification
Zhao, D.; Chen, Y.; and Lv, L. 2016 · 2016
Cited alongside, same era.
Event-triggered optimal control for partially unknown constrained-input systems via adaptive dynamic programming
Zhu, Y.; Zhao, D.; He, H.; and Ji, J. 2016 · 2016
Cited alongside, same era.
Chu, X.; and Ye, H. 2017 · 2017
Cited alongside, same era.
A survey of learning in multiagent environments: Dealing with non-stationarity
Hernandez-Leal, P.; Kaisers, M.; Baarslag, T.; and de Cote, E. M. 2017 · 2017
Cited alongside, same era.
Multi-agent deep reinforcement learning with extremely noisy observations
Kilinc, O.; and Montana, G. 2018 · 2018
Later among the works it cites.
Learning to communicate via supervised attentional message processing
Peng, Z.; Zhang, L.; and Luo, T. 2018 · 2018
Later among the works it cites.
QMIX: monotonic value function factorisation for deep multi-agent reinforcement learning
Rashid, T.; Samvelyan, M.; Schroeder, C.; Farquhar, G.; Foerster, J.; and Whiteson, S. 2018 · 2018
Later among the works it cites.
Starcraft micromanagement with reinforcement learning and curriculum transfer learning
Shao, K.; Zhu, Y.; and Zhao, D. 2018 · 2018
Later among the works it cites.
Learning when to communicate at scale in multiagent cooperative and competitive tasks
Singh, A.; Jain, T.; and Sukhbaatar, S. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Multi-agent actor-critic for mixed cooperative-competitive environments
Lowe, R.; Wu, Y. I.; Tamar, A.; Harb, J.; Abbeel, O. P.; and Mordatch, I. 2017 · 2017
Cited alongside, same era.
Mao, H.; Gong, Z.; Ni, Y.; and Xiao, Z. 2017 · 2017
Cited alongside, same era.
Peng, P.; Wen, Y.; Yang, Y.; Yuan, Q.; Tang, Z.; Long, H.; and Wang, J. 2017 · 2017
Cited alongside, same era.
Mastering the game of Go without human knowledge
Silver, D.; Schrittwieser, J.; Simonyan, K.; Antonoglou, I.; Huang, A.; Guez, A.; Hubert, T.; Baker, L.; Lai, M.; Bolton, A.; et al. 2017 · 2017
Cited alongside, same era.
Value-decomposition networks for cooperative multi-agent learning
Sunehag, P.; Lever, G.; Gruslys, A.; Czarnecki, W. M.; Zambaldi, V.; Jaderberg, M.; Lanctot, M.; Sonnerat, N.; Leibo, J. Z.; Tuyls, K.; et al. 2017 · 2017
Cited alongside, same era.
Counterfactual multi-agent policy gradients
Foerster, J. N.; Farquhar, G.; Afouras, T.; Nardelli, N.; and Whiteson, S. 2018 · 2018
Cited alongside, same era.
Learning attentional communication for multi-agent cooperation
Jiang, J.; and Lu, Z. 2018 · 2018
Cited alongside, same era.
Learning to cooperate via an attention-based communication neural network in decentralized multi-robot exploration
Geng, M.; Xu, K.; Zhou, X.; Ding, B.; Wang, H.; and Zhang, L. 2019 · 2019
Later among the works it cites.
Message-dropout: An efficient training method for multi-agent deep reinforcement learning
Kim, W.; Cho, M.; and Sung, Y. 2019 · 2019
Later among the works it cites.
Grandmaster level in StarCraft II using multi-agent reinforcement learning
Vinyals, O.; Babuschkin, I.; Czarnecki, W. M.; Mathieu, M.; Dudzik, A.; Chung, J.; Choi, D. H.; Powell, R. E.; Ewalds, T.; Georgiev, P.; et al. 2019 · 2019
Later among the works it cites.
Factorized Q-learning for large-scale multi-agent systems
Zhou, M.; Chen, Y.; Wen, Y.; Yang, Y.; Su, Y.; Zhang, W.; Zhang, D.; and Wang, J. 2019 · 2019
Later among the works it cites.
Learning multi-agent communication with double attentional deep reinforcement learning
Mao, H.; Zhang, Z.; Xiao, Z.; Gong, Z.; and Ni, Y. 2020 · 2020
Closest in time.
Improving coordination in small-scale multi-agent deep reinforcement learning through memory-driven communication
Pesce, E.; and Montana, G. 2020 · 2020
Closest in time.
Multi agent deep learning with cooperative communication
Simões, D.; Lau, N.; and Reis, L. P. 2020 · 2020
Closest in time.