Fetching the paper…
Reading the bibliography…
Recent studies have shown that introducing communication between agents can significantly improve overall performance in cooperative Multi-agent reinforcement learning (MARL).
Propagation measurements and models for wireless communications channels
J. B. Andersen, T. S. Rappaport, and S. Yoshida · 1995
Earlier work this paper cites.
Finite-state markov channel-a useful model for radio communication channels
H. S. Wang and N. Moayeri · 1995
Earlier work this paper cites.
Measurement and analysis of the error characteristics of an in-building wireless network
D. Eckhardt and P. Steenkiste · 1996
Earlier work this paper cites.
On channel modeling for delay analysis of packet communications over wireless links
M. Zorzi and R. R. Rao · 1998
Earlier work this paper cites.
Tcp and udp performance over a wireless lan
G. Xylomenos and G. C. Polyzos · 1999
Earlier work this paper cites.
Multi-agent reinforcement learning for traffic light control
M. Wiering · 2000
Earlier work this paper cites.
A markov-based channel model algorithm for wireless networks
A. Konrad, B. Y. Zhao, A. D. Joseph, and R. Ludwig · 2003
Earlier work this paper cites.
F-rto: an enhanced recovery algorithm for tcp retransmission timeouts
P. Sarolahti, M. Kojo, and K. Raatikainen · 2003
Earlier work this paper cites.
Fundamentals of wireless communication
D. Tse and P. Viswanath · 2005
Earlier work this paper cites.
Packet loss characterization in wifi-based long distance networks
A. Sheth, S. Nedevschi, R. Patra, S. Surana, E. Brewer, and L. Subramanian · 2007
Earlier work this paper cites.
Optimal and approximate q-value functions for decentralized pomdps
F. A. Oliehoek, M. T. Spaan, and N. Vlassis · 2008
Earlier work this paper cites.
Diagnosing wireless packet losses in 802.11: Separating collision from weak signal
S. Rayanchu, A. Mishra, D. Agrawal, S. Saha, and S. Banerjee · 2008
Earlier work this paper cites.
Classes of multiagent q-learning dynamics with epsilon-greedy exploration
M. Wunder, M. L. Littman, and M. Babes · 2010
Earlier work this paper cites.
"reinforcement learning in robotics: A survey."
J. Kober, J. A. Bagnell, and J. Peters · 2013
Cited alongside, same era.
"playing atari with deep reinforcement learning."
V. Mnih, K. Kavukcuoglu, D. Silver, A. Graves, I. Antonoglou, D. Wierstra, and M. Riedmiller · 2013
Cited alongside, same era.
Deep recurrent q-learning for partially observable mdps
M. Hausknecht and P. Stone · 2015
Cited alongside, same era.
Learning to communicate with deep multi-agent reinforcement learning
J. Foerster, I. A. Assael, N. de Freitas, and S. Whiteson · 2016
Cited alongside, same era.
"safe, multi-agent, reinforcement learning for autonomous driving."
S.-S. Shai, S. Shammah, and A. Shashua · 2016
Cited alongside, same era.
Multi-agent common knowledge reinforcement learning
J. N. Foerster, C. A. S. de Witt, G. Farquhar, P. H. Torr, W. Boehmer, and S. Whiteson · 2018
Later among the works it cites.
Counterfactual multi-agent policy gradients
J. N. Foerster, G. Farquhar, T. Afouras, N. Nardelli, and S. Whiteson · 2018
Later among the works it cites.
Actor-attention-critic for multi-agent reinforcement learning
S. Iqbal and F. Sha · 2018
Later among the works it cites.
Learning attentional communication for multi-agent cooperation
J. Jiang and Z. Lu · 2018
Later among the works it cites.
Qmix: Monotonic value function factorisation for deep multi-agent reinforcement learning
T. Rashid, M. Samvelyan, C. S. de Witt, G. Farquhar, J. Foerster, and S. Whiteson · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S. Shalev-Shwartz, S. Shammah, and A. Shashua · 2016
Cited alongside, same era.
Learning multiagent communication with backpropagation
S. Sukhbaatar, R. Fergus, et al · 2016
Cited alongside, same era.
A unified game-theoretic approach to multiagent reinforcement learning
M. Lanctot, V. Zambaldi, A. Gruslys, A. Lazaridou, K. Tuyls, J. Pérolat, D. Silver, and T. Graepel · 2017
Cited alongside, same era.
Multi-agent actor-critic for mixed cooperative-competitive environments
R. Lowe, Y. Wu, A. Tamar, J. Harb, O. P. Abbeel, and I. Mordatch · 2017
Cited alongside, same era.
P. Peng, Y. Wen, Y. Yang, Q. Yuan, Z. Tang, H. Long, and J. Wang · 2017
Cited alongside, same era.
Value-decomposition networks for cooperative multi-agent learning
P. Sunehag, G. Lever, A. Gruslys, W. M. Czarnecki, V. Zambaldi, M. Jaderberg, M. Lanctot, N. Sonnerat, J. Z. Leibo, K. Tuyls, et al · 2017
Cited alongside, same era.
Tarmac: Targeted multi-agent communication
A. Das, T. Gervet, J. Romoff, D. Batra, D. Parikh, M. Rabbat, and J. Pineau · 2018
Cited alongside, same era.
Designing multi-agent swarm of uav for precise agriculture
P. Skobelev, D. Budaev, N. Gusev, and G. Voschuk · 2018
Later among the works it cites.
Learning to schedule communication in multi-agent reinforcement learning
D. Kim, S. Moon, D. Hostallero, W. J. Kang, T. Lee, K. Son, and Y. Yi · 2019
Later among the works it cites.
Message-dropout: An efficient training method for multi-agent deep reinforcement learning
W. Kim, M. Cho, and Y. Sung · 2019
Later among the works it cites.
The StarCraft Multi-Agent Challenge
M. Samvelyan, T. Rashid, C. S. de Witt, G. Farquhar, N. Nardelli, T. G. J. Rudner, C.-M. Hung, P. H. S. Torr, J. Foerster, and S. Whiteson · 2019
Later among the works it cites.
Qtran: Learning to factorize with transformation for cooperative multi-agent reinforcement learning
K. Son, D. Kim, W. J. Kang, D. E. Hostallero, and Y. Yi · 2019
Later among the works it cites.
Efficient communication in multi-agent reinforcement learning via variance based control
S. Q. Zhang, Q. Zhang, and J. Lin · 2019
Later among the works it cites.
On the robustness of cooperative multi-agent reinforcement learning
J. Lin, K. Dzeparoska, S. Q. Zhang, A. Leon-Garcia, and N. Papernot · 2020
Closest in time.