Fetching the paper…
Reading the bibliography…
Decentralized cooperation in partially-observable multi-agent systems requires effective communications among agents.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald J Williams. 1992 · 1992
Earlier work this paper cites.
Multi-agent reinforcement learning: Independent vs. cooperative agents. In Proceedings of the tenth international conference on machine learning . 330–337
Ming Tan. 1993 · 1993
Earlier work this paper cites.
Binaryconnect: Training deep neural networks with binary weights during propagations. In Advances in neural information processing systems . 3123–3131
Matthieu Courbariaux, Yoshua Bengio, and Jean-Pierre David. 2015 · 2015
Earlier work this paper cites.
Deep recurrent q-learning for partially observable mdps. In 2015 aaai fall symposium series
Matthew Hausknecht and Peter Stone. 2015 · 2015
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, et al · 2015
Earlier work this paper cites.
Variable rate image compression with recurrent neural networks
George Toderici, Sean M O’Malley, Sung Jin Hwang, Damien Vincent, David Minnen, Shumeet Baluja, Michele Covell, and Rahul Sukthankar. 2015 · 2015
Earlier work this paper cites.
Learning to communicate with Deep multi-agent reinforcement learning. In Proceedings of the 30th International Conference on Neural Information Processing Systems . 2145–2153
Jakob N Foerster, Yannis M Assael, Nando de Freitas, and Shimon Whiteson. 2016 · 2016
Earlier work this paper cites.
Learning multiagent communication with backpropagation
Sainbayar Sukhbaatar and Rob Fergus. 2016 · 2016
Earlier work this paper cites.
VAIN: attentional multi-agent predictive modeling. In Proceedings of the 31st International Conference on Neural Information Processing Systems . 2698–2708
Yedid Hoshen. 2017 · 2017
Earlier work this paper cites.
Natural Language Does Not Emerge ‘Naturally’in Multi-Agent Dialog. In Proceedings of the 2017 Conference on Empirical Methods in Natural Language Processing . 2962–2967
Satwik Kottur, José Moura, Stefan Lee, and Dhruv Batra. 2017 · 2017
Earlier work this paper cites.
Peng Peng, Ying Wen, Yaodong Yang, Quan Yuan, Zhenkun Tang, Haitao Long, and Jun Wang. 2017 · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov. 2017 · 2017
Cited alongside, same era.
Deep Communicating Agents for Abstractive Summarization. In Proceedings of the 2018 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long Papers) . 1662–1675
Asli Celikyilmaz, Antoine Bosselut, Xiaodong He, and Yejin Choi. 2018 · 2018
Cited alongside, same era.
Learning Attentional Communication for Multi-Agent Cooperation
Jiechuan Jiang and Zongqing Lu. 2018 · 2018
Cited alongside, same era.
Emergence of grounded compositional language in multi-agent populations. In Thirty-second AAAI conference on artificial intelligence
Igor Mordatch and Pieter Abbeel. 2018 · 2018
Cited alongside, same era.
Qmix: Monotonic value function factorisation for deep multi-agent reinforcement learning. In International Conference on Machine Learning . PMLR, 4295–4304
The StarCraft Multi-Agent Challenge. In Proceedings of the 18th International Conference on Autonomous Agents and MultiAgent Systems . 2186–2188
Mikayel Samvelyan, Tabish Rashid, Christian Schroeder de Witt, Gregory Farquhar, Nantas Nardelli, Tim GJ Rudner, Chia-Man Hung, Philip HS Torr, Jakob Foerster, and Shimon Whiteson. 2019 · 2019
Later among the works it cites.
Distributed reinforcement learning for multi-robot decentralized collective construction
Guillaume Sartoretti, Yue Wu, William Paivine, TK Satish Kumar, Sven Koenig, and Howie Choset. 2019 · 2019
Later among the works it cites.
Open problems in cooperative AI
Allan Dafoe, Edward Hughes, Yoram Bachrach, Tantum Collins, Kevin R McKee, Joel Z Leibo, Kate Larson, and Thore Graepel. 2020 · 2020
Later among the works it cites.
Communication learning via backpropagation in discrete channels with unknown noise. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 34. 7160–7168
Benjamin Freed, Guillaume Sartoretti, Jiaheng Hu, and Howie Choset. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Tabish Rashid, Mikayel Samvelyan, Christian Schroeder, Gregory Farquhar, Jakob Foerster, and Shimon Whiteson. 2018 · 2018
Cited alongside, same era.
Learning when to Communicate at Scale in Multiagent Cooperative and Competitive Tasks. In International Conference on Learning Representations
Amanpreet Singh, Tushar Jain, and Sainbayar Sukhbaatar. 2018 · 2018
Cited alongside, same era.
Alphastar: An evolutionary computation perspective. In Proceedings of the genetic and evolutionary computation conference companion . 314–315
Kai Arulkumaran, Antoine Cully, and Julian Togelius. 2019 · 2019
Cited alongside, same era.
Dota 2 with large scale deep reinforcement learning
Christopher Berner, Greg Brockman, Brooke Chan, Vicki Cheung, Przemysław Dębiak, Christy Dennison, David Farhi, Quirin Fischer, Shariq Hashme, Chris Hesse, et al · 2019
Cited alongside, same era.
Tarmac: Targeted multi-agent communication. In International Conference on Machine Learning . PMLR, 1538–1546
Abhishek Das, Théophile Gervet, Joshua Romoff, Dhruv Batra, Devi Parikh, Mike Rabbat, and Joelle Pineau. 2019 · 2019
Cited alongside, same era.
Learning to Schedule Communication in Multi-agent Reinforcement Learning. In ICLR 2019: International Conference on Representation Learning
Daewoo Kim, Sangwoo Moon, David Hostallero, Wan Ju Kang, Taeyoung Lee, Kyunghwan Son, and Yung Yi. 2019 · 2019
Cited alongside, same era.
Multi-agent game abstraction via graph attention neural network. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 34. 7211–7218
Yong Liu, Weixun Wang, Yujing Hu, Jianye Hao, Xingguo Chen, and Yang Gao. 2020 · 2020
Later among the works it cites.
PRIMAL _ 2 \_2 : Pathfinding Via Reinforcement and Imitation Multi-Agent Learning-Lifelong
Mehul Damani, Zhiyao Luo, Emerson Wenzel, and Guillaume Sartoretti. 2021 · 2021
Later among the works it cites.
Deep reinforcement learning for autonomous driving: A survey
B Ravi Kiran, Ibrahim Sobh, Victor Talpaert, Patrick Mannion, Ahmad A Al Sallab, Senthil Yogamani, and Patrick Pérez. 2021 · 2021
Later among the works it cites.
Autonomous Bus Fleet Control Using Multiagent Reinforcement Learning
Sung-Jung Wang and SK Chang. 2021 · 2021
Later among the works it cites.
The surprising effectiveness of mappo in cooperative, multi-agent games
Chao Yu, Akash Velu, Eugene Vinitsky, Yu Wang, Alexandre Bayen, and Yi Wu. 2021 · 2021
Later among the works it cites.
Value-Decomposition Networks For Cooperative Multi-Agent Learning Based On Team Reward. In Proceedings of the 17th International Conference on Autonomous Agents and MultiAgent Systems . 2085–2087
Peter Sunehag, Guy Lever, Audrunas Gruslys, Wojciech Marian Czarnecki, Vinicius Zambaldi, Max Jaderberg, Marc Lanctot, Nicolas Sonnerat, Joel Z Leibo, Karl Tuyls, et al · 2087
Closest in time.