Fetching the paper…
Reading the bibliography…
Learning sparse coordination graphs adaptive to the coordination dynamics among agents is a long-standing problem in cooperative multi-agent learning.
Über algebraische gleichungen mit lauter reellen wurzeln
Julius Nagy · 1918
Earlier work this paper cites.
On optimal cooperation of knowledge sources: an empirical investigation
Miroslav Benda · 1986
Earlier work this paper cites.
Multi-agent reinforcement learning: Independent vs. cooperative agents
Ming Tan · 1993
Earlier work this paper cites.
The dynamics of reinforcement learning in cooperative multiagent systems
Caroline Claus and Craig Boutilier · 1998
Earlier work this paper cites.
Multiagent systems: A survey from a machine learning perspective
Peter Stone and Manuela Veloso · 2000
Earlier work this paper cites.
Dynamic programming for partially observable stochastic games
Eric A Hansen, Daniel S Bernstein, and Shlomo Zilberstein · 2004
Earlier work this paper cites.
Collaborative multiagent reinforcement learning by payoff propagation
Jelle R Kok and Nikos Vlassis · 2006
Earlier work this paper cites.
Biasing coevolutionary search for optimal multiagent behaviors
Liviu Panait, Sean Luke, and R Paul Wiegand · 2006
Earlier work this paper cites.
Decentralised coordination of mobile sensors using the max-sum algorithm
Ruben Stranders, Alessandro Farinelli, Alex Rogers, and Nick Jennings · 2009
Earlier work this paper cites.
Value-based planning for teams of agents in stochastic partially observable environments
Frans Oliehoek · 2010
Earlier work this paper cites.
Coordinated multi-agent reinforcement learning in networked distributed pomdps
Chongjie Zhang and Victor Lesser · 2011
Earlier work this paper cites.
Coordinating decentralized learning and conflict resolution across agent boundaries
Shanjun Cheng · 2012
Earlier work this paper cites.
Distributed sensor networks: A multiagent perspective , volume 9
Victor Lesser, Charles L Ortiz Jr, and Milind Tambe · 2012
Cited alongside, same era.
Coordinating multi-agent reinforcement learning with limited communication
Chongjie Zhang and Victor Lesser · 2013
Cited alongside, same era.
Learning phrase representations using rnn encoder–decoder for statistical machine translation
Kyunghyun Cho, Bart van Merriënboer, Caglar Gulcehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio · 2014
Cited alongside, same era.
Probabilistic reasoning in intelligent systems: networks of plausible inference
Judea Pearl · 2014
Cited alongside, same era.
A concise introduction to decentralized POMDPs , volume 1
Frans A Oliehoek, Christopher Amato, et al · 2016
Cited alongside, same era.
Lenient learning in independent-learner stochastic cooperative games
Qmix: Monotonic value function factorisation for deep multi-agent reinforcement learning
Tabish Rashid, Mikayel Samvelyan, Christian Schroeder Witt, Gregory Farquhar, Jakob Foerster, and Shimon Whiteson · 2018
Later among the works it cites.
Value-decomposition networks for cooperative multi-agent learning based on team reward
Peter Sunehag, Guy Lever, Audrunas Gruslys, Wojciech Marian Czarnecki, Vinicius Zambaldi, Max Jaderberg, Marc Lanctot, Nicolas Sonnerat, Joel Z Leibo, Karl Tuyls, et al · 2018
Later among the works it cites.
Multi-vehicle flocking control with deep deterministic policy gradient method
Zhao Xu, Yang Lyu, Quan Pan, Jinwen Hu, Chunhui Zhao, and Shuai Liu · 2018
Later among the works it cites.
The representational capacity of action-value networks for multi-agent reinforcement learning
Jacopo Castellini, Frans A Oliehoek, Rahul Savani, and Shimon Whiteson · 2019
Later among the works it cites.
The starcraft multi-agent challenge
Mikayel Samvelyan, Tabish Rashid, Christian Schroeder de Witt, Gregory Farquhar, Nantas Nardelli, Tim GJ Rudner, Chia-Man Hung, Philip HS Torr, Jakob Foerster, and Shimon Whiteson · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ermo Wei and Sean Luke · 2016
Cited alongside, same era.
Multi-agent actor-critic for mixed cooperative-competitive environments
Ryan Lowe, Yi Wu, Aviv Tamar, Jean Harb, OpenAI Pieter Abbeel, and Igor Mordatch · 2017
Cited alongside, same era.
Probability and computing: Randomization and probabilistic techniques in algorithms and data analysis
Michael Mitzenmacher and Eli Upfal · 2017
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Cited alongside, same era.
Counterfactual multi-agent policy gradients
Jakob N Foerster, Gregory Farquhar, Triantafyllos Afouras, Nantas Nardelli, and Shimon Whiteson · 2018
Cited alongside, same era.
Cooperative and distributed reinforcement learning of drones for field coverage, 2018
Huy Xuan Pham, Hung Manh La, David Feil-Seifer, and Aria Nefian · 2018
Cited alongside, same era.
Multiagent planning with factored mdps
Carlos Guestrin, Daphne Koller, and Ronald Parr
Cited in the paper.
Later among the works it cites.
Qtran: Learning to factorize with transformation for cooperative multi-agent reinforcement learning
Kyunghwan Son, Daewoo Kim, Wan Ju Kang, David Earl Hostallero, and Yung Yi · 2019
Later among the works it cites.
Deep coordination graphs
Wendelin Böhmer, Vitaly Kurin, and Shimon Whiteson · 2020
Later among the works it cites.
Graph convolutional value decomposition in multi-agent reinforcement learning
Navid Naderializadeh, Fan H Hung, Sean Soleyman, and Deepak Khosla · 2020
Later among the works it cites.
Weighted qmix: Expanding monotonic value function factorisation for deep multi-agent reinforcement learning
Tabish Rashid, Gregory Farquhar, Bei Peng, and Shimon Whiteson · 2020
Later among the works it cites.
Learning nearly decomposable value functions with communication minimization
Tonghan Wang, Jianhao Wang, Chongyi Zheng, and Chongjie Zhang · 2020
Later among the works it cites.
Deep implicit coordination graphs for multi-agent reinforcement learning
Sheng Li, Jayesh K Gupta, Peter Morales, Ross Allen, and Mykel J Kochenderfer · 2021
Closest in time.