Fetching the paper…
Reading the bibliography…
Value factorisation is a useful technique for multi-agent reinforcement learning (MARL) in global reward game, however its underlying mechanism is not yet fully understood.
Sur les opérations dans les ensembles abstraits et leur application aux équations intégrales
Stefan Banach · 1922
Earlier work this paper cites.
On the theory of dynamic programming
Richard Bellman · 1952
Earlier work this paper cites.
Cores of convex games
Lloyd S Shapley · 1971
Earlier work this paper cites.
Generalized network problems yielding totally balanced games
Ehud Kalai and Eitan Zemel · 1982
Earlier work this paper cites.
On the complexity of cooperative solution concepts
Xiaotie Deng and Christos H Papadimitriou · 1994
Earlier work this paper cites.
On the convergence of stochastic iterative dynamic programming algorithms
Tommi Jaakkola, Michael I Jordan, and Satinder P Singh · 1994
Earlier work this paper cites.
The dynamics of reinforcement learning in cooperative multiagent systems
Caroline Claus and Craig Boutilier · 1998
Earlier work this paper cites.
Algorithmic aspects of the core of combinatorial optimization games
Xiaotie Deng, Toshihide Ibaraki, and Hiroshi Nagamochi · 1999
Earlier work this paper cites.
Convergence of q-learning: A simple proof
Francisco S Melo · 2001
Earlier work this paper cites.
Optimal payoff functions for members of collectives
David H Wolpert and Kagan Tumer · 2002
Earlier work this paper cites.
Introduction to Banach algebras, operators, and harmonic analysis , volume 57
Harold Garth Dales, H Garth Dales, Pietro Aiena, Jörg Eschmeier, Kjeld Laursen, and George A Willis · 2003
Earlier work this paper cites.
Tree-based batch mode reinforcement learning
Damien Ernst, Pierre Geurts, and Louis Wehenkel · 2005
Earlier work this paper cites.
Towards understanding linear value decomposition in cooperative multi-agent q-learning
Jianhao Wang, Zhizhou Ren, Beining Han, Jianing Ye, and Chongjie Zhang · 2006
Earlier work this paper cites.
Decentralized receding horizon control and coordination of autonomous vehicle formations
Tamás Keviczky, Francesco Borrelli, Kingsley Fregene, Datta Godbole, and Gary J Balas · 2007
Earlier work this paper cites.
Optimal and approximate q-value functions for decentralized pomdps
Frans A Oliehoek, Matthijs TJ Spaan, and Nikos Vlassis · 2008
Earlier work this paper cites.
Qplex: Duplex dueling multi-agent q-learning
Jianhao Wang, Zhizhou Ren, Terry Liu, Yang Yu, and Chongjie Zhang · 2008
Cited alongside, same era.
Computational aspects of cooperative game theory
Georgios Chalkiadakis, Edith Elkind, and Michael Wooldridge · 2011
Cited alongside, same era.
Multiagent coordination enabling autonomous logistics
Arne Schuldt · 2012
Cited alongside, same era.
Decentralized pomdps
Frans A Oliehoek · 2012
Cited alongside, same era.
Sample size selection in optimization methods for machine learning
Richard H. Byrd, Gillian M. Chin, Jorge Nocedal, and Yuchen Wu · 2012
Cited alongside, same era.
Counterfactual multi-agent policy gradients
Jakob N Foerster, Gregory Farquhar, Triantafyllos Afouras, Nantas Nardelli, and Shimon Whiteson · 2018
Later among the works it cites.
Value-decomposition networks for cooperative multi-agent learning based on team reward
Peter Sunehag, Guy Lever, Audrunas Gruslys, Wojciech Marian Czarnecki, Vinícius Flores Zambaldi, Max Jaderberg, Marc Lanctot, Nicolas Sonnerat, Joel Z. Leibo, Karl Tuyls, and Thore Graepel · 2018
Later among the works it cites.
QMIX: monotonic value function factorisation for deep multi-agent reinforcement learning
Tabish Rashid, Mikayel Samvelyan, Christian Schröder de Witt, Gregory Farquhar, Jakob N. Foerster, and Shimon Whiteson · 2018
Later among the works it cites.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto · 2018
Later among the works it cites.
Multiagent soft q-learning
Ermo Wei, Drew Wicke, David Freelan, and Sean Luke · 2018
Later among the works it cites.
Learning to schedule communication in multi-agent reinforcement learning
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Alex Graves, Ioannis Antonoglou, Daan Wierstra, and Martin Riedmiller · 2013
Cited alongside, same era.
Empirical evaluation of gated recurrent neural networks on sequence modeling
Junyoung Chung, Caglar Gulcehre, KyungHyun Cho, and Yoshua Bengio · 2014
Cited alongside, same era.
Secure estimation and control for cyber-physical systems under adversarial attacks
Hamza Fawzi, Paulo Tabuada, and Suhas Diggavi · 2014
Cited alongside, same era.
Generative adversarial networks
Ian J Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Cited alongside, same era.
Variance reduced stochastic gradient descent with neighbors
Thomas Hofmann, Aurelien Lucchi, Simon Lacoste-Julien, and Brian McWilliams · 2015
Cited alongside, same era.
Learning multiagent communication with backpropagation
Sainbayar Sukhbaatar, arthur szlam, and Rob Fergus · 2016
Cited alongside, same era.
Lenient learning in independent-learner stochastic cooperative games
Ermo Wei and Sean Luke · 2016
Cited alongside, same era.
Daewoo Kim, Sangwoo Moon, David Hostallero, Wan Ju Kang, Taeyoung Lee, Kyunghwan Son, and Yung Yi · 2019
Later among the works it cites.
QTRAN: learning to factorize with transformation for cooperative multi-agent reinforcement learning
Kyunghwan Son, Daewoo Kim, Wan Ju Kang, David Hostallero, and Yung Yi · 2019
Later among the works it cites.
The starcraft multi-agent challenge
Mikayel Samvelyan, Tabish Rashid, Christian Schroeder de Witt, Gregory Farquhar, Nantas Nardelli, Tim GJ Rudner, Chia-Man Hung, Philip HS Torr, Jakob Foerster, and Shimon Whiteson · 2019
Later among the works it cites.
Reinforcement learning and optimal control
Dimitri Bertsekas · 2019
Later among the works it cites.
Actor-attention-critic for multi-agent reinforcement learning
Shariq Iqbal and Fei Sha · 2019
Later among the works it cites.
Learning implicit credit assignment for multi-agent actor-critic
Meng Zhou, Ziyu Liu, Pengwei Sui, Yixuan Li, and Yuk Ying Chung · 2020
Later among the works it cites.
Deep coordination graphs
Wendelin Böhmer, Vitaly Kurin, and Shimon Whiteson · 2020
Later among the works it cites.
Weighted qmix: Expanding monotonic value function factorisation for deep multi-agent reinforcement learning
Tabish Rashid, Gregory Farquhar, Bei Peng, and Shimon Whiteson · 2020
Later among the works it cites.
Multi-agent reinforcement learning for active voltage control on power distribution networks
Jianhong Wang, Wangkun Xu, Yunjie Gu, Wenbin Song, and Tim Green · 2021
Closest in time.
Benchmarking multi-agent deep reinforcement learning algorithms in cooperative tasks
Georgios Papoudakis, Filippos Christianos, Lukas Schäfer, and Stefano V. Albrecht · 2021
Closest in time.