Fetching the paper…
Reading the bibliography…
Cooperative Multi-agent Reinforcement Learning (MARL) has attracted significant attention and played the potential for many real-world applications.
S. Wold, K. Esbensen, and P. Geladi, “Principal component analysis,” Chemometrics and Intelligent Laboratory Systems , vol. 2, no. 1-3, pp. 37–52, 1987
1987
Earlier work this paper cites.
F. Christianos, G. Papoudakis, M. A. Rahman, and S. V. Albrecht, “Scaling multi-agent reinforcement learning with selective parameter sharing,” in ICML , 2021, pp. 1989–1998
1998
Earlier work this paper cites.
H. Jeffreys, The Theory of Probability . OUP Oxford, 1998
1998
Earlier work this paper cites.
G. E. Hinton, “Training products of experts by minimizing contrastive divergence,” Neural computation , vol. 14, no. 8, pp. 1771–1800, 2002
2002
Earlier work this paper cites.
S. Chopra, R. Hadsell, and Y. LeCun, “Learning a similarity metric discriminatively, with application to face verification,” in CVPR , 2005, pp. 539–546
2005
Earlier work this paper cites.
L. Busoniu, R. Babuska, and B. De Schutter, “A comprehensive survey of multiagent reinforcement learnin,” IEEE Transactions on Systems, Man, and Cybernetics, Part C (Applications and Reviews) , vol. 38, no. 2, pp. 156–172, 2008
2008
Earlier work this paper cites.
2014
Earlier work this paper cites.
F. A. Oliehoek and C. Amato, A Concise Introduction to Decentralized POMDPs . Springer, 2016
2016
Earlier work this paper cites.
R. Lowe, Y. Wu, A. Tamar, J. H. P. Abbeel, and I. Mordatch, “Multi-agent actor-critic for mixed cooperative-competitive environments,” in NIPS , 2017, pp. 6379–6390
2017
Earlier work this paper cites.
J. Kirkpatrick, R. Pascanu, N. Rabinowitz, J. Veness, G. Desjardins, A. A. Rusu, K. Milan, J. Quan, T. Ramalho, A. Grabska-Barwinska et al. , “Overcoming catastrophic forgetting in neural networks,” Proceedings of the national academy of sciences , vol. 114, no. 13, pp. 3521–3526, 2017
2017
Earlier work this paper cites.
D. Lopez-Paz and M. Ranzato, “Gradient episodic memory for continual learning,” in NIPS , 2017, pp. 6467–6476
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin, “Attention is all you need,” in NIPS , 2017, pp. 5998–6008
2017
Earlier work this paper cites.
T. Rashid, M. Samvelyan, C. Schroeder, G. Farquhar, J. Foerster, and S. Whiteson, “Qmix: Monotonic value function factorisation for deep multi-agent reinforcement learning,” in ICML , 2018, pp. 4295–4304
2018
Earlier work this paper cites.
G. Sartoretti, J. Kerr, Y. Shi, G. Wagner, T. S. Kumar, S. Koenig, and H. Choset, “Primal: Pathfinding via reinforcement and imitation multi-agent learning,” IEEE Robotics and Automation Letters , vol. 4, no. 3, pp. 2378–2385, 2019
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
G. I. Parisi, R. Kemker, J. L. Part, C. Kanan, and S. Wermter, “Continual lifelong learning with neural networks: A review,” Neural Networks , vol. 113, pp. 54–71, 2019
2019
Earlier work this paper cites.
C. Kaplanis, M. Shanahan, and C. Clopath, “Policy consolidation for continual reinforcement learning,” in ICML , 2019, pp. 3242–3251
2019
Earlier work this paper cites.
M. Samvelyan, T. Rashid, C. S. de Witt, G. Farquhar, N. Nardelli, T. G. J. Rudner, C. Hung, P. H. S. Torr, J. N. Foerster, and S. Whiteson, “The Starcraft multi-agent challenge,” in AAMAS , 2019, pp. 2186–2188
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
D. Rolnick, A. Ahuja, J. Schwarz, T. P. Lillicrap, and G. Wayne, “Experience replay for continual learning,” in NeurIPS , 2019, pp. 348–358
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
K. Son, D. Kim, W. J. Kang, D. E. Hostallero, and Y. Yi, “QTRAN: Learning to factorize with transformation for cooperative multi-agent reinforcement learning,” in ICML , 2019, pp. 5887–5896
2019
Cited alongside, same era.
H. Hu, A. Lerer, A. Peysakhovich, and J. N. Foerster, “”other-play” for zero-shot coordination,” in ICML , 2020, pp. 4399–4410
2020
Cited alongside, same era.
L. Caccia, E. Belilovsky, M. Caccia, and J. Pineau, “Online learned continual compression with adaptive quantization modules,” in ICML , 2020, pp. 1240–1250
2020
Cited alongside, same era.
2020
Cited alongside, same era.
S. Lee, J. Ha, D. Zhang, and G. Kim, “A neural dirichlet process mixture model for task-free continual learning,” in ICLR , 2020
2022
Later among the works it cites.
2022
Later among the works it cites.
C. Yu, A. Velu, E. Vinitsky, J. Gao, Y. Wang, A. Bayen, and Y. Wu, “The surprising effectiveness of PPO in cooperative multi-agent games,” in NeurIPS , 2022
2022
Later among the works it cites.
R. Gorsane, O. Mahjoub, R. J. de Kock, R. Dubb, S. Singh, and A. Pretorius, “Towards a standardised performance evaluation protocol for cooperative MARL,” in NeurIPS , 2022
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2020
Cited alongside, same era.
J. Wang, W. Xu, Y. Gu, W. Song, and T. C. Green, “Multi-agent reinforcement learning for active voltage control on power distribution networks,” in NeurIPS , 2021, pp. 3271–3284
2021
Cited alongside, same era.
J. Wang, Z. Ren, B. Han, J. Ye, and C. Zhang, “Towards understanding cooperative multi-agent q-learning with value factorization,” in NeurIPS , 2021, pp. 29 142–29 155
2021
Cited alongside, same era.
E. Lecarpentier, D. Abel, K. Asadi, Y. Jinnai, E. Rachelson, and M. L. Littman, “Lipschitz lifelong reinforcement learning,” in AAAI , 2021, pp. 8270–8278
2021
Cited alongside, same era.
K. Zhang, Z. Yang, and T. Başar, “Multi-agent reinforcement learning: A selective overview of theories and algorithms,” Handbook of Reinforcement Learning and Control , pp. 321–384, 2021
2021
Cited alongside, same era.
H. Nekoei, A. Badrinaaraayanan, A. C. Courville, and S. Chandar, “Continuous coordination as a realistic scenario for lifelong learning,” in ICML , 2021, pp. 8016–8024
2021
Cited alongside, same era.
G. Papoudakis, F. Christianos, L. Schäfer, and S. V. Albrecht, “Benchmarking multi-agent deep reinforcement learning algorithms in cooperative tasks,” in NeurIPS , 2021
2021
Cited alongside, same era.
2021
Cited alongside, same era.
K. Khetarpal, M. Riemer, I. Rish, and D. Precup, “Towards continual reinforcement learning: A review and perspectives,” Journal of Artificial Intelligence Research , vol. 75, pp. 1401–1476, 2022
2022
Later among the works it cites.
S. Sodhani, F. Meier, J. Pineau, and A. Zhang, “Block contextual mdps for continual learning,” in L4DC , 2022, pp. 608–623
2022
Later among the works it cites.
S. Kessler, J. Parker-Holder, P. J. Ball, S. Zohren, and S. J. Roberts, “Same state, different task: Continual reinforcement learning without interference,” in AAAI , 2022, pp. 7143–7151
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
H. Wang, Y. Yu, and Y. Jiang, “Fully decentralized multiagent communication via causal inference,” IEEE Transactions on Neural Networks and Learning Systems , 2022
2022
Later among the works it cites.
M. Wen, J. G. Kuba, R. Lin, W. Zhang, Y. Wen, J. Wang, and Y. Yang, “Multi-agent reinforcement learning is a sequence modeling problem,” in NeurIPS , 2022
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
D. Kudithipudi, M. Aguilar-Simon, J. Babb, M. Bazhenov, D. Blackiston, J. Bongard, A. P. Brna, S. Chakravarthi Raja, N. Cheney, J. Clune et al. , “Biological underpinnings for lifelong learning machines,” Nature Machine Intelligence , vol. 4, no. 3, pp. 196–210, 2022
2022
Later among the works it cites.
2022
Later among the works it cites.
Z. Wang, C. Chen, and D. Dong, “Lifelong incremental reinforcement learning with online bayesian inference,” IEEE Transactions on Neural Networks and Learning Systems , vol. 33, no. 8, pp. 4003–4016, 2022
2022
Later among the works it cites.
S. Powers, E. Xing, E. Kolve, R. Mottaghi, and A. Gupta, “Cora: Benchmarks, baselines, and metrics as a platform for continual reinforcement learning agents,” in CoLLAs , 2022, pp. 705–743
2022
Later among the works it cites.
2022
Later among the works it cites.
F. Zhang, C. Jia, Y.-C. Li, L. Yuan, Y. Yu, and Z. Zhang, “Discovering generalizable multi-agent coordination skills from multi-task offline data,” in ICLR , 2023
2023
Closest in time.
P. Sunehag, G. Lever, A. Gruslys, W. M. Czarnecki, V. F. Zambaldi, M. Jaderberg, M. Lanctot, N. Sonnerat, J. Z. Leibo, K. Tuyls, and T. Graepel, “Value-decomposition networks for cooperative multi-agent learning based on team reward,” in AAMAS , 2018, pp. 2085–2087
2087
Closest in time.