Fetching the paper…
Reading the bibliography…
Centralized Training with Decentralized Execution (CTDE) has emerged as a widely adopted paradigm in multi-agent reinforcement learning, emphasizing the utilization of global information for learning an enhanced joint $Q$-function or centralized critic.
Cumulated gain-based evaluation of ir techniques
Kalervo Järvelin and Jaana Kekäläinen · 2002
Earlier work this paper cites.
Learning to rank for information retrieval
Tie-Yan Liu et al · 2009
Earlier work this paper cites.
Introducing letor 4.0 datasets
Tao Qin and Tie-Yan Liu · 2013
Earlier work this paper cites.
Distilling the knowledge in a neural network
Geoffrey Hinton, Oriol Vinyals, Jeff Dean, et al · 2015
Earlier work this paper cites.
Andrei A Rusu, Sergio Gomez Colmenarejo, Caglar Gulcehre, Guillaume Desjardins, James Kirkpatrick, Razvan Pascanu, Volodymyr Mnih, Koray Kavukcuoglu, and Raia Hadsell · 2015
Earlier work this paper cites.
A concise introduction to decentralized POMDPs
Frans A Oliehoek and Christopher Amato · 2016
Earlier work this paper cites.
Multi-agent actor-critic for mixed cooperative-competitive environments
Ryan Lowe, Yi I Wu, Aviv Tamar, Jean Harb, OpenAI Pieter Abbeel, and Igor Mordatch · 2017
Earlier work this paper cites.
Value-decomposition networks for cooperative multi-agent learning
Peter Sunehag, Guy Lever, Audrunas Gruslys, Wojciech Marian Czarnecki, Vinicius Zambaldi, Max Jaderberg, Marc Lanctot, Nicolas Sonnerat, Joel Z Leibo, Karl Tuyls, et al · 2017
Earlier work this paper cites.
Counterfactual multi-agent policy gradients
Jakob Foerster, Gregory Farquhar, Triantafyllos Afouras, Nantas Nardelli, and Shimon Whiteson · 2018
Earlier work this paper cites.
Towards optimally decentralized multi-robot collision avoidance via deep reinforcement learning
Pinxin Long, Tingxiang Fan, Xinyi Liao, Wenxi Liu, Hao Zhang, and Jia Pan · 2018
Earlier work this paper cites.
Qmix: Monotonic value function factorisation for deep multi-agent reinforcement learning
Tabish Rashid, Mikayel Samvelyan, Christian Schroeder, Gregory Farquhar, Jakob Foerster, and Shimon Whiteson · 2018
Earlier work this paper cites.
Actor-attention-critic for multi-agent reinforcement learning
Shariq Iqbal and Fei Sha · 2019
Cited alongside, same era.
Joint optimization of multi-uav target assignment and path planning based on multi-agent reinforcement learning
Han Qie, Dianxi Shi, Tianlong Shen, Xinhai Xu, Yuan Li, and Liujing Wang · 2019
Cited alongside, same era.
The StarCraft Multi-Agent Challenge
Mikayel Samvelyan, Tabish Rashid, Christian Schroeder de Witt, Gregory Farquhar, Nantas Nardelli, Tim G. J. Rudner, Chia-Man Hung, Philiph H. S. Torr, Jakob Foerster, and Shimon Whiteson · 2019
Cited alongside, same era.
Cooperative multi-robot navigation in dynamic environment with deep reinforcement learning
Ruihua Han, Shengduo Chen, and Qi Hao · 2020
Cited alongside, same era.
Google research football: A novel reinforcement learning environment
Karol Kurach, Anton Raichuk, et al · 2020
Cited alongside, same era.
Coach-player multi-agent reinforcement learning for dynamic team composition
Bo Liu, Qiang Liu, Peter Stone, Animesh Garg, Yuke Zhu, and Anima Anandkumar · 2021
Later among the works it cites.
Seihai: A sample-efficient hierarchical ai for the minerl competition
Hangyu Mao, Chao Wang, Xiaotian Hao, Yihuan Mao, Yiming Lu, Chengjie Wu, Jianye Hao, Dong Li, and Pingzhong Tang · 2021
Later among the works it cites.
The surprising effectiveness of ppo in cooperative, multi-agent games
Chao Yu, Akash Velu, Eugene Vinitsky, Yu Wang, Alexandre Bayen, and Yi Wu · 2021
Later among the works it cites.
Commander-soldiers reinforcement learning for cooperative multi-agent systems
Yiqun Chen, Wei Yang, Tianle Zhang, Shiguang Wu, and Hongxing Chang · 2022
Closest in time.
Efficient policy generation in multi-agent systems via hypergraph neural network
Bin Zhang, Yunpeng Bai, Zhiwei Xu, Dapeng Li, and Guoliang Fan · 2022
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Learning agent communication under limited bandwidth by message pruning
Hangyu Mao, Zhengchao Zhang, Zhen Xiao, Zhibo Gong, and Yan Ni · 2020
Cited alongside, same era.
Qplex: Duplex dueling multi-agent q-learning
Jianhao Wang, Zhizhou Ren, Terry Liu, Yang Yu, and Chongjie Zhang · 2020
Cited alongside, same era.
Unmas: Multiagent reinforcement learning for unshaped cooperative scenarios
Jiajun Chai, Weifan Li, Yuanheng Zhu, Dongbin Zhao, Zhe Ma, Kewu Sun, and Jishiyu Ding · 2021
Cited alongside, same era.
Knowledge distillation: A survey
Jianping Gou, Baosheng Yu, Stephen J Maybank, and Dacheng Tao · 2021
Cited alongside, same era.
Towards robust and domain agnostic reinforcement learning competitions: Minerl 2020
William Hebgen Guss, Stephanie Milani, Nicholay Topin, Brandon Houghton, Sharada Mohanty, Andrew Melnik, Augustin Harter, Benoit Buschmaas, Bjarne Jaster, Christoph Berganski, Hangyu Mao, et al · 2021
Cited alongside, same era.
Rethinking the implementation tricks and monotonicity constraint in cooperative multi-agent reinforcement learning
Jian Hu, Siyang Jiang, Seth Austin Harding, Haibin Wu, and Shih-wei Liao · 2021
Cited alongside, same era.
Jian Zhao, Xunhan Hu, Mingyu Yang, Wengang Zhou, Jiangcheng Zhu, and Houqiang Li · 2022
Closest in time.
Inducing stackelberg equilibrium through spatio-temporal sequential decision-making in multi-agent reinforcement learning, 2023
Bin Zhang, Lijuan Li, Zhiwei Xu, Dapeng Li, and Guoliang Fan · 2023
Closest in time.
Stackelberg decision transformer for asynchronous action coordination in multi-agent systems
Bin Zhang, Hangyu Mao, Lijuan Li, Zhiwei Xu, Dapeng Li, Rui Zhao, and Guoliang Fan · 2023
Closest in time.
Ma4div: Multi-agent reinforcement learning for search result diversification
Yiqun Chen, Jiaxin Mao, Yi Zhang, Dehong Ma, Long Xia, Jun Fan, Daiting Shi, Zhicong Cheng, and Dawei Yin · 2024
Closest in time.
Measuring policy distance for multi-agent reinforcement learning
Tianyi Hu, Zhiqiang Pu, Xiaolin Ai, Tenghai Qiu, and Jianqiang Yi · 2024
Closest in time.