Fetching the paper…
Reading the bibliography…
In fully cooperative multi-agent reinforcement learning (MARL) settings, the environments are highly stochastic due to the partial observability of each agent and the continuously changing policies of the other agents.
Multi-agent reinforcement learning: Independent versus cooperative agents
Tan, M · 1993
Earlier work this paper cites.
Multiagent planning with factored mdps
Guestrin, C., Koller, D., and Parr, R · 2001
Earlier work this paper cites.
Estimation of quantile mixtures via l-moments and trimmed l-moments
Karvanen, J · 2005
Earlier work this paper cites.
Hysteretic q-learning : an algorithm for decentralized reinforcement learning in cooperative multi-agent teams
Matignon, L., Laurent, G., and Fort-Piat, N · 2007
Earlier work this paper cites.
Universal value function approximators
Schaul, T., Horgan, D., Gregor, K., and Silver, D · 2015
Earlier work this paper cites.
A Concise Introduction to Decentralized POMDPs
Oliehoek, F. A. and Amato, C · 2016
Earlier work this paper cites.
A concise introduction to decentralized POMDPs , volume 1
Oliehoek, F. A., Amato, C., et al · 2016
Earlier work this paper cites.
A distributional perspective on reinforcement learning
Bellemare, M. G., Dabney, W., and Munos, R · 2017
Earlier work this paper cites.
QMIX: Monotonic value function factorisation for deep multi-agent reinforcement learning
Rashid, T. et al · 2018
Earlier work this paper cites.
Value-decomposition networks for cooperative multi-agent learning based on team reward
Sunehag, P. et al · 2018
Cited alongside, same era.
Distributional reinforcement learning with linear function approximation
Bellemare, M. G., Roux, N. L., Castro, P. S., and Moitra, S · 2019
Cited alongside, same era.
Multi-agent common knowledge reinforcement learning
de Witt, C. S., Foerster, J., Farquhar, G., Torr, P., Böhmer, W., and Whiteson, S · 2019
Cited alongside, same era.
Liir: Learning individual intrinsic reward in multi-agent reinforcement learning
Du, Y., Han, L., Fang, M., Liu, J., Dai, T., and Tao, D · 2019
Cited alongside, same era.
Distributional reward decomposition for reinforcement learning
Lin, Z., Zhao, L., Yang, D., Qin, T., Liu, T.-Y., and Yang, G · 2019
Cited alongside, same era.
The starcraft multi-agent challenge
Samvelyan, M. et al · 2019
Later among the works it cites.
QTRAN: Learning to factorize with transformation for cooperative multi-agent reinforcement learning
Son, K., Kim, D., Kang, W. J., Hostallero, D. E., and Yi, Y · 2019
Later among the works it cites.
Learning nearly decomposable value functions via communication minimization
Wang, T., Wang, J., Zheng, C., and Zhang, C · 2019
Later among the works it cites.
Fully parameterized quantile function for distributional reinforcement learning
Yang, D. et al · 2019
Later among the works it cites.
Quota: The quantile option architecture for reinforcement learning
Zhang, S. and Yao, H · 2019
Later among the works it cites.
Efficient communication in multi-agent reinforcement learning via variance based control
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Lyle, C., Bellemare, M. G., and Castro, P. S · 2019
Cited alongside, same era.
Distributional reinforcement learning for efficient exploration
Mavrin, B., Yao, H., Kong, L., Wu, K., and Yu, Y · 2019
Cited alongside, same era.
Information-directed exploration for deep reinforcement learning
Nikolov, N., Kirschner, J., Berkenkamp, F., and Krause, A · 2019
Cited alongside, same era.
Statistics and samples in distributional reinforcement learning
Rowland, M. et al · 2019
Cited alongside, same era.
Implicit quantile networks for distributional reinforcement learning
Dabney, W., Ostrovski, G., Silver, D., and Munos, R
Cited in the paper.
Distributional reinforcement learning with quantile regression
Dabney, W., Rowland, M., Bellemare, M. G., and Munos, R
Cited in the paper.
Zhang, S. Q., Zhang, Q., and Lin, J · 2019
Later among the works it cites.
Likelihood quantile networks for coordinating multi-agent reinforcement learning
Lyu, X. and Amato, C · 2020
Later among the works it cites.
Multi-agent reinforcement learning with emergent roles
Wang, T., Dong, H., Lesser, V., and Zhang, C · 2020
Later among the works it cites.