Learning attentional communication for multi-agent cooperation
Jiang, J. and Lu, Z · 2018
Later among the works it cites.
Neural relational inference for interacting systems
Kipf, T., Fetaya, E., Wang, K.-C., Welling, M., and Zemel, R · 2018
Later among the works it cites.
Role-based modeling for designing agent behavior in self-organizing multi-agent systems
Lhaksmana, K. M., Murakami, Y., and Ishida, T · 2018
Later among the works it cites.
Emergence of grounded compositional language in multi-agent populations
Mordatch, I. and Abbeel, P · 2018
Later among the works it cites.
Credit assignment for collective multiagent rl with global rewards
Nguyen, D. T., Kumar, A., and Lau, H. C · 2018
Later among the works it cites.
Qmix: Monotonic value function factorisation for deep multi-agent reinforcement learning
Rashid, T., Samvelyan, M., Witt, C. S., Farquhar, G., Foerster, J., and Whiteson, S · 2018
Later among the works it cites.
Value-decomposition networks for cooperative multi-agent learning based on team reward
Sunehag, P., Lever, G., Gruslys, A., Czarnecki, W. M., Zambaldi, V., Jaderberg, M., Lanctot, M., Sonnerat, N., Leibo, J. Z., Tuyls, K., et al · 2018
Later among the works it cites.
Glomo: unsupervised learning of transferable relational graphs
Yang, Z., Zhao, J., Dhingra, B., He, K., Cohen, W. W., Salakhutdinov, R. R., and LeCun, Y · 2018
Later among the works it cites.
Tarmac: Targeted multi-agent communication
Das, A., Gervet, T., Romoff, J., Batra, D., Parikh, D., Rabbat, M., and Pineau, J · 2019
Later among the works it cites.
Actor-attention-critic for multi-agent reinforcement learning
Iqbal, S. and Sha, F · 2019
Later among the works it cites.
Learning fairness in multi-agent systems
Jiang, J. and Lu, Z · 2019
Later among the works it cites.
Learning to schedule communication in multi-agent reinforcement learning
Kim, D., Moon, S., Hostallero, D., Kang, W. J., Lee, T., Son, K., and Yi, Y · 2019
Later among the works it cites.
Maven: Multi-agent variational exploration
Mahajan, A., Rashid, T., Samvelyan, M., and Whiteson, S · 2019
Later among the works it cites.
The starcraft multi-agent challenge
Original
Samvelyan, M., Rashid, T., de Witt, C. S., Farquhar, G., Nardelli, N., Rudner, T. G., Hung, C.-M., Torr, P. H., Foerster, J., and Whiteson, S · 2019
Later among the works it cites.
Learning when to communicate at scale in multiagent cooperative and competitive tasks
Singh, A., Jain, T., and Sukhbaatar, S · 2019
Later among the works it cites.
Qtran: Learning to factorize with transformation for cooperative multi-agent reinforcement learning
Son, K., Kim, D., Kang, W. J., Hostallero, D. E., and Yi, Y · 2019
Later among the works it cites.
Grandmaster level in starcraft ii using multi-agent reinforcement learning
Vinyals, O., Babuschkin, I., Czarnecki, W. M., Mathieu, M., Dudzik, A., Chung, J., Choi, D. H., Powell, R., Ewalds, T., Georgiev, P., et al · 2019
Later among the works it cites.
Probabilistic recursive reasoning for multi-agent reinforcement learning
Wen, Y., Yang, Y., Luo, R., Wang, J., and Pan, W · 2019
Later among the works it cites.
Emergent tool use from multi-agent autocurricula
Baker, B., Kanitscheider, I., Markov, T., Wu, Y., Powell, G., McGrew, B., and Mordatch, I · 2020
Closest in time.
Incorporating pragmatic reasoning communication into emergent language
Original
Kang, Y., Wang, T., and de Melo, G · 2020
Closest in time.