Learning social conventions in markov games
Original
Lerer, A. and Peysakhovich, A · 2018
Later among the works it cites.
Prosocial learning agents solve generalized stag hunts better than selfish ones
Peysakhovich, A. and Lerer, A · 2018
Later among the works it cites.
Vehicle community strategies
Original
Resnick, C., Kulikov, I., Cho, K., and Weston, J · 2018
Later among the works it cites.
Value-decomposition networks for cooperative multi-agent learning based on team reward
Sunehag, P., Lever, G., Gruslys, A., Czarnecki, W. M., Zambaldi, V., Jaderberg, M., Lanctot, M., Sonnerat, N., Leibo, J. Z., Tuyls, K., et al · 2018
Later among the works it cites.
Shaping representations through communication
Tieleman, O., Lazaridou, A., Mourad, S., Blundell, C., and Precup, D · 2018
Later among the works it cites.
On the utility of learning about humans for human-ai coordination
Carroll, M., Shah, R., Ho, M. K., Griffiths, T., Seshia, S., Abbeel, P., and Dragan, A · 2019
Later among the works it cites.
To populate is to regulate
Original
Fitzgerald, N · 2019
Later among the works it cites.
Predicting and understanding initial play
Fudenberg, D. and Liang, A · 2019
Later among the works it cites.
A survey and critique of multiagent deep reinforcement learning
Hernandez-Leal, P., Kartal, B., and Taylor, M. E · 2019
Later among the works it cites.
Simplified action decoder for deep multi-agent reinforcement learning
Original
Hu, H. and Foerster, J. N · 2019
Later among the works it cites.
Improving policies via search in cooperative partially observable games
Original
Lerer, A., Hu, H., Foerster, J., and Brown, N · 2019
Later among the works it cites.
Learning to learn to communicate, 2019
Lowe, R., Gupta, A., Foerster, J., Kiela, D., and Pineau, J · 2019
Later among the works it cites.
Theory of minds: Understanding behavior in groups through inverse planning
Shum, M., Kleiman-Weiner, M., Littman, M. L., and Tenenbaum, J. B · 2019
Later among the works it cites.
The hanabi challenge: A new frontier for ai research
Bard, N., Foerster, J. N., Chandar, S., Burch, N., Lanctot, M., Song, H. F., Parisotto, E., Dumoulin, V., Moitra, S., Hughes, E., et al · 2020
Closest in time.
Adversarially guided self-play for adopting social conventions
Original
Tucker, M., Zhou, Y., and Shah, J · 2020
Closest in time.
Plannable approximations to mdp homomorphisms: Equivariance under actions
Original
van der Pol, E., Kipf, T., Oliehoek, F. A., and Welling, M · 2020
Closest in time.