Fetching the paper…
Reading the bibliography…
Zero-shot coordination in cooperative artificial intelligence (AI) remains a significant challenge, which means effectively coordinating with a wide range of unseen partners.
Cores of convex games
Shapley, L. S · 1971
Earlier work this paper cites.
Centrality in social networks conceptual clarification
Freeman, L. C · 1978
Earlier work this paper cites.
Td-gammon, a self-teaching backgammon program, achieves master-level play
Tesauro, G · 1994
Earlier work this paper cites.
The theory of learning in games , volume 2
Fudenberg, D., Drew, F., Levine, D. K., and Levine, D. K · 1998
Earlier work this paper cites.
The pagerank citation ranking: Bringing order to the web
Page, L., Brin, S., Motwani, R., and Winograd, T · 1999
Earlier work this paper cites.
Analyzing complex strategic interactions in multi-agent systems
Walsh, W. E., Das, R., Tesauro, G., and Kephart, J. O · 2002
Earlier work this paper cites.
Weighted pagerank algorithm
Xing, W. and Ghorbani, A · 2004
Earlier work this paper cites.
Universal intelligence: A definition of machine intelligence
Legg, S. and Hutter, M · 2007
Earlier work this paper cites.
Introduction to the theory of cooperative games
Peleg, B. and Sudhölter, P · 2007
Earlier work this paper cites.
Polynomial calculation of the shapley value based on sampling
Castro, J., Gómez, D., and Tejada, J · 2009
Earlier work this paper cites.
Computational aspects of cooperative game theory
Chalkiadakis, G., Elkind, E., and Wooldridge, M · 2011
Earlier work this paper cites.
Continually adding self-invented problems to the repertoire: First experiments with powerplay
Srivastava, R., Steunebrink, B., Stollenga, M., and Schmidhuber, J · 2012
Earlier work this paper cites.
Population based training of neural networks
Jaderberg, M., Dalibard, V., Osindero, S., Czarnecki, W. M., Donahue, J., Razavi, A., Vinyals, O., Green, T., Dunning, I., Simonyan, K., et al · 2017
Cited alongside, same era.
A unified game-theoretic approach to multiagent reinforcement learning
Lanctot, M., Zambaldi, V., Gruslys, A., Lazaridou, A., Tuyls, K., Perolat, J., Silver, D., and Graepel, T · 2017
Cited alongside, same era.
Proximal policy optimization algorithms
Schulman, J., Wolski, F., Dhariwal, P., Radford, A., and Klimov, O · 2017
Cited alongside, same era.
Learning social conventions in markov games
Lerer, A. and Peysakhovich, A · 2018
Cited alongside, same era.
A generalised method for empirical game theoretic analysis
Tuyls, K., Perolat, J., Lanctot, M., Leibo, J. Z., and Graepel, T · 2018
Towards unifying behavioral and response diversity for open-ended learning in zero-sum games
Liu, X., Jia, H., Wen, Y., Hu, Y., Chen, Y., Fan, C., Hu, Z., and Yang, Y · 2021
Later among the works it cites.
Trajectory diversity for zero-shot coordination
Lupu, A., Cui, B., Hu, H., and Foerster, J · 2021
Later among the works it cites.
Collaborating with humans without human data
Strouse, D., McKee, K., Botvinick, M., Hughes, E., and Everett, R · 2021
Later among the works it cites.
Open-ended learning leads to generally capable agents
Team, O. E. L., Stooke, A., Mahajan, A., Barros, C., Deck, C., Bauer, J., Sygnowski, J., Trebacz, M., Jaderberg, M., Mathieu, M., et al · 2021
Later among the works it cites.
Diverse auto-curriculum is critical for successful real-world multiagent learning systems
Yang, Y., Luo, J., Wen, Y., Slumbers, O., Graves, D., Bou Ammar, H., Wang, J., and Taylor, M. E · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Open-ended learning in symmetric zero-sum games, 2019
Balduzzi, D., Garnelo, M., Bachrach, Y., Czarnecki, W. M., Perolat, J., Jaderberg, M., and Graepel, T · 2019
Cited alongside, same era.
On the utility of learning about humans for human-ai coordination
Carroll, M., Shah, R., Ho, M. K., Griffiths, T., Seshia, S., Abbeel, P., and Dragan, A · 2019
Cited alongside, same era.
The hanabi challenge: A new frontier for ai research
Bard, N., Foerster, J. N., Chandar, S., Burch, N., Lanctot, M., Song, H. F., Parisotto, E., Dumoulin, V., Moitra, S., Hughes, E., et al · 2020
Cited alongside, same era.
Investigating partner diversification methods in cooperative multi-agent deep reinforcement learning
Charakorn, R., Manoonpong, P., and Dilokthanakul, N · 2020
Cited alongside, same era.
”other-play” for zero-shot coordination
Hu, H., Lerer, A., Peysakhovich, A., and Foerster, J. N · 2020
Cited alongside, same era.
Pipeline psro: A scalable approach for finding approximate nash equilibria in large games
McAleer, S., Lanier, J. B., Fox, R., and Baldi, P · 2020
Cited alongside, same era.
Evaluating the robustness of collaborative agents
Knott, P., Carroll, M., Devlin, S., Ciosek, K., Hofmann, K., Dragan, A. D., and Shah, R · 2021
Cited alongside, same era.
Generating and adapting to diverse ad-hoc partners in hanabi
Canaan, R., Gao, X., Togelius, J., Nealen, A., and Menzel, S · 2022
Later among the works it cites.
Generalization in cooperative multi-agent systems
Mahajan, A., Samvelyan, M., Gupta, T., Ellis, B., Sun, M., Rocktäschel, T., and Whiteson, S · 2022
Later among the works it cites.
Self-play psro: Toward optimal populations in two-player zero-sum games
McAleer, S., Lanier, J., Wang, K., Baldi, P., Fox, R., and Sandholm, T · 2022
Later among the works it cites.
Open-ended reinforcement learning with neural reward functions
Meier, R. and Mujika, A · 2022
Later among the works it cites.
Heterogeneous multi-agent zero-shot coordination by coevolution
Xue, K., Wang, Y., Yuan, L., Guan, C., Qian, C., and Yu, Y · 2022
Later among the works it cites.
Maximum entropy population based training for zero-shot human-ai coordination
Zhao, R., Song, J., Haifeng, H., Gao, Y., Wu, Y., Sun, Z., and Wei, Y · 2023
Closest in time.