Fetching the paper…
Reading the bibliography…
Promoting behavioural diversity is critical for solving games with non-transitive dynamics where strategic cycles exist, and there is no consistent winner (e.g., Rock-Paper-Scissors).
Equilibrium points in n-person games
Nash, J. F. et al · 1950
Earlier work this paper cites.
Iterative solution of games by fictitious play
Brown, G. W · 1951
Earlier work this paper cites.
The fermion process—a model of stochastic point process with repulsive points
Macchi, O · 1977
Earlier work this paper cites.
Coevolution: Genes, culture, and human diversity
Durham, W. H · 1991
Earlier work this paper cites.
Adaptation in natural and artificial systems: an introductory analysis with applications to biology, control, and artificial intelligence
Holland, J. H. et al · 1992
Earlier work this paper cites.
Consistency and cautious fictitious play
Fudenberg, D. and Levine, D · 1995
Earlier work this paper cites.
Coevolutionary computation
Paredis, J · 1995
Earlier work this paper cites.
The theory of learning in games , volume 2
Fudenberg, D., Drew, F., Levine, D. K., and Levine, D. K · 1998
Earlier work this paper cites.
Planning in the presence of cost functions controlled by an adversary
McMahan, H. B., Gordon, G. J., and Blum, A · 2003
Earlier work this paper cites.
Open-ended artificial evolution
Standish, R. K · 2003
Earlier work this paper cites.
Diversity creation methods: a survey and categorisation
Brown, G., Wyatt, J., Harris, R., and Yao, X · 2005
Earlier work this paper cites.
Evolutionary computation: toward a new philosophy of machine intelligence , volume 1
Fogel, D. B · 2006
Earlier work this paper cites.
Generalised weakened fictitious play
Leslie, D. S. and Collins, E. J · 2006
Earlier work this paper cites.
The colonel blotto game
Roberson, B · 2006
Earlier work this paper cites.
Methods for empirical game-theoretic analysis
Wellman, M. P · 2006
Earlier work this paper cites.
Exploiting open-endedness to solve problems through the search for novelty
Lehman, J. and Stanley, K. O · 2008
Earlier work this paper cites.
Settling the complexity of computing two-player nash equilibria
Chen, X., Deng, X., and Teng, S.-H · 2009
Earlier work this paper cites.
The complexity of computing a nash equilibrium
Daskalakis, C., Goldberg, P. W., and Papadimitriou, C. H · 2009
Earlier work this paper cites.
Flows and decompositions of games: Harmonic and potential games
Candogan, O., Menache, I., Ozdaglar, A., and Parrilo, P. A · 2011
Cited alongside, same era.
Determinantal point processes for machine learning
Kulesza, A., Taskar, B., et al · 2012
Cited alongside, same era.
Intrinsic motivation and reinforcement learning
Barto, A. G · 2013
Cited alongside, same era.
Learning the parameters of determinantal point process kernels
Affandi, R. H., Fox, E., Adams, R., and Taskar, B · 2014
Cited alongside, same era.
Using response functions to measure strategy strength
Davis, T., Burch, N., and Bowling, M · 2014
Cited alongside, same era.
Robots that can adapt like animals
Cully, A., Clune, J., Tarapore, D., and Mouret, J.-B · 2015
Cited alongside, same era.
Emergent coordination through competition
Liu, S., Lever, G., Merel, J., Tunyasuvunakool, S., Heess, N., and Graepel, T · 2018
Later among the works it cites.
A generalised method for empirical game theoretic analysis
Tuyls, K., Perolat, J., Lanctot, M., Leibo, J. Z., and Graepel, T · 2018
Later among the works it cites.
Emergent tool use from multi-agent autocurricula
Baker, B., Kanitscheider, I., Markov, T., Wu, Y., Powell, G., McGrew, B., and Mordatch, I · 2019
Later among the works it cites.
Open-ended learning in symmetric zero-sum games
Balduzzi, D., Garnelo, M., Bachrach, Y., Czarnecki, W., Pérolat, J., Jaderberg, M., and Graepel, T · 2019
Later among the works it cites.
Human-level performance in 3d multiplayer games with population-based reinforcement learning
Jaderberg, M., Czarnecki, W. M., Dunning, I., Marris, L., Lever, G., Castaneda, A. G., Beattie, C., Rabinowitz, N. C., Morcos, A. S., Ruderman, A., et al · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Illuminating search spaces by mapping elites
Mouret, J.-B. and Clune, J · 2015
Cited alongside, same era.
Defining and simulating open-ended novelty: requirements, guidelines, and challenges
Banzhaf, W., Baumgaertner, B., Beslon, G., Doursat, R., Foster, J. A., McMullin, B., De Melo, V. V., Miconi, T., Spector, L., Stepney, S., et al · 2016
Cited alongside, same era.
Unifying count-based exploration and intrinsic motivation
Bellemare, M., Srinivasan, S., Ostrovski, G., Schaul, T., Saxton, D., and Munos, R · 2016
Cited alongside, same era.
Quality diversity: A new frontier for evolutionary computation
Pugh, J. K., Soros, L. B., and Stanley, K. O · 2016
Cited alongside, same era.
Variational intrinsic control
Gregor, K., Rezende, D. J., and Wierstra, D · 2017
Cited alongside, same era.
Reinforcement learning with deep energy-based policies
Haarnoja, T., Tang, H., Abbeel, P., and Levine, S · 2017
Cited alongside, same era.
Leibo, J. Z., Hughes, E., Lanctot, M., and Graepel, T · 2019
Later among the works it cites.
A generalized training approach for multiagent learning
Muller, P., Omidshafiei, S., Rowland, M., Tuyls, K., Perolat, J., Liu, S., Hennes, D., Marris, L., Lanctot, M., Hughes, E., et al · 2019
Later among the works it cites.
α \alpha -rank: Multi-agent evaluation by evolution
Omidshafiei, S., Papadimitriou, C., Piliouras, G., Tuyls, K., Rowland, M., Lespiau, J.-B., Czarnecki, W. M., Lanctot, M., Perolat, J., and Munos, R · 2019
Later among the works it cites.
Real world games look like spinning tops
Czarnecki, W. M., Gidel, G., Tracey, B., Tuyls, K., Omidshafiei, S., Balduzzi, D., and Jaderberg, M · 2020
Later among the works it cites.
Harnessing distribution ratio estimators for learning agents with quality and diversity
Gangwani, T., Peng, J., and Zhou, Y · 2020
Later among the works it cites.
Google research football: A novel reinforcement learning environment
Kurach, K., Raichuk, A., Stańczyk, P., Zajac, M., Bachem, O., Espeholt, L., Riquelme, C., Vincent, D., Michalski, M., Bousquet, O., et al · 2020
Later among the works it cites.
Pipeline psro: A scalable approach for finding approximate nash equilibria in large games
McAleer, S., Lanier, J., Fox, R., and Baldi, P · 2020
Later among the works it cites.
Effective diversity in population-based reinforcement learning
Parker-Holder, J., Pacchiano, A., Choromanski, K., and Roberts, S · 2020
Later among the works it cites.
An overview of multi-agent reinforcement learning from game theoretical perspective
Yang, Y. and Wang, J · 2020
Later among the works it cites.
Towards playing full moba games with deep reinforcement learning
Ye, D., Chen, G., Zhang, W., Chen, S., Yuan, B., Liu, B., Chen, J., Liu, Z., Qiu, F., Yu, H., et al · 2020
Later among the works it cites.
Smarts: Scalable multi-agent reinforcement learning training school for autonomous driving
Zhou, M., Luo, J., Villela, J., Yang, Y., Rusu, D., Miao, J., Zhang, W., Alban, M., Fadakar, I., Chen, Z., et al · 2020
Later among the works it cites.
Dinh, L. C., Yang, Y., Tian, Z., Nieves, N. P., Slumbers, O., Mguni, D. H., and Wang, J · 2021
Closest in time.
Diverse auto-curriculum is critical for successful real-world multiagent learning systems
Yang, Y., Luo, J., Wen, Y., Slumbers, O., Graves, D., Bou Ammar, H., Wang, J., and Taylor, M. E · 2021
Closest in time.