Fetching the paper…
Reading the bibliography…
Driven by recent successes in two-player, zero-sum game solving and playing, artificial intelligence work on games has increasingly focused on algorithms that produce equilibrium-based strategies.
OpenSpiel: A Framework for Reinforcement Learning in Games
Lanctot, M.; Lockhart, E.; Lespiau, J.-B.; Zambaldi, V.; Upadhyay, S.; Pérolat, J.; Srinivasan, S.; Timbers, F.; Tuyls, K.; Omidshafiei, S.; Hennes, D.; Morrill, D.; Muller, P.; Ewalds, T.; Faulkner, R.; Kramár, J.; Vylder, B. D.; Saeta, B.; Bradbury, J.; Ding, D.; Borgeaud, S.; Lai, M.; Schrittwieser, J.; Anthony, T.; Hughes, E.; Danihelka, I.; and Ryan-Davis, J. 2019 · 1908
Earlier work this paper cites.
Non-Cooperative Games
Nash, J. 1951 · 1951
Earlier work this paper cites.
Extensive Games and the Problem of Information
Kuhn, H. W. 1953 · 1953
Earlier work this paper cites.
Approximation to Bayes risk in repeated play
Hannan, J. 1957 · 1957
Earlier work this paper cites.
Some Topics in Two-Person Games
Shapley, L. 1964 · 1964
Earlier work this paper cites.
Subjectivity and correlation in randomized strategies
Aumann, R. J. 1974 · 1974
Earlier work this paper cites.
Reexamination of the Perfectness Concept for Equilibrium Points in Extensive Games
Selten, R. 1974 · 1974
Earlier work this paper cites.
Strategically zero-sum games: the class of games whose completely mixed equilibria cannot be improved upon
Moulin, H.; and Vial, J.-P. 1978 · 1978
Earlier work this paper cites.
Sequential equilibria
Kreps, D. M.; and Wilson, R. 1982 · 1982
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Williams, R. J. 1992 · 1992
Earlier work this paper cites.
Calibrated learning and correlated equilibrium
Foster, D. P.; and Vohra, R. V. 1997 · 1997
Earlier work this paper cites.
Regret in the on-line decision problem
Foster, D. P.; and Vohra, R. 1999 · 1999
Cited alongside, same era.
A Simple Adaptive Procedure Leading to Correlated Equilibrium
Hart, S.; and Mas-Colell, A. 2000 · 2000
Cited alongside, same era.
Policy Gradient Methods for Reinforcement Learning with Function Approximation
Sutton, R. S.; McAllester, D.; Singh, S.; and Mansour, Y. 2000 · 2000
Cited alongside, same era.
Computionally Efficient Coordination in Games Trees
Forges, F.; and von Stengel, B. 2002 · 2002
Cited alongside, same era.
A general class of no-regret learning algorithms and game-theoretic equilibria
Greenwald, A.; Jafari, A.; and Marks, C. 2003 · 2003
Cited alongside, same era.
Extensive-form correlated equilibrium: Definition and computational complexity
von Stengel, B.; and Forges, F. 2008 · 2008
Cited alongside, same era.
Superhuman AI for Heads-Up No-Limit Poker: Libratus Beats Top Professionals
Brown, N.; and Sandholm, T. 2018 · 2018
Later among the works it cites.
A general reinforcement learning algorithm that masters chess, shogi, and Go through self-play
Silver, D.; Hubert, T.; Schrittwieser, J.; Antonoglou, I.; Lai, M.; Guez, A.; Lanctot, M.; Sifre, L.; Kumaran, D.; Graepel, T.; Lillicrap, T.; Simonyan, K.; and Hassabis, D. 2018 · 2018
Later among the works it cites.
Superhuman AI for Multiplayer Poker
Brown, N.; and Sandholm, T. 2019 · 2019
Later among the works it cites.
Revisiting CFR+ and alternating updates
Burch, N.; Moravcik, M.; and Schmid, M. 2019 · 2019
Later among the works it cites.
Learning to correlate in multi-player general-sum sequential games
Celli, A.; Marchesi, A.; Bianchi, T.; and Gatti, N. 2019 · 2019
Later among the works it cites.
Grandmaster level in StarCraft II using multi-agent reinforcement learning
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A Sampling-Based Approach to Computing Equilibria in Succinct Extensive-Form Games
Dudík, M.; and Gordon, G. J. 2009 · 2009
Cited alongside, same era.
Hindsight and Sequential Rationality of Correlated Play: Corrections
Morrill, D.; D’Orazio, R.; Sarfati, R.; Lanctot, M.; Wright, J. R.; Greenwald, A.; and Bowling, M. 2022 · 2012
Cited alongside, same era.
Mastering the Game of Go with Deep Neural Networks and Tree Search
Silver, D.; Huang, A.; Maddison, C. J.; Guez, A.; Sifre, L.; van den Driessche, G.; Schrittwieser, J.; Antonoglou, I.; Panneershelvam, V.; Lanctot, M.; Dieleman, S.; Grewe, D.; Nham, J.; Kalchbrenner, N.; Sutskever, I.; Lillicrap, T.; Leach, M.; Kavukcuoglu, K.; Graepel, T.; and Hassabis, D. 2016 · 2016
Cited alongside, same era.
Deepstack: Expert-Level Artificial Intelligence in Heads-Up No-Limit Poker
Moravčík, M.; Schmid, M.; Burch, N.; Lisỳ, V.; Morrill, D.; Bard, N.; Davis, T.; Waugh, K.; Johanson, M.; and Bowling, M. 2017 · 2017
Cited alongside, same era.
Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Silver, D.; Hubert, T.; Schrittwieser, J.; Antonoglou, I.; Lai, M.; Guez, A.; Lanctot, M.; Sifre, L.; Kumaran, D.; Graepel, T.; Lillicrap, T. P.; Simonyan, K.; and Hassabis, D. 2017a
Cited in the paper.
Mastering the game of Go without human knowledge
Silver, D.; Schrittwieser, J.; Simonyan, K.; Antonoglou, I.; Huang, A.; Guez, A.; Hubert, T.; Baker, L.; Lai, M.; Bolton, A.; Chen, Y.; Lillicrap, T.; Hui, F.; Sifre, L.; van den Driessche, G.; Graepel, T.; and Hassabis, D. 2017b
Cited in the paper.
Vinyals, O.; Babuschkin, I.; Czarnecki, W. M.; Mathieu, M.; Dudzik, A.; Chung, J.; Choi, D. H.; Powell, R.; Ewalds, T.; Georgiev, P.; et al. 2019 · 2019
Later among the works it cites.
No-regret learning dynamics for extensive-form correlated equilibrium
Celli, A.; Marchesi, A.; Farina, G.; and Gatti, N. 2020 · 2020
Closest in time.
Coarse Correlation in Extensive-Form Games
Farina, G.; Bianchi, T.; and Sandholm, T. 2020 · 2020
Closest in time.
Hindsight and Sequential Rationality of Correlated Play
Morrill, D.; D’Orazio, R.; Sarfati, R.; Lanctot, M.; Wright, J. R.; Greenwald, A.; and Bowling, M. 2021 · 2021
Closest in time.
Personal communication
MacQueen, R. 2022 · 2022
Closest in time.