Fetching the paper…
Reading the bibliography…
Solving strategic games with huge action space is a critical yet under-explored topic in economics, operations research and artificial intelligence.
Zur theorie der gesellschaftsspiele
J v Neumann · 1928
Earlier work this paper cites.
Equilibrium points in n-person games
John F Nash et al · 1950
Earlier work this paper cites.
Iterative solution of games by fictitious play
George W Brown · 1951
Earlier work this paper cites.
Iterative solution of games by fictitious play
GW Brown · 1951
Earlier work this paper cites.
Theory of games and economic behavior
Oskar Morgenstern and John Von Neumann · 1953
Earlier work this paper cites.
Stability and perfection of Nash equilibria
Eric Van Damme · 1991
Earlier work this paper cites.
Adaptive game playing using multiplicative weights
Yoav Freund and Robert E Schapire · 1999
Earlier work this paper cites.
Rational and convergent learning in stochastic games
Michael Bowling and Manuela Veloso · 2001
Earlier work this paper cites.
A general class of adaptive strategies
Sergiu Hart and Andreu Mas-Colell · 2001
Earlier work this paper cites.
The nonstochastic multiarmed bandit problem
Peter Auer, Nicolo Cesa-Bianchi, Yoav Freund, and Robert E Schapire · 2002
Earlier work this paper cites.
Planning in the presence of cost functions controlled by an adversary
H Brendan McMahan, Geoffrey J Gordon, and Avrim Blum · 2003
Earlier work this paper cites.
On the optimal strategy in a random game
Johan Jonasson et al · 2004
Earlier work this paper cites.
Prediction, learning, and games
Nicolo Cesa-Bianchi and Gábor Lugosi · 2006
Cited alongside, same era.
Settling the complexity of two-player nash equilibrium
Xi Chen and Xiaotie Deng · 2006
Cited alongside, same era.
Generalised weakened fictitious play
David S Leslie and Edmund J Collins · 2006
Cited alongside, same era.
Learning, regret minimization, and equilibria
Avrim Blum and Yishay Monsour · 2007
Cited alongside, same era.
Improved second-order bounds for prediction with expert advice
Nicolo Cesa-Bianchi, Yishay Mansour, and Gilles Stoltz · 2007
Cited alongside, same era.
Regret minimization in games with incomplete information
Martin Zinkevich, Michael Johanson, Michael Bowling, and Carmelo Piccione · 2007
Cited alongside, same era.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto · 2018
Later among the works it cites.
Openspiel: A framework for reinforcement learning in games
Marc Lanctot, Edward Lockhart, Jean-Baptiste Lespiau, Vinicius Zambaldi, Satyaki Upadhyay, Julien Pérolat, Sriram Srinivasan, Finbarr Timbers, Karl Tuyls, Shayegan Omidshafiei, et al · 2019
Later among the works it cites.
α \alpha -rank: Multi-agent evaluation by evolution
Shayegan Omidshafiei, Christos Papadimitriou, Georgios Piliouras, Karl Tuyls, Mark Rowland, Jean-Baptiste Lespiau, Wojciech M Czarnecki, Marc Lanctot, Julien Perolat, and Remi Munos · 2019
Later among the works it cites.
Grandmaster level in starcraft ii using multi-agent reinforcement learning
Oriol Vinyals, Igor Babuschkin, Wojciech M Czarnecki, Michaël Mathieu, Andrew Dudzik, Junyoung Chung, David H Choi, Richard Powell, Timo Ewalds, Petko Georgiev, et al · 2019
Later among the works it cites.
Real world games look like spinning tops
Wojciech M Czarnecki, Gauthier Gidel, Brendan Tracey, Karl Tuyls, Shayegan Omidshafiei, David Balduzzi, and Max Jaderberg · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Monte carlo sampling for regret minimization in extensive games
Marc Lanctot, Kevin Waugh, Martin Zinkevich, and Michael H Bowling · 2009
Cited alongside, same era.
On the rate of convergence of fictitious play
Felix Brandt, Felix Fischer, and Paul Harrenstein · 2010
Cited alongside, same era.
Online learning and online convex optimization
Shai Shalev-Shwartz et al · 2011
Cited alongside, same era.
Using response functions to measure strategy strength
Trevor Davis, Neil Burch, and Michael Bowling · 2014
Cited alongside, same era.
A unified game-theoretic approach to multiagent reinforcement learning
Marc Lanctot, Vinicius Zambaldi, Audrunas Gruslys, Angeliki Lazaridou, Karl Tuyls, Julien Pérolat, David Silver, and Thore Graepel · 2017
Cited alongside, same era.
Later among the works it cites.
Pipeline PSRO: A scalable approach for finding approximate nash equilibria in large games
Stephen McAleer, John Lanier, Roy Fox, and Pierre Baldi · 2020
Later among the works it cites.
A deterministic linear program solver in current matrix multiplication time
Jan van den Brand · 2020
Later among the works it cites.
α \alpha α \alpha -rank: Practically scaling α \alpha -rank through stochastic optimisation
Yaodong Yang, Rasul Tutunov, Phu Sakulwongtana, and Haitham Bou Ammar · 2020
Later among the works it cites.
An overview of multi-agent reinforcement learning from game theoretical perspective
Yaodong Yang and Jun Wang · 2020
Later among the works it cites.
XDO: A double oracle algorithm for extensive-form games
Stephen McAleer, John Lanier, Pierre Baldi, and Roy Fox · 2021
Closest in time.
Modelling behavioural diversity for learning in open-ended games
Nicolas Perez Nieves, Yaodong Yang, Oliver Slumbers, David Henry Mguni, and Jun Wang · 2021
Closest in time.