Fetching the paper…
Reading the bibliography…
In this paper we investigate the Follow the Regularized Leader dynamics in sequential imperfect information games (IIG).
Iterative solutions of games by fictitious play
Brown, G. W · 1951
Earlier work this paper cites.
Evolutionarily stable strategies and game dynamics
Taylor and Jonker · 1978
Earlier work this paper cites.
Population dynamics from game theory
Zeeman, E · 1980
Earlier work this paper cites.
Dynamics of the evolution of animal conflicts
Zeeman, E · 1981
Earlier work this paper cites.
Markov games as a framework for multi-agent reinforcement learning
Littman, M. L · 1994
Earlier work this paper cites.
Quantal response equilibria for normal form games
McKelvey, R. D. and Palfrey, T. R · 1995
Earlier work this paper cites.
Evolutionary game theory
Weibull, J · 1997
Earlier work this paper cites.
The Theory of Learning in Games
Fudenberg, D. and Levine, D · 1998
Earlier work this paper cites.
Nash convergence of gradient dynamics in general-sum games
Singh, S. P., Kearns, M. J., and Mansour, Y · 2000
Earlier work this paper cites.
Friend-or-foe Q-learning in general-sum games
Littman, M. L · 2001
Earlier work this paper cites.
On the global convergence of stochastic fictitious play
Hofbauer, J. and Sandholm, W. H · 2002
Earlier work this paper cites.
Value function approximation in zero-sum Markov games
Lagoudakis, M. G. and Parr, R · 2002
Earlier work this paper cites.
Nash Q-learning for general-sum stochastic games
Hu, J. and Wellman, M · 2003
Earlier work this paper cites.
A selection-mutation model for Q-learning in multi-agent systems
Tuyls, K., Verbeeck, K., and Lenaerts, T · 2003
Earlier work this paper cites.
Individual Q-learning in normal form games
Leslie, D. S. and Collins, E. J · 2005
Earlier work this paper cites.
Evolutionary game theory and multi-agent reinforcement learning
Tuyls, K. and Nowé, A · 2005
Earlier work this paper cites.
Predition, Learning, and Games
Cesa-Biachi, N. and Lugosi, G · 2006
Earlier work this paper cites.
A new algorithm for generating equilibria in massive zero-sum games
M. Zinkevich, M. Bowling, N. B · 2007
Earlier work this paper cites.
Regret minimization in games with incomplete information
Zinkevich, M., Johanson, M., Bowling, M., and Piccione, C · 2008
Earlier work this paper cites.
Game Theory Evolving
Gintis, H · 2009
Earlier work this paper cites.
Time average replicator and best-reply dynamics
Hofbauer, J., Sorin, S., and Viossat, Y · 2009
Cited alongside, same era.
Monte Carlo sampling for regret minimization in extensive games
Lanctot, M., Waugh, K., Zinkevich, M., and Bowling, M · 2009
Cited alongside, same era.
Frequency adjusted multi-agent Q-learning
Kaisers, M. and Tuyls, K · 2010
Cited alongside, same era.
FAQ-learning in matrix games: Demonstrating convergence near Nash equilibria, and bifurcation of attractors in the battle of sexes
Kaisers, M. and Tuyls, K · 2011
Cited alongside, same era.
Independent reinforcement learners in cooperative markov games: A survey regarding coordination problems
Matignon, L., Laurent, G. J., and Le Fort-Piat, N · 2012
Cited alongside, same era.
Online learning and online convex optimization
Shalev-Shwartz, S. et al · 2012
The numerics of GANs
Mescheder, L., Nowozin, S., and Geiger, A · 2017
Later among the works it cites.
GANGs: Generative adversarial network games
Oliehoek, F. A., Savani, R., Gallego-Posada, J., Van der Pol, E., De Jong, E. D., and Groß, R · 2017
Later among the works it cites.
Learning Nash equilibrium for general-sum Markov games from batch data
Pérolat, J., Strub, F., Piot, B., and Pietquin, O · 2017
Later among the works it cites.
Faster rates for convex-concave games
Abernethy, J., Lai, K. A., Levy, K. Y., and Wang, J.-K · 2018
Later among the works it cites.
Multiplicative weights update in zero-sum games
Bailey, J. P. and Piliouras, G · 2018
Later among the works it cites.
The mechanics of n n -player differentiable games
Balduzzi, D., Racaniere, S., Martens, J., Foerster, J., Tuyls, K., and Graepel, T · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Monte Carlo Sampling and Regret Minimization for Equilibrium Computation and Decision-Making in Large Extensive Form Games
Lanctot, M · 2013
Cited alongside, same era.
Optimization despite chaos: Convex relaxations to complex limit sets via Poincaré recurrence
Piliouras, G. and Shamma, J. S · 2014
Cited alongside, same era.
Evolutionary dynamics of multi-agent learning: a survey
Bloembergen, D., Tuyls, K., Hennes, D., and Kaisers, M · 2015
Cited alongside, same era.
Fictitious self-play in extensive-form games
Heinrich, J., Lanctot, M., and Silver, D · 2015
Cited alongside, same era.
Approximate dynamic programming for two-player zero-sum Markov games
Pérolat, J., Scherrer, B., Piot, B., and Pietquin, O · 2015
Cited alongside, same era.
Fast convergence of regularized learning in games
Syrgkanis, V., Agarwal, A., Luo, H., and Schapire, R. E · 2015
Cited alongside, same era.
IMPALA: Scalable distributed deep-RL with importance weighted actor-learner architectures
Espeholt, L., Soyer, H., Munos, R., Simonyan, K., Mnih, V., Ward, T., Doron, Y., Firoiu, V., Harley, T., Dunning, I., et al · 2018
Later among the works it cites.
An online learning approach to generative adversarial networks
Grnarova, P., Levy, K. Y., Lucchi, A., Hofmann, T., and Krause, A · 2018
Later among the works it cites.
Cycles in adversarial regularized learning
Mertikopoulos, P., Papadimitriou, C., and Piliouras, G · 2018
Later among the works it cites.
Modeling friends and foes
Ortega, P. A. and Legg, S · 2018
Later among the works it cites.
Actor-critic policy optimization in partially observable multiagent environments
Srinivasan, S., Lanctot, M., Zambaldi, V., Pérolat, J., Tuyls, K., Munos, R., and Bowling, M · 2018
Later among the works it cites.
From Darwin to Poincaré and von Neumann: Recurrence and cycles in evolutionary and algorithmic game theory
Boone, V. and Piliouras, G · 2019
Later among the works it cites.
A theory of regularized Markov decision processes
Geist, M., Scherrer, B., and Pietquin, O · 2019
Later among the works it cites.
Negative momentum for improved game dynamics
Gidel, G., Hemmat, R. A., Pezeshki, M., Huang, G., Lepriol, R., Lacoste-Julien, S., and Mitliagkas, I · 2019
Later among the works it cites.
Openspiel: A framework for reinforcement learning in games
Lanctot, M., Lockhart, E., Lespiau, J.-B., Zambaldi, V., Upadhyay, S., Pérolat, J., Srinivasan, S., Timbers, F., Tuyls, K., Omidshafiei, S., et al · 2019
Later among the works it cites.
Stable opponent shaping in differentiable games
Letcher, A., Foerster, J., Balduzzi, D., Rocktäschel, T., and Whiteson, S · 2019
Later among the works it cites.
Computing approximate equilibria in sequential adversarial games by exploitability descent
Lockhart, E., Lanctot, M., Pérolat, J., Lespiau, J.-B., Morrill, D., Timbers, F., and Tuyls, K · 2019
Later among the works it cites.
Neural replicator dynamics
Omidshafiei, S., Hennes, D., Morrill, D., Munos, R., Perolat, J., Lanctot, M., Gruslys, A., Lespiau, J.-B., and Tuyls, K · 2019
Later among the works it cites.
Variance reduction in Monte Carlo counterfactual regret minimization (VR-MCCFR) for extensive form games using baselines
Schmid, M., Burch, N., Lanctot, M., Moravcik, M., Kadlec, R., and Bowling, M · 2019
Later among the works it cites.