Error detecting and error correcting codes
Richard W Hamming · 1950
Earlier work this paper cites.
Non-cooperative games
John Nash · 1951
Earlier work this paper cites.
Stochastic games
Lloyd S Shapley · 1953
Earlier work this paper cites.
Markov games as a framework for multi-agent reinforcement learning
Michael L Littman · 1994
Earlier work this paper cites.
A simple adaptive procedure leading to correlated equilibrium
Sergiu Hart and Andreu Mas-Colell · 2000
Earlier work this paper cites.
Friend-or-foe q-learning in general-sum games
Michael L Littman · 2001
Earlier work this paper cites.
Nash q-learning for general-sum stochastic games
Junling Hu and Michael P Wellman · 2003
Earlier work this paper cites.
Adaptive heuristics
Sergiu Hart · 2005
Earlier work this paper cites.
Incomplete information and internal regret in prediction of individual sequences
Gilles Stoltz · 2005
Earlier work this paper cites.
Prediction, learning, and games
Nicolo Cesa-Bianchi and Gábor Lugosi · 2006
Earlier work this paper cites.
From external to internal regret
Avrim Blum and Yishay Mansour · 2007
Earlier work this paper cites.
Algorithmic Game Theory
Noam Nisan, Tim Roughgarden, Eva Tardos, and Vijay V Vazirani · 2007
Earlier work this paper cites.
Extensive-form correlated equilibrium: Definition and computational complexity
Bernhard Von Stengel and Françoise Forges · 2008
Earlier work this paper cites.
Multiplicative updates outperform generic no-regret learning in congestion games
Robert Kleinberg, Georgios Piliouras, and Éva Tardos · 2009
Earlier work this paper cites.
On the complexity of approximating a nash equilibrium
Constantinos Daskalakis · 2013
Earlier work this paper cites.
Strategy iteration is strongly polynomial for 2-player turn-based stochastic games with a constant discount factor
Thomas Dueholm Hansen, Peter Bro Miltersen, and Uri Zwick · 2013
Earlier work this paper cites.
Query complexity of correlated equilibrium
Yakov Babichenko and Siddharth Barman · 2015
Earlier work this paper cites.
Well-supported versus approximate nash equilibria: Query complexity of large games
Original
Xi Chen, Yu Cheng, and Bo Tang · 2015
Earlier work this paper cites.
Sample complexity of episodic fixed-horizon reinforcement learning
Original
Christoph Dann and Emma Brunskill · 2015
Earlier work this paper cites.
Learning equilibria of games via payoff queries
John Fearnley, Martin Gairing, Paul W Goldberg, and Rahul Savani · 2015
Earlier work this paper cites.
Explore no more: Improved high-probability regret bounds for non-stochastic bandits
Original
Gergely Neu · 2015
Earlier work this paper cites.
Fast convergence of regularized learning in games
Original
Vasilis Syrgkanis, Alekh Agarwal, Haipeng Luo, and Robert E Schapire · 2015
Earlier work this paper cites.