Fetching the paper…
Reading the bibliography…
Game theory has been increasingly applied in settings where the game is not known outright, but has to be estimated by sampling.
α \alpha α \alpha -Rank: Practically Scaling α \alpha -Rank through Stochastic Optimisation
Yang, Y.; Tutunov, R.; Sakulwongtana, P.; and Ammar, H. B. 2019 · 1909
Earlier work this paper cites.
Dota 2 with Large Scale Deep Reinforcement Learning
Berner, C.; Brockman, G.; Chan, B.; Cheung, V.; Debiak, P.; Dennison, C.; Farhi, D.; Fischer, Q.; Hashme, S.; Hesse, C.; et al. 2019 · 1912
Earlier work this paper cites.
Non-cooperative games
Nash, J. 1951 · 1951
Earlier work this paper cites.
The rating of chessplayers, past and present
Elo, A. E. 1978 · 1978
Earlier work this paper cites.
Game theory, 1991
Fudenberg, D.; and Tirole, J. 1991 · 1991
Earlier work this paper cites.
Entropy and inference, revisited
Nemenman, I.; Shafee, F.; and Bialek, W. 2002 · 2002
Earlier work this paper cites.
Choosing samples to compute heuristic-strategy Nash equilibrium
Walsh, W. E.; Parkes, D. C.; and Das, R. 2003 · 2003
Earlier work this paper cites.
Comparing partial rankings
Fagin, R.; Kumar, R.; Mahdian, M.; Sivakumar, D.; and Vee, E. 2006 · 2006
Earlier work this paper cites.
Methods for empirical game-theoretic analysis
Wellman, M. P. 2006 · 2006
Earlier work this paper cites.
TrueSkill™: a Bayesian skill rating system
Herbrich, R.; Minka, T.; and Graepel, T. 2007 · 2007
Earlier work this paper cites.
Searching for approximate equilibria in empirical games
Jordan, P. R.; Vorobeychik, Y.; and Wellman, M. P. 2008 · 2008
Cited alongside, same era.
Bayesian probabilistic matrix factorization using Markov chain Monte Carlo
Salakhutdinov, R.; and Mnih, A. 2008 · 2008
Cited alongside, same era.
Optimal transport: old and new , volume 338
Villani, C. 2008 · 2008
Cited alongside, same era.
The complexity of computing a Nash equilibrium
Daskalakis, C.; Goldberg, P. W.; and Papadimitriou, C. H. 2009 · 2009
Cited alongside, same era.
Gaussian process optimization in the bandit setting: No regret and experimental design
Srinivas, N.; Krause, A.; Kakade, S. M.; and Seeger, M. 2009 · 2009
Cited alongside, same era.
A survey of collaborative filtering techniques
Su, X.; and Khoshgoftaar, T. M. 2009 · 2009
A unified game-theoretic approach to multiagent reinforcement learning
Lanctot, M.; Zambaldi, V.; Gruslys, A.; Lazaridou, A.; Tuyls, K.; Pérolat, J.; Silver, D.; and Graepel, T. 2017 · 2017
Later among the works it cites.
Mastering chess and shogi by self-play with a general reinforcement learning algorithm
Silver, D.; Hubert, T.; Schrittwieser, J.; Antonoglou, I.; Lai, M.; Guez, A.; Lanctot, M.; Sifre, L.; Kumaran, D.; Graepel, T.; et al. 2017 · 2017
Later among the works it cites.
Re-evaluating evaluation
Balduzzi, D.; Tuyls, K.; Perolat, J.; and Graepel, T. 2018 · 2018
Later among the works it cites.
Trueskill 2: An improved bayesian skill rating system
Minka, T.; Cleven, R.; and Zaykov, Y. 2018 · 2018
Later among the works it cites.
Learning to optimize via information-directed sampling
Russo, D.; and van Roy, B. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Improved Algorithms for Linear Stochastic Bandits
Abbasi-yadkori, Y.; Pál, D.; and Szepesvári, C. 2011 · 2011
Cited alongside, same era.
Displacement interpolation using Lagrangian mass transport
Bonneel, N.; Van De Panne, M.; Paris, S.; and Heidrich, W. 2011 · 2011
Cited alongside, same era.
Modeling intransitivity in matchup and comparison data
Chen, S.; and Joachims, T. 2016 · 2016
Cited alongside, same era.
POT Python Optimal Transport library
Flamary, R.; and Courty, N. 2017 · 2017
Cited alongside, same era.
ndd - Bayesian entropy estimation from discrete data
Simone, M. ????
Cited in the paper.
Omidshafiei, S.; Papadimitriou, C.; Piliouras, G.; Tuyls, K.; Rowland, M.; Lespiau, J.-B.; Czarnecki, W. M.; Lanctot, M.; Perolat, J.; and Munos, R. 2019 · 2019
Later among the works it cites.
Multiagent Evaluation under Incomplete Information
Rowland, M.; Omidshafiei, S.; Tuyls, K.; Perolat, J.; Valko, M.; Piliouras, G.; and Munos, R. 2019 · 2019
Later among the works it cites.
Grandmaster level in StarCraft II using multi-agent reinforcement learning
Vinyals, O.; Babuschkin, I.; Czarnecki, W. M.; Mathieu, M.; Dudzik, A.; Chung, J.; Choi, D. H.; Powell, R.; Ewalds, T.; Georgiev, P.; et al. 2019 · 2019
Later among the works it cites.
A Generalized Training Approach for Multiagent Learning
Muller, P.; Omidshafiei, S.; Rowland, M.; Tuyls, K.; Perolat, J.; Liu, S.; Hennes, D.; Marris, L.; Lanctot, M.; Hughes, E.; Wang, Z.; Lever, G.; Heess, N.; Graepel, T.; and Munos, R. 2020 · 2020
Later among the works it cites.
Bounds and dynamics for empirical game theoretic analysis
Tuyls, K.; Perolat, J.; Lanctot, M.; Hughes, E.; Everett, R.; Leibo, J. Z.; Szepesvári, C.; and Graepel, T. 2020 · 2020
Later among the works it cites.