Fetching the paper…
Reading the bibliography…
Real-world competitive games, such as chess, go, or StarCraft II, rely on Elo models to measure the strength of their players.
The USCF rating system
Arpad Elo · 1961
Earlier work this paper cites.
Linear algebra , volume 23
Werner H. Greub · 1975
Earlier work this paper cites.
The rating of Chess players, past and present (Arco, New York)
Arpad Elo · 1978
Earlier work this paper cites.
Bradley-Terry-Luce models with an ordered response
Gerhard Tutz · 1986
Earlier work this paper cites.
On the limited memory bfgs method for large scale optimization
Dong C. Liu and Jorge Nocedal · 1989
Earlier work this paper cites.
Asian american preferences for counselor characteristics: Application of the Bradley-Terry-Luce model to paired comparison data
Donald R Atkinson, Bruce E Wampold, Susana M Lowe, Linda Matthews, and Hyun-Nie Ahn · 1998
Earlier work this paper cites.
Dynamic paired comparison models with stochastic variances
Mark E Glickman · 2001
Earlier work this paper cites.
Choosing samples to compute heuristic-strategy nash equilibrium
William E Walsh, David C Parkes, and Rajarshi Das · 2003
Earlier work this paper cites.
Convex optimization
Stephen Boyd and Lieven Vandenberghe · 2004
Earlier work this paper cites.
Trueskill™: a bayesian skill rating system
Ralf Herbrich, Tom Minka, and Thore Graepel · 2006
Earlier work this paper cites.
Methods for empirical game-theoretic analysis
Michael P Wellman · 2006
Earlier work this paper cites.
Exact matrix completion via convex optimization
Emmanuel J. Candès and Benjamin Recht · 2008
Earlier work this paper cites.
Exercices de mathématiques oraux x-ens algèbre , volume 3
Serge Francinou, Hervé Gianella, and Serge Nicolas · 2008
Earlier work this paper cites.
Matrix completion with noise
Emmanuel J. Candès and Yaniv Plan · 2009
Earlier work this paper cites.
Settling the complexity of computing two-player nash equilibria
Xi Chen, Xiaotie Deng, and Shang-Hua Teng · 2009
Cited alongside, same era.
The complexity of computing a nash equilibrium
Constantinos Daskalakis, Paul W. Goldberg, and Christos H. Papadimitriou · 2009
Cited alongside, same era.
How I won the" chess ratings-elo vs the rest of the world" competition
Yannis Sismanis · 2010
Cited alongside, same era.
Statistical ranking and combinatorial hodge theory
Xiaoye Jiang, Lek-Heng Lim, Yuan Yao, and Yinyu Ye · 2011
Cited alongside, same era.
Iterative ranking from pair-wise comparisons
Sahand Negahban, Sewoong Oh, and Devavrat Shah · 2012
Cited alongside, same era.
Matrix estimation by universal singular value thresholding
Sourav Chatterjee · 2015
Open-ended learning in symmetric zero-sum games
David Balduzzi, Marta Garnelo, Yoram Bachrach, Wojciech Czarnecki, Julien Perolat, Max Jaderberg, and Thore Graepel · 2019
Later among the works it cites.
Dota 2 with large scale deep reinforcement learning
Christopher Berner, Greg Brockman, Brooke Chan, Vicki Cheung, Przemysław Dębiak, Christy Dennison, David Farhi, Quirin Fischer, Shariq Hashme, Chris Hesse, et al · 2019
Later among the works it cites.
Competing against equilibria in zero-sum games with evolving payoffs
Adrian Rivera Cardoso, Jacob Abernethy, He Wang, and Huan Xu · 2019
Later among the works it cites.
Human-level performance in 3d multiplayer games with population-based reinforcement learning
Max Jaderberg, Wojciech M. Czarnecki, Iain Dunning, Luke Marris, Guy Lever, Antonio Garcia Castaneda, Charles Beattie, Neil C. Rabinowitz, Ari S Morcos, Avraham Ruderman, et al · 2019
Later among the works it cites.
α \alpha -rank: Multi-agent evaluation by evolution
Shayegan Omidshafiei, Christos Papadimitriou, Georgios Piliouras, Karl Tuyls, Mark Rowland, Jean-Baptiste Lespiau, Wojciech M. Czarnecki, Marc Lanctot, Julien Perolat, and Remi Munos · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
A continuous model for ratings
Pierre-Emmanuel Jabin and Stéphane Junca · 2015
Cited alongside, same era.
A unified game-theoretic approach to multiagent reinforcement learning
Marc Lanctot, Vinicius Zambaldi, Audrunas Gruslys, Angeliki Lazaridou, Karl Tuyls, Julien Pérolat, David Silver, and Thore Graepel · 2017
Cited alongside, same era.
Simple, robust and optimal ranking from pairwise comparisons
Nihar B Shah and Martin J Wainwright · 2017
Cited alongside, same era.
Starcraft II: A new challenge for reinforcement learning
Oriol Vinyals, Timo Ewalds, Sergey Bartunov, Petko Georgiev, Alexander Sasha Vezhnevets, Michelle Yeo, Alireza Makhzani, Heinrich Küttler, John Agapiou, Julian Schrittwieser, et al · 2017
Cited alongside, same era.
Re-evaluating evaluation
David Balduzzi, Karl Tuyls, Julien Perolat, and Thore Graepel · 2018
Cited alongside, same era.
Emergent complexity via multi-agent competition
Trapit Bansal, Jakub Pachocki, Szymon Sidor, Ilya Sutskever, and Igor Mordatch · 2018
Cited alongside, same era.
Later among the works it cites.
Grandmaster level in starcraft ii using multi-agent reinforcement learning
Oriol Vinyals, Igor Babuschkin, Wojciech M. Czarnecki, Michaël Mathieu, Andrew Dudzik, Junyoung Chung, David H. Choi, Richard Powell, Timo Ewalds, Petko Georgiev, et al · 2019
Later among the works it cites.
Real world games look like spinning tops
Wojciech M. Czarnecki, Gauthier Gidel, Brendan Tracey, Karl Tuyls, Shayegan Omidshafiei, David Balduzzi, and Max Jaderberg · 2020
Later among the works it cites.
Scipy 1.0: fundamental algorithms for scientific computing in python
Pauli Virtanen, Ralf Gommers, Travis E. Oliphant, Matt Haberland, Tyler Reddy, David Cournapeau, Evgeni Burovski, Pearu Peterson, Warren Weckesser, Jonathan Bright, et al · 2020
Later among the works it cites.
Eigengame: PCA as a nash equilibrium
Ian Gemp, Brian McWilliams, Claire Vernade, and Thore Graepel · 2021
Later among the works it cites.
From motor control to team play in simulated humanoid football
Siqi Liu, Guy Lever, Zhe Wang, Josh Merel, SM Eslami, Daniel Hennes, Wojciech M. Czarnecki, Yuval Tassa, Shayegan Omidshafiei, Abbas Abdolmaleki, et al · 2021
Later among the works it cites.
Estimating α \alpha -rank by maximizing information gain
Tabish Rashid, Cheng Zhang, and Kamil Ciosek · 2021
Later among the works it cites.
Playing for the legend in the age of empires II online community
Matthew Horrigan · 2022
Closest in time.
Learning to identify top Elo ratings: A dueling bandits approach
Xue Yan, Yali Du, Binxin Ru, Jun Wang, Haifeng Zhang, and Xu Chen · 2022
Closest in time.