Fetching the paper…
Reading the bibliography…
We study the problem of training a principal in a multi-agent general-sum game using reinforcement learning (RL).
Variabilità e mutabilità
Gini, C. 1912 · 1912
Earlier work this paper cites.
The theory of risk aversion
Arrow, K. J. 1971 · 1971
Earlier work this paper cites.
From substantive to procedural rationality
Simon, H. A. 1976 · 1976
Earlier work this paper cites.
Reputation and imperfect information
Kreps, D. M.; and Wilson, R. 1982 · 1982
Earlier work this paper cites.
Learning Quickly When Irrelevant Attributes Abound: A New Linear-threshold Algorithm
Littlestone, N. 1987 · 1987
Earlier work this paper cites.
Reputation and Equilibrium Selection in Games with a Patient Player
Fudenberg, D.; and Levine, D. K. 1989 · 1989
Earlier work this paper cites.
Artificial Adaptive Agents in Economic Theory
Holland, J. H.; and Miller, J. H. 1991 · 1991
Earlier work this paper cites.
Security games with interval uncertainty
Kiekintveld, C.; Islam, T.; and Kreinovich, V. 2013 · 1993
Earlier work this paper cites.
Calibrated Learning and Correlated Equilibrium
Foster, D. P.; and Vohra, R. V. 1997 · 1997
Earlier work this paper cites.
A Simple Adaptive Procedure Leading to Correlated Equilibrium
Hart, S.; and Mas-Colell, A. 2000 · 2000
Earlier work this paper cites.
Robust Reinforcement Learning
Morimoto, J.; and Doya, K. 2000 · 2000
Earlier work this paper cites.
Agent-based modeling: Methods and techniques for simulating human systems
Bonabeau, E. 2002 · 2002
Earlier work this paper cites.
The AI Economist: Improving Equality and Productivity with AI-Driven Tax Policies
Zheng, S.; Trott, A.; Srinivasa, S.; Naik, N.; Gruesbeck, M.; Parkes, D. C.; and Socher, R. 2020 · 2004
Earlier work this paper cites.
Robust game theory
Aghassi, M.; and Bertsimas, D. 2006 · 2006
Cited alongside, same era.
Prediction, Learning, and Games
Cesa-Bianchi, N.; and Lugosi, G. 2006 · 2006
Cited alongside, same era.
Computing the optimal strategy to commit to
Conitzer, V.; and Sandholm, T. 2006 · 2006
Cited alongside, same era.
Robust Reinforcement Learning with Wasserstein Constraint
Hou, L.; Pang, L.; Hong, X.; Lan, Y.; Ma, Z.; and Yin, D. 2020 · 2006
Cited alongside, same era.
Computing correlated equilibria in multi-player games
Papadimitriou, C. H.; and Roughgarden, T. 2008 · 2008
Cited alongside, same era.
The Complexity of Computing a Nash Equilibrium
Daskalakis, C.; Goldberg, P. W.; and Papadimitriou, C. H. 2009 · 2009
Cited alongside, same era.
Intrinsic Robustness of the Price of Anarchy
Roughgarden, T. 2015 · 2015
Later among the works it cites.
The computational power of optimization in online learning
Hazan, E.; and Koren, T. 2016 · 2016
Later among the works it cites.
Mechanism Design
Myerson, R. B. 2016 · 2016
Later among the works it cites.
CAPTURE: A new predictive anti-poaching tool for wildlife protection
Nguyen, T. H.; Sinha, A.; Gholami, S.; Plumptre, A.; Joppa, L.; Tambe, M.; Driciru, M.; Wanyama, F.; Rwetsiba, A.; and Critchlow, R. 2016 · 2016
Later among the works it cites.
Robust Adversarial Reinforcement Learning
Pinto, L.; Davidson, J.; Sukthankar, R.; and Gupta, A. 2017 · 2017
Later among the works it cites.
Proximal Policy Optimization Algorithms
Schulman, J.; Wolski, F.; Dhariwal, P.; Radford, A.; and Klimov, O. 2017 · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Using Game Theory for Los Angeles Airport Security
Pita, J.; Jain, M.; Ordóñez, F.; Portway, C.; Tambe, M.; Western, C.; Paruchuri, P.; and Kraus, S. 2009 · 2009
Cited alongside, same era.
Robust solutions to Stackelberg games: Addressing bounded rationality and limited observations in human cognition
Pita, J.; Jain, M.; Tambe, M.; Ordóñez, F.; and Kraus, S. 2010 · 2010
Cited alongside, same era.
A General Framework for Computing Optimal Correlated Equilibria in Compact Games
Jiang, A. X.; and Leyton-Brown, K. 2011 · 2011
Cited alongside, same era.
Security Games with Multiple Attacker Resources
Korzhyk, D.; Conitzer, V.; and Parr, R. 2011 · 2011
Cited alongside, same era.
A robust approach to addressing human adversaries in security games
Pita, J.; John, R.; Maheswaran, R.; Tambe, M.; and Kraus, S. 2012 · 2012
Cited alongside, same era.
Analyzing the Effectiveness of Adversary Modeling in Security Games
Nguyen, T. H.; Yang, R.; Azaria, A.; Kraus, S.; and Tambe, M. 2013 · 2013
Cited alongside, same era.
Later among the works it cites.
Optimal Auctions through Deep Learning
Duetting, P.; Feng, Z.; Narasimhan, H.; Parkes, D. C.; and Ravindranath, S. S. 2019 · 2019
Later among the works it cites.
Robust Multi-Agent Reinforcement Learning via Minimax Deep Deterministic Policy Gradient
Li, S.; Wu, Y.; Cui, X.; Dong, H.; Fang, F.; and Russell, S. J. 2019 · 2019
Later among the works it cites.
Action Robust Reinforcement Learning and Applications in Continuous Control
Tessler, C.; Efroni, Y.; and Mannor, S. 2019 · 2019
Later among the works it cites.
Bilevel programming methods for computing single-leader-multi-follower equilibria in normal-form and polymatrix games
Basilico, N.; Coniglio, S.; Gatti, N.; and Marchesi, A. 2020 · 2020
Later among the works it cites.
Implicit competitive regularization in GANs
Schäfer, F.; Zheng, H.; and Anandkumar, A. 2020 · 2020
Later among the works it cites.
Human-Level Performance in No-Press Diplomacy via Equilibrium Search
Gray, J.; Lerer, A.; Bakhtin, A.; and Brown, N. 2021 · 2021
Closest in time.