Fetching the paper…
Reading the bibliography…
Autocurricular training is an important sub-area of multi-agent reinforcement learning~(MARL) that allows multiple agents to learn emergent skills in an unsupervised co-evolving scheme.
Fast exact multiplication by the hessian
B. A. Pearlmutter · 1994
Earlier work this paper cites.
Dynamic noncooperative game theory
T. Başar and G. J. Olsder · 1998
Earlier work this paper cites.
The implicit function theorem: history, theory, and applications
S. G. Krantz and H. R. Parks · 2002
Earlier work this paper cites.
Market structure and equilibrium
H. Von Stackelberg · 2010
Earlier work this paper cites.
Competitive Markov decision processes
J. Filar and K. Vrieze · 2012
Earlier work this paper cites.
Continuous control with deep reinforcement learning
T. P. Lillicrap, J. J. Hunt, A. Pritzel, N. Heess, T. Erez, Y. Tassa, D. Silver, and D. Wierstra · 2015
Earlier work this paper cites.
Benchmarking deep reinforcement learning for continuous control
Y. Duan, X. Chen, R. Houthooft, J. Schulman, and P. Abbeel · 2016
Earlier work this paper cites.
Mastering chess and shogi by self-play with a general reinforcement learning algorithm
D. Silver, T. Hubert, J. Schrittwieser, I. Antonoglou, M. Lai, A. Guez, M. Lanctot, L. Sifre, D. Kumaran, T. Graepel, et al · 2017
Earlier work this paper cites.
Robust adversarial reinforcement learning
L. Pinto, J. Davidson, R. Sukthankar, and A. Gupta · 2017
Earlier work this paper cites.
Emergent complexity via multi-agent competition
T. Bansal, J. Pachocki, S. Sidor, I. Sutskever, and I. Mordatch · 2017
Cited alongside, same era.
Multi-agent actor-critic for mixed cooperative-competitive environments
R. Lowe, Y. WU, A. Tamar, J. Harb, O. Pieter Abbeel, and I. Mordatch · 2017
Cited alongside, same era.
Learning with opponent-learning awareness
J. Foerster, R. Y. Chen, M. Al-Shedivat, S. Whiteson, P. Abbeel, and I. Mordatch · 2018
Cited alongside, same era.
Multi-agent reinforcement learning: A selective overview of theories and algorithms
K. Zhang, Z. Yang, and T. Başar · 2019
Cited alongside, same era.
Dota 2 with large scale deep reinforcement learning
C. Berner, G. Brockman, B. Chan, V. Cheung, P. P. Debiak, C. Dennison, D. Farhi, Q. Fischer, S. Hashme, C. Hesse, et al · 2019
Competitive policy optimization
M. Prajapat, K. Azizzadenesheli, A. Liniger, Y. Yue, and A. Anandkumar · 2020
Later among the works it cites.
Near-optimal algorithms for minimax optimization
T. Lin, C. Jin, and M. I. Jordan · 2020
Later among the works it cites.
Implicit learning dynamics in stackelberg games: Equilibria characterization, convergence analysis, and empirical study
T. Fiez, B. Chasnov, and L. J. Ratliff · 2020
Later among the works it cites.
An overview of multi-agent reinforcement learning from game theoretical perspective
Y. Yang and J. Wang · 2020
Later among the works it cites.
A game theoretic framework for model based reinforcement learning
A. Rajeswaran, I. Mordatch, and V. Kumar · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Grandmaster level in starcraft ii using multi-agent reinforcement learning
O. Vinyals, I. Babuschkin, W. M. Czarnecki, M. Mathieu, A. Dudzik, J. Chung, D. H. Choi, R. Powell, T. Ewalds, P. Georgiev, et al · 2019
Cited alongside, same era.
Emergent tool use from multi-agent autocurricula
B. Baker, I. Kanitscheider, T. Markov, Y. Wu, G. Powell, B. McGrew, and I. Mordatch · 2019
Cited alongside, same era.
Meta-learning with implicit gradients
A. Rajeswaran, C. Finn, S. M. Kakade, and S. Levine · 2019
Cited alongside, same era.
Emergent complexity and zero-shot transfer via unsupervised environment design
M. Dennis, N. Jaques, E. Vinitsky, A. Bayen, S. Russell, A. Critch, and S. Levine · 2020
Cited alongside, same era.
Motivating physical activity via competitive human-robot interaction
B. Yang, G. Habibi, P. Lancaster, B. Boots, and J. Smith · 2021
Later among the works it cites.
Control strategies for physically simulated characters performing two-player competitive sports
J. Won, D. Gopinath, and J. Hodgins · 2021
Later among the works it cites.
Stackelberg actor-critic: Game-theoretic reinforcement learning algorithms
L. Zheng, T. Fiez, Z. Alumbaugh, B. Chasnov, and L. J. Ratliff · 2021
Later among the works it cites.