Fetching the paper…
Reading the bibliography…
Reinforcement learning solutions have great success in the 2-player general sum setting.
An experimental study of n-person iterated prisoner’s dilemma games
Xin Yao and Paul Darwen · 2001
Earlier work this paper cites.
Evolutionary dynamics of collective action in n-person stag hunt dilemmas
Jorge Pacheco, Francisco Santos, and Max Souza · 2008
Earlier work this paper cites.
Evolution strategies as a scalable alternative to reinforcement learning, 2017
Tim Salimans, Jonathan Ho, Xi Chen, Szymon Sidor, and Ilya Sutskever · 2017
Earlier work this paper cites.
Proximal policy optimization algorithms, 2017
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Earlier work this paper cites.
Prisoner’s Dilemma
Steven Kuhn · 2019
Earlier work this paper cites.
Stable opponent shaping in differentiable games
Alistair Letcher, Jakob N. Foerster, David Balduzzi, Tim Rocktäschel, and Shimon Whiteson · 2019
Earlier work this paper cites.
A policy gradient algorithm for learning to learn in multiagent reinforcement learning
Dong-Ki Kim, Miao Liu, Matthew Riemer, Chuangchuang Sun, Marwa Abdulhai, Golnaz Habibi, Sebastian Lopez-Cot, Gerald Tesauro, and Jonathan P. How · 2021
Cited alongside, same era.
Replicator dynamics of an n-player snowdrift game with delayed payoffs
Thomas A. Wettergren · 2021
Cited alongside, same era.
Model-free opponent shaping
Christopher Lu, Timon Willi, Christian A. Schröder de Witt, and Jakob N. Foerster · 2022
Cited alongside, same era.
COLA: consistent learning with opponent-learning awareness
Timon Willi, Alistair Letcher, Johannes Treutlein, and Jakob N. Foerster · 2022
Cited alongside, same era.
Proximal learning with opponent-learning awareness
Stephen Zhao, Chris Lu, Roger B Grosse, and Jakob Foerster · 2022
Cited alongside, same era.
Learning with opponent-learning awareness, 2018a
Jakob N. Foerster, Richard Y. Chen, Maruan Al-Shedivat, Shimon Whiteson, Pieter Abbeel, and Igor Mordatch
Analyzing the sample complexity of model-free opponent shaping
Kitty Fung, Qizhen Zhang, Chris Lu, Timon Willi, and Jakob Nicolaus Foerster · 2023
Closest in time.
Scaling opponent shaping to high dimensional games
Akbir Khan, Timon Willi, Newton Kwan, Andrea Tacchetti, Chris Lu, Edward Grefenstette, Tim Rocktäschel, and Jakob Foerster · 2023
Closest in time.
Adversarial cheap talk
Chris Lu, Timon Willi, Alistair Letcher, and Jakob Nicolaus Foerster · 2023
Closest in time.
Jaxmarl: Multi-agent rl environments in jax, 2023
Alexander Rutherford, Benjamin Ellis, Matteo Gallici, Jonathan Cook, Andrei Lupu, Gardar Ingvarsson, Timon Willi, Akbir Khan, Christian Schroeder de Witt, Alexandra Souly, Saptarashmi Bandyopadhyay, Mikayel Samvelyan, Minqi Jiang, Robert Tjarko Lange, Shimon Whiteson, Bruno Lacerda, Nick Hawes, Tim Rocktaschel, Chris Lu, and Jakob Nicolaus Foerster · 2023
Closest in time.
Pax: Multi-agent learning in jax
Timon Willi, Akbir Khan, Newton Kwan, Mikayel Samvelyan, Chris Lu, and Jakob Foerster · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited in the paper.
Dice: The infinitely differentiable monte carlo estimator
Jakob N. Foerster, Gregory Farquhar, Maruan Al-Shedivat, Tim Rocktäschel, Eric P. Xing, and Shimon Whiteson
Cited in the paper.