Prosocial learning agents solve generalized Stag Hunts better than selfish ones
Original
Alexander Peysakhovich and Adam Lerer · 2017
Later among the works it cites.
Multi-agent Reinforcement Learning in Sequential Social Dilemmas
Original
Joel Z. Leibo, Vinicius Zambaldi, Marc Lanctot, Janusz Marecki, and Thore Graepel · 2017
Later among the works it cites.
Multi-view decision processes: The helper-ai problem
Christos Dimitrakakis, David C. Parkes, Goran Radanovic, and Paul Tylkin · 2017
Later among the works it cites.
Reinforcement mechanism design
Pingzhong Tang · 2017
Later among the works it cites.
The positronic economist: A computational system for analyzing economic mechanisms
David R. M. Thompson, Neil Newman, and Kevin Leyton-Brown · 2017
Later among the works it cites.
Proximal policy optimization algorithms
Original
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Later among the works it cites.
Intergenerational mobility and preferences for redistribution
Alberto Alesina, Stefanie Stantcheva, and Edoardo Teso · 2018
Later among the works it cites.
Theory and Agent-based Modeling of Taxpayer Preference and Behavior
Shree Krishna Subburaj and Shrisha Rao · 2018
Later among the works it cites.
Human-level performance in first-person multiplayer games with population-based deep reinforcement learning
Original
Max Jaderberg, Wojciech M. Czarnecki, Iain Dunning, Luke Marris, Guy Lever, Antonio Garcia Castaneda, Charles Beattie, Neil C. Rabinowitz, Ari S. Morcos, Avraham Ruderman, Nicolas Sonnerat, Tim Green, Louise Deason, Joel Z. Leibo, David Silver, Demis Hassabis, Koray Kavukcuoglu, and Thore Graepel · 2018
Later among the works it cites.
Openai five
OpenAI · 2018
Later among the works it cites.
Relational Forward Models for Multi-Agent Learning
Original
Andrea Tacchetti, H. Francis Song, Pedro A. M. Mediano, Vinicius Zambaldi, Neil C. Rabinowitz, Thore Graepel, Matthew Botvinick, and Peter W. Battaglia · 2018
Later among the works it cites.
M^3rl: Mind-aware Multi-agent Management Reinforcement Learning
Original
Tianmin Shu and Yuandong Tian · 2018
Later among the works it cites.
Inequity aversion improves cooperation in intertemporal social dilemmas
Original
Edward Hughes, Joel Z. Leibo, Matthew G. Phillips, Karl Tuyls, Edgar A. Duéñez Guzmán, Antonio García Castañeda, Iain Dunning, Tina Zhu, Kevin R. McKee, Raphael Koster, Heather Roff, and Thore Graepel · 2018
Later among the works it cites.
Stable Opponent Shaping in Differentiable Games
Original
Alistair Letcher, Jakob Foerster, David Balduzzi, Tim Rocktäschel, and Shimon Whiteson · 2018
Later among the works it cites.
The Mechanics of n-Player Differentiable Games
Original
David Balduzzi, Sebastien Racaniere, James Martens, Jakob Foerster, Karl Tuyls, and Thore Graepel · 2018
Later among the works it cites.
Social Influence as Intrinsic Motivation for Multi-Agent Deep Reinforcement Learning
Original
Natasha Jaques, Angeliki Lazaridou, Edward Hughes, Caglar Gulcehre, Pedro A. Ortega, D. J. Strouse, Joel Z. Leibo, and Nando de Freitas · 2018
Later among the works it cites.
Deep learning for revenue-optimal auctions with budgets
Z. Feng, H. Narasimhan, and D. C. Parkes · 2018
Later among the works it cites.
Deep learning for multi-facility location mechanism design
N. Golowich, H. Narasimhan, and D. C. Parkes · 2018
Later among the works it cites.
The sample complexity of up-to-epsilon multi-dimensional revenue maximization
Y. A. Gonczarowski and S. M. Weinberg · 2018
Later among the works it cites.
Designing core-selecting payment rules: A computational search approach
Benedikt Bünz, Benjamin Lubin, and Sven Seuken · 2018
Later among the works it cites.
RLlib: Abstractions for Distributed Reinforcement Learning
Eric Liang, Richard Liaw, Philipp Moritz, Robert Nishihara, Roy Fox, Ken Goldberg, Joseph E. Gonzalez, Michael I. Jordan, and Ion Stoica · 2018
Later among the works it cites.
Optimal Auctions through Deep Learning
Paul Dütting, Zhe Feng, Harikrishna Narasimhan, David C. Parkes, and Sai Srivatsa Ravindranath · 2019
Later among the works it cites.
A neural architecture for designing truthful and efficient auctions
Original
A. Tacchetti, D.J. Strouse, M. Garnelo, T. Graepel, and Y. Bachrach · 2019
Later among the works it cites.
Automated mechanism design via neural networks
W. Shen, P. Tang, and S. Zuo · 2019
Later among the works it cites.
Deep reinforcement learning for green security games with real-time information
Yufei Wang, Zheyuan Ryan Shi, Lantao Yu, Yi Wu, Rohit Singh, Lucas Joppa, and Fei Fang · 2019
Later among the works it cites.
On the utility of learning about humans for human-ai coordination
Micah Carroll, Rohin Shah, Mark K. Ho, Tom Griffiths, Sanjit A. Seshia, Pieter Abbeel, and Anca D. Dragan · 2019
Later among the works it cites.
Open-ended learning in symmetric zero-sum games
David Balduzzi, Marta Garnelo, Yoram Bachrach, Wojciech Czarnecki, Julien Pérolat, Max Jaderberg, and Thore Graepel · 2019
Later among the works it cites.
Multi-player Atari: Designing Helper-AIs in Rich Environments
Paul Tylkin, David C. Parkes, and Goran Radanovic · 2020
Closest in time.
Reinforcement mechanism design, with applications to dynamic pricing in sponsored search auctions
Weiran Shen, Binghui Peng, Hanpeng Liu, Michael Zhang, Ruohan Qian, Yan Hong, Zhi Guo, Zongyao Ding, Pengjun Lu, and Pingzhong Tang · 2020
Closest in time.