Fetching the paper…
Reading the bibliography…
In social psychology, Social Value Orientation (SVO) describes an individual's propensity to allocate resources between themself and others.
Asynchronous methods for deep reinforcement learning. In International conference on machine learning . PMLR, 1928–1937
Volodymyr Mnih, Adria Puigdomenech Badia, Mehdi Mirza, Alex Graves, Timothy Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu. 2016 · 1937
Earlier work this paper cites.
Stochastic games
Lloyd S Shapley. 1953 · 1953
Earlier work this paper cites.
Toward a model of interpersonal motivation in experimental games
Donald W Griesinger and James W Livingston Jr. 1973 · 1973
Earlier work this paper cites.
Prisoner’s dilemma—recollections and observations
Anatol Rapoport. 1974 · 1974
Earlier work this paper cites.
The ring measure of social values: A computerized procedure for assessing individual differences in information processing and social value orientation
Wim BG Liebrand and Charles G McClintock. 1988 · 1988
Earlier work this paper cites.
Markov games as a framework for multi-agent reinforcement learning
Michael L Littman. 1994 · 1994
Earlier work this paper cites.
Evolutionary game theory
Jörgen W Weibull. 1997 · 1997
Earlier work this paper cites.
Measuring social value orientation
Ryan O Murphy, Kurt A Ackermann, and Michel JJ Handgraaf. 2011 · 2011
Earlier work this paper cites.
Multi-agent Reinforcement Learning in Sequential Social Dilemmas. In Proceedings of the 16th Conference on Autonomous Agents and MultiAgent Systems . 464–473
Joel Z Leibo, Vinicius Zambaldi, Marc Lanctot, Janusz Marecki, and Thore Graepel. 2017 · 2017
Earlier work this paper cites.
Maintaining cooperation in complex social dilemmas using deep reinforcement learning
Adam Lerer and Alexander Peysakhovich. 2017 · 2017
Earlier work this paper cites.
Impala: Scalable distributed deep-rl with importance weighted actor-learner architectures. In International conference on machine learning . PMLR, 1407–1416
Lasse Espeholt, Hubert Soyer, Remi Munos, Karen Simonyan, Vlad Mnih, Tom Ward, Yotam Doron, Vlad Firoiu, Tim Harley, Iain Dunning, et al · 2018
Earlier work this paper cites.
Inequity aversion improves cooperation in intertemporal social dilemmas
Edward Hughes, Joel Z Leibo, Matthew Phillips, Karl Tuyls, Edgar Dueñez-Guzman, Antonio García Castañeda, Iain Dunning, Tina Zhu, Kevin McKee, Raphael Koster, et al · 2018
Cited alongside, same era.
Representation learning with contrastive predictive coding
Aaron van den Oord, Yazhe Li, and Oriol Vinyals. 2018 · 2018
Cited alongside, same era.
Consequentialist conditional cooperation in social dilemmas with imperfect information (short workshop version). In Workshops at the Thirty-Second AAAI Conference on Artificial Intelligence
Alexander Peysakhovich and Adam Lerer. 2018 · 2018
Cited alongside, same era.
Evolving intrinsic motivations for altruistic behavior
Jane X Wang, Edward Hughes, Chrisantha Fernando, Wojciech M Czarnecki, Edgar A Duéñez-Guzmán, and Joel Z Leibo. 2018 · 2018
Cited alongside, same era.
Pick Your Battles: Interaction Graphs as Population-Level Objectives for Strategic Diversity. In Proceedings of the 20th International Conference on Autonomous Agents and MultiAgent Systems . 1501–1503
Marta Garnelo, Wojciech Marian Czarnecki, Siqi Liu, Dhruva Tirumala, Junhyuk Oh, Gauthier Gidel, Hado van Hasselt, and David Balduzzi. 2021 · 2021
Later among the works it cites.
Off-belief learning. In International Conference on Machine Learning . PMLR, 4369–4379
Hengyuan Hu, Adam Lerer, Brandon Cui, Luis Pineda, Noam Brown, and Jakob Foerster. 2021 · 2021
Later among the works it cites.
Scalable evaluation of multi-agent reinforcement learning with melting pot. In International Conference on Machine Learning . PMLR, 6187–6199
Joel Z Leibo, Edgar A Dueñez-Guzman, Alexander Vezhnevets, John P Agapiou, Peter Sunehag, Raphael Koster, Jayd Matyas, Charlie Beattie, Igor Mordatch, and Thore Graepel. 2021 · 2021
Later among the works it cites.
Trajectory diversity for zero-shot coordination. In International Conference on Machine Learning . PMLR, 7204–7213
Andrei Lupu, Brandon Cui, Hengyuan Hu, and Jakob Foerster. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
David Balduzzi, Marta Garnelo, Yoram Bachrach, Wojciech Czarnecki, Julien Perolat, Max Jaderberg, and Thore Graepel. 2019 · 2019
Cited alongside, same era.
Multi-task deep reinforcement learning with popart. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 33. 3796–3803
Matteo Hessel, Hubert Soyer, Lasse Espeholt, Wojciech Czarnecki, Simon Schmitt, and Hado van Hasselt. 2019 · 2019
Cited alongside, same era.
Social behavior for autonomous vehicles
Wilko Schwarting, Alyssa Pierson, Javier Alonso-Mora, Sertac Karaman, and Daniela Rus. 2019 · 2019
Cited alongside, same era.
“Other-Play” for Zero-Shot Coordination. In International Conference on Machine Learning . PMLR, 4399–4410
Hengyuan Hu, Adam Lerer, Alex Peysakhovich, and Jakob Foerster. 2020 · 2020
Cited alongside, same era.
Social Diversity and Social Preferences in Mixed-Motive Reinforcement Learning. In Proceedings of the 19th International Conference on Autonomous Agents and MultiAgent Systems . 869–877
Kevin R McKee, Ian Gemp, Brian McWilliams, Edgar A Duèñez-Guzmán, Edward Hughes, and Joel Z Leibo. 2020 · 2020
Cited alongside, same era.
Adaptable agent populations via a generative model of policies
Kenneth Derek and Phillip Isola. 2021 · 2021
Cited alongside, same era.
Statistical discrimination in learning agents
Edgar A Duéñez-Guzmán, Kevin R McKee, Yiran Mao, Ben Coppin, Silvia Chiappa, Alexander Sasha Vezhnevets, Michiel A Bakker, Yoram Bachrach, Suzanne Sadedin, William Isaac, et al · 2021
Cited alongside, same era.
Warmth and Competence in Human-Agent Cooperation. In Proceedings of the 21st International Conference on Autonomous Agents and Multiagent Systems . 898–907
Kevin R McKee, Xuechunzi Bai, and Susan T Fiske. 2022a
Cited in the paper.
Modelling behavioural diversity for learning in open-ended games. In International Conference on Machine Learning . PMLR, 8514–8524
Nicolas Perez-Nieves, Yaodong Yang, Oliver Slumbers, David H Mguni, Ying Wen, and Jun Wang. 2021 · 2021
Later among the works it cites.
Collaborating with humans without human data
DJ Strouse, Kevin McKee, Matt Botvinick, Edward Hughes, and Richard Everett. 2021 · 2021
Later among the works it cites.
Discovering Diverse Multi-Agent Strategic Behavior via Reward Randomization. In International Conference on Learning Representations
Zhenggang Tang, Chao Yu, Boyuan Chen, Huazhe Xu, Xiaolong Wang, Fei Fang, Simon Shaolei Du, Yu Wang, and Yi Wu. 2021 · 2021
Later among the works it cites.
John P Agapiou, Alexander Sasha Vezhnevets, Edgar A Duéñez-Guzmán, Jayd Matyas, Yiran Mao, Peter Sunehag, Raphael Köster, Udari Madhushani, Kavya Kopparapu, Ramona Comanescu, et al · 2022
Later among the works it cites.
Quantifying the effects of environment and population diversity in multi-agent reinforcement learning
Kevin R McKee, Joel Z Leibo, Charlie Beattie, and Richard Everett. 2022b · 2022
Later among the works it cites.
Discovering policies with domino: Diversity optimization maintaining near optimality
Tom Zahavy, Yannick Schroecker, Feryal Behbahani, Kate Baumli, Sebastian Flennerhag, Shaobo Hou, and Satinder Singh. 2022 · 2022
Later among the works it cites.