Fetching the paper…
Reading the bibliography…
Multi-agent artificial intelligence research promises a path to develop intelligent technologies that are more human-like and more human-compatible than those produced by "solipsistic" approaches, which do not consider interactions between agents.
The Evolution of Cooperation
R. Axelrod · 1984
Earlier work this paper cites.
Games and decisions: Introduction and critical survey
R. D. Luce and H. Raiffa · 1989
Earlier work this paper cites.
Evolutionary game theory
J. W. Weibull · 1997
Earlier work this paper cites.
The dynamics of reinforcement learning in cooperative multiagent systems
C. Claus and C. Boutilier · 1998
Earlier work this paper cites.
Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning
R. S. Sutton, D. Precup, and S. Singh · 1999
Earlier work this paper cites.
The elements of statistical learning: data mining, inference, and prediction , volume 2
T. Hastie, R. Tibshirani, J. H. Friedman, and J. H. Friedman · 2009
Earlier work this paper cites.
Understanding institutional diversity
E. Ostrom · 2009
Earlier work this paper cites.
Meta-analysis in medical research
A.-B. Haidich · 2010
Earlier work this paper cites.
Lab experiments for the study of social-ecological systems
M. A. Janssen, R. Holahan, A. Lee, and E. Ostrom · 2010
Earlier work this paper cites.
Horde: A scalable real-time architecture for learning knowledge from unsupervised sensorimotor interaction
R. S. Sutton, J. Modayil, M. Delp, T. Degris, P. M. Pilarski, A. White, and D. Precup · 2011
Earlier work this paper cites.
A treatise of human nature
D. Hume · 2012
Earlier work this paper cites.
On the difficulty of training recurrent neural networks
R. Pascanu, T. Mikolov, and Y. Bengio · 2013
Earlier work this paper cites.
Aligning superintelligence with human interests: A technical research agenda
N. Soares and B. Fallenstein · 2014
Earlier work this paper cites.
Imagenet large scale visual recognition challenge
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. Bernstein, et al · 2015
Earlier work this paper cites.
Research priorities for robust and beneficial artificial intelligence
S. Russell, D. Dewey, and M. Tegmark · 2015
Earlier work this paper cites.
Concrete problems in AI safety
D. Amodei, C. Olah, J. Steinhardt, P. Christiano, J. Schulman, and D. Mané · 2016
Earlier work this paper cites.
Cooperative inverse reinforcement learning
D. Hadfield-Menell, S. J. Russell, P. Abbeel, and A. Dragan · 2016
Earlier work this paper cites.
Reinforcement learning with unsupervised auxiliary tasks
M. Jaderberg, V. Mnih, W. M. Czarnecki, T. Schaul, J. Z. Leibo, D. Silver, and K. Kavukcuoglu · 2016
Earlier work this paper cites.
Asynchronous methods for deep reinforcement learning
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu · 2016
Earlier work this paper cites.
Mastering the game of Go with deep neural networks and tree search
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot, et al · 2016
Earlier work this paper cites.
Evaluation in artificial intelligence: from task-oriented to ability-oriented measurement
J. Hernández-Orallo · 2017
Earlier work this paper cites.
Six challenges for neural machine translation
P. Koehn and R. Knowles · 2017
Earlier work this paper cites.
A unified game-theoretic approach to multiagent reinforcement learning
M. Lanctot, V. Zambaldi, A. Gruslys, A. Lazaridou, K. Tuyls, J. Pérolat, D. Silver, and T. Graepel · 2017
Earlier work this paper cites.
Multi-agent Reinforcement Learning in Sequential Social Dilemmas
J. Z. Leibo, V. Zambaldi, M. Lanctot, J. Marecki, and T. Graepel · 2017
Earlier work this paper cites.
Maintaining cooperation in complex social dilemmas using deep reinforcement learning
A. Lerer and A. Peysakhovich · 2017
Earlier work this paper cites.
Multi-agent actor-critic for mixed cooperative-competitive environments
R. Lowe, Y. Wu, A. Tamar, J. Harb, P. Abbeel, and I. Mordatch · 2017
Earlier work this paper cites.
A multi-agent reinforcement learning model of common-pool resource appropriation
J. Perolat, J. Z. Leibo, V. Zambaldi, C. Beattie, K. Tuyls, and T. Graepel · 2017
Earlier work this paper cites.
Prosocial learning agents solve generalized stag hunts better than selfish ones
A. Peysakhovich and A. Lerer · 2017
Cited alongside, same era.
Mastering the game of go without human knowledge
D. Silver, J. Schrittwieser, K. Simonyan, I. Antonoglou, A. Huang, A. Guez, T. Hubert, L. Baker, M. Lai, A. Bolton, et al · 2017
Cited alongside, same era.
Feudal networks for hierarchical reinforcement learning
A. S. Vezhnevets, S. Osindero, T. Schaul, N. Heess, M. Jaderberg, D. Silver, and K. Kavukcuoglu · 2017
Cited alongside, same era.
Impala: Scalable distributed deep-rl with importance weighted actor-learner architectures
L. Espeholt, H. Soyer, R. Munos, K. Simonyan, V. Mnih, T. Ward, Y. Doron, V. Firoiu, T. Harley, I. Dunning, et al · 2018
Cited alongside, same era.
Generalization and regularization in DQN
J. Farebrother, M. C. Machado, and M. Bowling · 2018
Cited alongside, same era.
A meta-analysis of overfitting in machine learning
R. Roelofs, V. Shankar, B. Recht, S. Fridovich-Keil, M. Hardt, J. Miller, and L. Schmidt · 2019
Later among the works it cites.
Human compatible: Artificial intelligence and the problem of control
S. Russell · 2019
Later among the works it cites.
Grandmaster level in starcraft II using multi-agent reinforcement learning
O. Vinyals, I. Babuschkin, W. M. Czarnecki, M. Mathieu, A. Dudzik, J. Chung, D. H. Choi, R. Powell, T. Ewalds, P. Georgiev, et al · 2019
Later among the works it cites.
Open problems in cooperative AI
A. Dafoe, E. Hughes, Y. Bachrach, T. Collins, K. R. McKee, J. Z. Leibo, K. Larson, and T. Graepel · 2020
Later among the works it cites.
Model-free conventions in multi-agent reinforcement learning with heterogeneous preferences
R. Köster, K. R. McKee, R. Everett, L. Weidinger, W. S. Isaac, E. Hughes, E. A. Duéñez-Guzmán, T. Graepel, M. Botvinick, and J. Z. Leibo · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Learning with opponent-learning awareness
J. Foerster, R. Y. Chen, M. Al-Shedivat, S. Whiteson, P. Abbeel, and I. Mordatch · 2018
Cited alongside, same era.
Inequity aversion improves cooperation in intertemporal social dilemmas
E. Hughes, J. Z. Leibo, M. G. Philips, K. Tuyls, E. A. Duéñez-Guzmán, A. G. Castañeda, I. Dunning, T. Zhu, K. R. McKee, R. Koster, H. Roff, and T. Graepel · 2018
Cited alongside, same era.
Representation learning with contrastive predictive coding
A. v. d. Oord, Y. Li, and O. Vinyals · 2018
Cited alongside, same era.
Qmix: Monotonic value function factorisation for deep multi-agent reinforcement learning
T. Rashid, M. Samvelyan, C. Schroeder, G. Farquhar, J. Foerster, and S. Whiteson · 2018
Cited alongside, same era.
A general reinforcement learning algorithm that masters chess, shogi, and go through self-play
D. Silver, T. Hubert, J. Schrittwieser, I. Antonoglou, M. Lai, A. Guez, M. Lanctot, L. Sifre, D. Kumaran, T. Graepel, et al · 2018
Cited alongside, same era.
Value-decomposition networks for cooperative multi-agent learning based on team reward
P. Sunehag, G. Lever, A. Gruslys, W. M. Czarnecki, V. Zambaldi, M. Jaderberg, M. Lanctot, N. Sonnerat, J. Z. Leibo, K. Tuyls, et al · 2018
Cited alongside, same era.
Reinforcement learning: An introduction
R. S. Sutton and A. G. Barto · 2018
Cited alongside, same era.
Later among the works it cites.
Tracking the impact and evolution of ai: The aicollaboratory
F. Martınez-Plumed, J. Hernández-Orallo, and E. Gómez · 2020
Later among the works it cites.
Social diversity and social preferences in mixed-motive reinforcement learning
K. R. McKee, I. Gemp, B. McWilliams, E. A. Duèñez-Guzmán, E. Hughes, and J. Z. Leibo · 2020
Later among the works it cites.
V-mpo: On-policy maximum a posteriori policy optimization for discrete and continuous control
H. F. Song, A. Abdolmaleki, J. T. Springenberg, A. Clark, H. Soyer, J. W. Rae, S. Noury, A. Ahuja, S. Liu, D. Tirumala, N. Heess, D. Belov, M. Riedmiller, and M. M. Botvinick · 2020
Later among the works it cites.
Options as responses: Grounding behavioural hierarchies in multi-agent reinforcement learning
A. Vezhnevets, Y. Wu, M. Eckstein, R. Leblond, and J. Z. Leibo · 2020
Later among the works it cites.
Too many cooks: Coordinating multi-agent collaboration through inverse planning
R. E. Wang, S. A. Wu, J. A. Evans, J. B. Tenenbaum, D. C. Parkes, and M. Kleiman-Weiner · 2020
Later among the works it cites.
Towards playing full moba games with deep reinforcement learning
D. Ye, G. Chen, W. Zhang, S. Chen, B. Yuan, B. Liu, J. Chen, Z. Liu, F. Qiu, H. Yu, et al · 2020
Later among the works it cites.
Harms of AI
D. Acemoglu · 2021
Later among the works it cites.
Modelling cooperation in network games with spatio-temporal complexity
M. A. Bakker, R. Everett, L. Weidinger, I. Gabriel, W. S. Isaac, J. Z. Leibo, and E. Hughes · 2021
Later among the works it cites.
Statistical discrimination in learning agents
E. A. Duéñez-Guzmán, K. R. McKee, Y. Mao, B. Coppin, S. Chiappa, A. S. Vezhnevets, M. A. Bakker, Y. Bachrach, S. Sadedin, W. Isaac, et al · 2021
Later among the works it cites.
The origins and psychology of human cooperation
J. Henrich and M. Muthukrishna · 2021
Later among the works it cites.
Podracer architectures for scalable reinforcement learning
M. Hessel, M. Kroiss, A. Clark, I. Kemaev, J. Quan, T. Keck, F. Viola, and H. van Hasselt · 2021
Later among the works it cites.
Scalable evaluation of multi-agent reinforcement learning with Melting Pot
J. Z. Leibo, E. A. Dueñez-Guzman, A. Vezhnevets, J. P. Agapiou, P. Sunehag, R. Koster, J. Matyas, C. Beattie, I. Mordatch, and T. Graepel · 2021
Later among the works it cites.
A survey on bias and fairness in machine learning
N. Mehrabi, F. Morstatter, N. Saxena, K. Lerman, and A. Galstyan · 2021
Later among the works it cites.
Collaborating with humans without human data
D. Strouse, K. McKee, M. Botvinick, E. Hughes, and R. Everett · 2021
Later among the works it cites.
Agent-based computational economics: Overview and brief history
L. Tesfatsion · 2021
Later among the works it cites.
Ethical and social risks of harm from language models
L. Weidinger, J. Mellor, M. Rauh, C. Griffin, J. Uesato, P.-S. Huang, M. Cheng, M. Glaese, B. Balle, A. Kasirzadeh, et al · 2021
Later among the works it cites.
Emergent bartering behaviour in multi-agent reinforcement learning
M. B. Johanson, E. Hughes, F. Timbers, and J. Z. Leibo · 2022
Closest in time.
Spurious normativity enhances learning of compliance and enforcement behavior in artificial agents
R. Köster, D. Hadfield-Menell, R. Everett, L. Weidinger, G. K. Hadfield, and J. Z. Leibo · 2022
Closest in time.
Mapping global dynamics of benchmark creation and saturation in artificial intelligence
S. Ott, A. Barbosa-Silva, K. Blagec, J. Brauner, and M. Samwald · 2022
Closest in time.
Rethink reporting of evaluation results in ai
R. Burnell, W. Schellaert, J. Burden, T. D. Ullman, F. Martinez-Plumed, J. B. Tenenbaum, D. Rutar, L. G. Cheke, J. Sohl-Dickstein, M. Mitchell, D. Kiela, M. Shanahan, E. M. Voorhees, A. G. Cohn, J. Z. Leibo, and J. Hernandez-Orallo · 2023
Closest in time.
A learning agent that acquires social norms from public sanctions in decentralized multi-agent settings
E. Vinitsky, R. Köster, J. P. Agapiou, E. A. Duéñez-Guzmán, A. S. Vezhnevets, and J. Z. Leibo · 2023
Closest in time.