Fetching the paper…
Reading the bibliography…
The real world is awash with multi-agent problems that require collective action by self-interested agents, from the routing of packets across a computer network to the management of irrigation systems.
Nicolas Anastassacos, Steve Hailes, and Mirco Musolesi. 2019 · 1902
Earlier work this paper cites.
Learning Reciprocity in Complex Sequential Social Dilemmas
Tom Eccles, Edward Hughes, János Kramár, Steven Wheelwright, and Joel Z. Leibo. 2019 · 1903
Earlier work this paper cites.
A Neural Architecture for Designing Truthful and Efficient Auctions
Andrea Tacchetti, DJ Strouse, Marta Garnelo, Thore Graepel, and Yoram Bachrach. 2019 · 1907
Earlier work this paper cites.
Asynchronous methods for deep reinforcement learning. In International conference on machine learning . 1928–1937
Volodymyr Mnih, Adria Puigdomenech Badia, Mehdi Mirza, Alex Graves, Timothy Lillicrap, Tim Harley, David Silver, and Koray Kavukcuoglu. 2016 · 1937
Earlier work this paper cites.
Optimising Worlds to Evaluate and Influence Reinforcement Learning Agents. In Proceedings of the 18th International Conference on Autonomous Agents and MultiAgent Systems (Montreal QC, Canada) (AAMAS ’19) . International Foundation for Autonomous Agents and Multiagent Systems, Richland, SC, 1943–1945
Richard Everett, Adam Cobb, Andrew Markham, and Stephen Roberts. 2019 · 1945
Earlier work this paper cites.
Notes on the n-Person Game—II: The Value of an n-Person Game
Lloyd S Shapley. 1951 · 1951
Earlier work this paper cites.
Industrial Dynamics. A major breakthrough for decision makers
Jay W Forrester. 1958 · 1958
Earlier work this paper cites.
Prisoner’s Dilemma: A Study in Conflict and Cooperation
A.R.A.M. Chammah, A. Rapoport, A.M. Chammah, and C.J. Orwant. 1965 · 1965
Earlier work this paper cites.
The Evolution of Reciprocal Altruism
Robert Trivers. 1971 · 1971
Earlier work this paper cites.
A simple expression for the Shapley value in a special case
Stephen C Littlechild and Guillermo Owen. 1973 · 1973
Earlier work this paper cites.
The evolution of cooperation
R Axelrod and WD Hamilton. 1981 · 1981
Earlier work this paper cites.
Governing the Commons: The Evolution of Institutions for Collective Action
Elinor Ostrom. 1990 · 1990
Earlier work this paper cites.
Moral Boundaries: A Political Argument for an Ethic of Care
Joan C. Tronto. 1993 · 1993
Earlier work this paper cites.
A Theory of Fairness, Competition, and Cooperation
Ernst Fehr and Klaus M. Schmidt. 1999 · 1999
Earlier work this paper cites.
Complexity of mechanism design. In Proceedings of the Eighteenth conference on Uncertainty in artificial intelligence . 103–110
Vincent Conitzer and Tuomas Sandholm. 2002 · 2002
Earlier work this paper cites.
What Really Matters in Auction Design
Paul Klemperer. 2002 · 2002
Earlier work this paper cites.
Global supply chain management: A reinforcement learning approach
Pierpaolo Pontrandolfo, Abhijit Gosavi, O. Okogbaa, and Tapas Das. 2002 · 2002
Earlier work this paper cites.
Collective intelligence, data routing and braess’ paradox
David H Wolpert and Kagan Tumer. 2002 · 2002
Earlier work this paper cites.
WES: Agent-based User Interaction Simulation on Real Infrastructure
John Ahlgren, Maria Eugenia Berezin, Kinga Bojarczuk, Elena Dulskyte, Inna Dvortsova, Johann George, Natalija Gucevska, Mark Harman, Ralf Lämmel, Erik Meijer, Silvia Sapora, and Justin Spahr-Summers. 2020 · 2004
Earlier work this paper cites.
The AI Economist: Improving Equality and Productivity with AI-Driven Tax Policies
Stephan Zheng, Alexander Trott, Sunil Srinivasa, Nikhil Naik, Melvin Gruesbeck, David C. Parkes, and Richard Socher. 2020 · 2004
Cited alongside, same era.
Mechanism design via machine learning. In 46th Annual IEEE Symposium on Foundations of Computer Science (FOCS’05) . 605–614
M. . Balcan, A. Blum, J. D. Hartline, and Y. Mansour. 2005 · 2005
Cited alongside, same era.
The ethics of care: Personal, political, and global
Virginia Held et al · 2006
Cited alongside, same era.
Methods for empirical game-theoretic analysis. In proceedings of the 21st national conference on Artificial intelligence-Volume 2 . 1552–1555
Michael P Wellman. 2006 · 2006
Cited alongside, same era.
Direct reciprocity on graphs
Hisashi Ohtsuki and Martin A Nowak. 2007 · 2007
Cited alongside, same era.
Consequentialist conditional cooperation in social dilemmas with imperfect information
Alexander Peysakhovich and Adam Lerer. 2017 · 2017
Later among the works it cites.
Reinforcement mechanism design. In Proceedings of the 26th International Joint Conference on Artificial Intelligence . 5146–5150
Pingzhong Tang. 2017 · 2017
Later among the works it cites.
Adaptive Mechanism Design: Learning to Promote Cooperation
Tobias Baumann, Thore Graepel, and John Shawe-Taylor. 2018 · 2018
Later among the works it cites.
IMPALA: Scalable Distributed Deep-RL with Importance Weighted Actor-Learner Architectures. In International Conference on Machine Learning . 1407–1416
Lasse Espeholt, Hubert Soyer, Remi Munos, Karen Simonyan, Vlad Mnih, Tom Ward, Yotam Doron, Vlad Firoiu, Tim Harley, Iain Dunning, et al · 2018
Later among the works it cites.
Inequity aversion improves cooperation in intertemporal social dilemmas
Edward Hughes, Joel Z Leibo, Matthew Phillips, Karl Tuyls, Edgar Dueñez-Guzman, Antonio García Castañeda, Iain Dunning, Tina Zhu, Kevin McKee, Raphael Koster, et al · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Mechanism design for a multicommodity flow game in service network alliances
Richa Agarwal and Özlem Ergun. 2008 · 2008
Cited alongside, same era.
A General Framework for Analyzing Sustainability of Social-Ecological Systems
Elinor Ostrom. 2009 · 2009
Cited alongside, same era.
A theory of justice
John Rawls. 2009 · 2009
Cited alongside, same era.
The Shapley value for airport and irrigation games
Judit Márkus, Péter Miklós Pintér, and Anna Radványi. 2012 · 2012
Cited alongside, same era.
Mechanism theory
Matthew O Jackson. 2014 · 2014
Cited alongside, same era.
Chapter 3 - Games on Networks
Matthew O. Jackson and Yves Zenou. 2015 · 2015
Cited alongside, same era.
Optimal Auctions through Deep Learning
Paul Dütting, Zhe Feng, Harikrishna Narasimhan, and David C. Parkes. 2017 · 2017
Cited alongside, same era.
Later among the works it cites.
Procedural Level Generation Improves Generality of Deep Reinforcement Learning
Niels Justesen, Ruben Rodriguez Torrado, Philip Bontrager, Ahmed Khalifa, Julian Togelius, and Sebastian Risi. 2018 · 2018
Later among the works it cites.
Reinforcement learning for supply chain optimization. In European Workshop on Reinforcement Learning 14 . 1–9
Lukas Kemmer, Henrik von Kleist, Diego de Rochebouët, Nikolaos Tziortziotis, and Jesse Read. 2018 · 2018
Later among the works it cites.
M3̂RL: Mind-aware Multi-agent Management Reinforcement Learning
Tianmin Shu and Yuandong Tian. 2018 · 2018
Later among the works it cites.
Emergent tool use from multi-agent autocurricula
Bowen Baker, Ingmar Kanitscheider, Todor Markov, Yi Wu, Glenn Powell, Bob McGrew, and Igor Mordatch. 2019 · 2019
Later among the works it cites.
Human-level performance in 3D multiplayer games with population-based reinforcement learning
Max Jaderberg, Wojciech M Czarnecki, Iain Dunning, Luke Marris, Guy Lever, Antonio Garcia Castaneda, Charles Beattie, Neil C Rabinowitz, Ari S Morcos, Avraham Ruderman, et al · 2019
Later among the works it cites.
Social influence as intrinsic motivation for multi-agent deep reinforcement learning. In International Conference on Machine Learning . PMLR, 3040–3049
Natasha Jaques, Angeliki Lazaridou, Edward Hughes, Caglar Gulcehre, Pedro Ortega, DJ Strouse, Joel Z Leibo, and Nando De Freitas. 2019 · 2019
Later among the works it cites.
Introduction to mechanism design and implementation
Eric Maskin. 2019 · 2019
Later among the works it cites.
Learning to teach in cooperative multiagent reinforcement learning. In Proceedings of the AAAI Conference on Artificial Intelligence , Vol. 33. 6128–6136
Shayegan Omidshafiei, Dong-Ki Kim, Miao Liu, Gerald Tesauro, Matthew Riemer, Christopher Amato, Murray Campbell, and Jonathan P How. 2019 · 2019
Later among the works it cites.
Social behavior for autonomous vehicles
Wilko Schwarting, Alyssa Pierson, Javier Alonso-Mora, Sertac Karaman, and Daniela Rus. 2019 · 2019
Later among the works it cites.
Artificial Intelligence, Values, and Alignment
Iason Gabriel. 2020 · 2020
Later among the works it cites.
Gifting in Multi-Agent Reinforcement Learning. In Proceedings of the 19th Conference on Autonomous Agents and MultiAgent Systems . International Foundation for Autonomous Agents and Multiagent Systems, 789–797
Andrei Lupu and Doina Precup. 2020 · 2020
Later among the works it cites.
Learning to Incentivize Other Learning Agents. In AAMAS 2020 - Adaptive and Learning Agents (ALA) Workshop
Jiachen Yang, Ang Li, Mehrdad Farajtabar, Peter Sunehag, Edward Hughes, and Hongyuan Zha. 2020 · 2020
Later among the works it cites.
Value-Decomposition Networks For Cooperative Multi-Agent Learning Based On Team Reward. In Proceedings of the 17th International Conference on Autonomous Agents and MultiAgent Systems . 2085–2087
Peter Sunehag, Guy Lever, Audrunas Gruslys, Wojciech Marian Czarnecki, Vinicius Zambaldi, Max Jaderberg, Marc Lanctot, Nicolas Sonnerat, Joel Z Leibo, Karl Tuyls, et al · 2087
Closest in time.