Fetching the paper…
Reading the bibliography…
Zero-sum games have long guided artificial intelligence research, since they possess both a rich strategy space of best-responses and a clear evaluation metric.
Open-ended Learning in Symmetric Zero-sum Games
David Balduzzi, Marta Garnelo, Yoram Bachrach, Wojciech M. Czarnecki, Julien Pérolat, Max Jaderberg, and Thore Graepel. 2019 · 1901
Earlier work this paper cites.
The Hanabi Challenge: A New Frontier for AI Research
Nolan Bard, Jakob N. Foerster, Sarath Chandar, Neil Burch, Marc Lanctot, H. Francis Song, Emilio Parisotto, Vincent Dumoulin, Subhodeep Moitra, Edward Hughes, Iain Dunning, Shibl Mourad, Hugo Larochelle, Marc G. Bellemare, and Michael Bowling. 2019 · 1902
Earlier work this paper cites.
Learning Reciprocity in Complex Sequential Social Dilemmas
Tom Eccles, Edward Hughes, János Kramár, Steven Wheelwright, and Joel Z. Leibo. 2019 · 1903
Earlier work this paper cites.
Options as responses: Grounding behavioural hierarchies in multi-agent RL
Alexander Sasha Vezhnevets, Yuhuai Wu, Rémi Leblond, and Joel Z. Leibo. 2019 · 1906
Earlier work this paper cites.
No Press Diplomacy: Modeling Multi-Agent Gameplay
Philip Paquette, Yuchen Lu, Steven Bocco, Max O. Smith, Satya Ortiz-Gagne, Jonathan K. Kummerfeld, Satinder Singh, Joelle Pineau, and Aaron Courville. 2019 · 1909
Earlier work this paper cites.
ColosseumRL: A Framework for Multiagent Reinforcement Learning in N N -Player Games
Alexander Shmakov, John Lanier, Stephen McAleer, Rohan Achar, Cristina Lopes, and Pierre Baldi. 2019 · 1912
Earlier work this paper cites.
Zur Theorie der Gesellschaftsspiele
J. v. Neumann. 1928 · 1928
Earlier work this paper cites.
XXII. Programming a computer for playing chess
Claude E. Shannon. 1950 · 1950
Earlier work this paper cites.
Stochastic Games
L. S. Shapley. 1953 · 1953
Earlier work this paper cites.
Theory of Games and Economic Behavior / J. von Neumann, O. Morgenstern ; introd. de Harold W. Kuhn
John VON NEUMANN and Oskar Morgenstern. 1953 · 1953
Earlier work this paper cites.
Prisoner’s Dilemma: A Study in Conflict and Cooperation
A.R.A.M. Chammah, A. Rapoport, A.M. Chammah, and C.J. Orwant. 1965 · 1965
Earlier work this paper cites.
The contract net protocol: High-level communication and control in a distributed problem solver
Reid G Smith. 1980 · 1980
Earlier work this paper cites.
Evolution as a zero-sum game for energy
L Van Valen. 1980 · 1980
Earlier work this paper cites.
An automated Diplomacy player
Sarit Kraus, Daniel Lehmann, and E. Ephrati. 1989 · 1989
Earlier work this paper cites.
A world championship caliber checkers program
Jonathan Schaeffer, Joseph Culberson, Norman Treloar, Brent Knight, Paul Lu, and Duane Szafron. 1992 · 1992
Earlier work this paper cites.
Rules of encounter: designing conventions for automated negotiation among computers
Jeffrey S Rosenschein and Gilad Zlotkin. 1994 · 1994
Earlier work this paper cites.
Coalition, cryptography, and stability: Mechanisms for coalition formation in task oriented domains
Gilad Zlotkin and Jeffrey S Rosenschein. 1994 · 1994
Earlier work this paper cites.
AgenTalk: Coordination Protocol Description for Multiagent Systems.. In ICMAS
Kazuhiro Kuwabara, Toru Ishida, and Nobuyasu Osato. 1995 · 1995
Earlier work this paper cites.
Issues in automated negotiation and electronic commerce: Extending the contract net framework. In ICMAS
Tuomas Sandholm, Victor R Lesser, et al · 1995
Earlier work this paper cites.
Temporal Difference Learning and TD-Gammon
Gerald Tesauro. 1995 · 1995
Earlier work this paper cites.
Advantages of a leveled commitment contracting protocol. In AAAI/IAAI, Vol. 1
Tuomas W Sandholm and Victor R Lesser. 1996 · 1996
Earlier work this paper cites.
Negotiation and cooperation in multi-agent environments
Sarit Kraus. 1997 · 1997
Earlier work this paper cites.
Methods for task allocation via agent coalition formation
Onn Shehory and Sarit Kraus. 1998 · 1998
Earlier work this paper cites.
Coalition structure generation with worst case guarantees
Tuomas Sandholm, Kate Larson, Martin Andersson, Onn Shehory, and Fernando Tohmé. 1999 · 1999
Earlier work this paper cites.
Generating functions for computing power indices efficiently
JM Bilbao, JR Fernandez, A Jiménez Losada, and JJ Lopez. 2000 · 2000
Earlier work this paper cites.
Customer coalitions in the electronic marketplace. In AAAI/IAAI
Maksim Tsvetovat, Katia Sycara, Yian Chen, and James Ying. 2000 · 2000
Earlier work this paper cites.
Automated negotiation: prospects, methods and challenges
Nicholas R Jennings, Peyman Faratin, Alessio R Lomuscio, Simon Parsons, Michael J Wooldridge, and Carles Sierra. 2001 · 2001
Earlier work this paper cites.
Strategic negotiation in multiagent environments
Sarit Kraus and Ronald C Arkin. 2001 · 2001
Earlier work this paper cites.
Deep Blue
Murray Campbell, A.Joseph Hoane, and Feng hsiung Hsu. 2002 · 2002
Earlier work this paper cites.
Operational specification of a commitment-based agent communication language. In Proceedings of the first international joint conference on Autonomous agents and multiagent systems: part 2
Nicoletta Fornara and Marco Colombetti. 2002 · 2002
Cited alongside, same era.
Dynamic coalition formation among rational agents
Matthias Klusch and Andreas Gerber. 2002 · 2002
Cited alongside, same era.
Learning dynamics in social dilemmas
Michael W. Macy and Andreas Flache. 2002 · 2002
Cited alongside, same era.
Social intelligence, innovation, and enhanced brain size in primates
Simon M. Reader and Kevin N. Laland. 2002 · 2002
Cited alongside, same era.
Contract Theory and the Limits of Contract Law
Alan Schwartz and Robert E. Scott. 2003 · 2003
Cited alongside, same era.
Computing Shapley values, manipulating value division schemes, and checking core membership in multi-issue domains. In AAAI
Making friends on the fly: Cooperating with new teammates
Samuel Barrett, Avi Rosenfeld, Sarit Kraus, and Peter Stone. 2017 · 2016
Later among the works it cites.
Learning to Communicate with Deep Multi-Agent Reinforcement Learning
Jakob N. Foerster, Yannis M. Assael, Nando de Freitas, and Shimon Whiteson. 2016 · 2016
Later among the works it cites.
Mastering the Game of Go with Deep Neural Networks and Tree Search
David Silver, Aja Huang, Chris J. Maddison, Arthur Guez, Laurent Sifre, George van den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, Sander Dieleman, Dominik Grewe, John Nham, Nal Kalchbrenner, Ilya Sutskever, Timothy Lillicrap, Madeleine Leach, Koray Kavukcuoglu, Thore Graepel, and Demis Hassabis. 2016 · 2016
Later among the works it cites.
Thinking Fast and Slow with Deep Learning and Tree Search
Thomas Anthony, Zheng Tian, and David Barber. 2017 · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Vincent Conitzer and Tuomas Sandholm. 2004 · 2004
Cited alongside, same era.
Agent-organized networks for dynamic team formation. In Proceedings of the fourth international joint conference on Autonomous agents and multiagent systems
Matthew E Gaston and Marie DesJardins. 2005 · 2005
Cited alongside, same era.
Marginal contribution nets: a compact representation scheme for coalitional games. In Proceedings of the 6th ACM conference on Electronic commerce
Samuel Ieong and Yoav Shoham. 2005 · 2005
Cited alongside, same era.
Complexity of constructing solutions in the core based on synergies among coalitions
Vincent Conitzer and Tuomas Sandholm. 2006 · 2006
Cited alongside, same era.
Cooperation in Symmetric and Asymmetric Prisoner’s Dilemma Games
Martin Beckenkamp, Heike Hennig-Schmidt, and Frank Maier-Rigaud. 2007 · 2007
Cited alongside, same era.
Cooperative control of distributed multi-agent systems
Jeff S Shamma. 2007 · 2007
Cited alongside, same era.
Models in cooperative game theory
Rodica Branzei, Dinko Dimitrov, and Stef Tijs. 2008 · 2008
Cited alongside, same era.
Jakob N. Foerster, Richard Y. Chen, Maruan Al-Shedivat, Shimon Whiteson, Pieter Abbeel, and Igor Mordatch. 2017 · 2017
Later among the works it cites.
A Unified Game-Theoretic Approach to Multiagent Reinforcement Learning
Marc Lanctot, Vinícius Flores Zambaldi, Audrunas Gruslys, Angeliki Lazaridou, Karl Tuyls, Julien Pérolat, David Silver, and Thore Graepel. 2017 · 2017
Later among the works it cites.
Multi-agent Reinforcement Learning in Sequential Social Dilemmas
Joel Z. Leibo, Vinícius Flores Zambaldi, Marc Lanctot, Janusz Marecki, and Thore Graepel. 2017 · 2017
Later among the works it cites.
Maintaining cooperation in complex social dilemmas using deep reinforcement learning
Adam Lerer and Alexander Peysakhovich. 2017 · 2017
Later among the works it cites.
Contract Theory
David Martimort. 2017 · 2017
Later among the works it cites.
How to form winning coalitions in mixed human-computer settings. In Proceedings of the 26th international joint conference on artificial intelligence (IJCAI)
Moshe Mash, Yoram Bachrach, and Yair Zick. 2017 · 2017
Later among the works it cites.
DeepStack: Expert-Level Artificial Intelligence in No-Limit Poker
Matej Moravcík, Martin Schmid, Neil Burch, Viliam Lisý, Dustin Morrill, Nolan Bard, Trevor Davis, Kevin Waugh, Michael Johanson, and Michael H. Bowling. 2017 · 2017
Later among the works it cites.
Emergence of Grounded Compositional Language in Multi-Agent Populations
Igor Mordatch and Pieter Abbeel. 2017 · 2017
Later among the works it cites.
A multi-agent reinforcement learning model of common-pool resource appropriation
Julien Pérolat, Joel Z. Leibo, Vinícius Flores Zambaldi, Charles Beattie, Karl Tuyls, and Thore Graepel. 2017 · 2017
Later among the works it cites.
Trust in Social Dilemmas
P.A.M. van Lange, B. Rockenbach, and T. Yamagishi. 2017 · 2017
Later among the works it cites.
StarCraft II: A New Challenge for Reinforcement Learning
Oriol Vinyals, Timo Ewalds, Sergey Bartunov, Petko Georgiev, Alexander Sasha Vezhnevets, Michelle Yeo, Alireza Makhzani, Heinrich Küttler, John Agapiou, Julian Schrittwieser, John Quan, Stephen Gaffney, Stig Petersen, Karen Simonyan, Tom Schaul, Hado van Hasselt, David Silver, Timothy P. Lillicrap, Kevin Calderone, Paul Keet, Anthony Brunasso, David Lawrence, Anders Ekermo, Jacob Repp, and Rodney Tsing. 2017 · 2017
Later among the works it cites.
Towards Optimal Play of Three-Player Piglet and Pig
François Bonnet, Todd W Neller, and Simon Viennot. 2018 · 2018
Later among the works it cites.
Emergent Communication through Negotiation
Kris Cao, Angeliki Lazaridou, Marc Lanctot, Joel Z. Leibo, Karl Tuyls, and Stephen Clark. 2018 · 2018
Later among the works it cites.
IMPALA: Scalable Distributed Deep-RL with Importance Weighted Actor-Learner Architectures
Lasse Espeholt, Hubert Soyer, Rémi Munos, Karen Simonyan, Volodymyr Mnih, Tom Ward, Yotam Doron, Vlad Firoiu, Tim Harley, Iain Dunning, Shane Legg, and Koray Kavukcuoglu. 2018 · 2018
Later among the works it cites.
Inequity aversion resolves intertemporal social dilemmas
Edward Hughes, Joel Z. Leibo, Matthew G. Philips, Karl Tuyls, Edgar A. Duéñez-Guzmán, Antonio García Castañeda, Iain Dunning, Tina Zhu, Kevin R. McKee, Raphael Koster, Heather Roff, and Thore Graepel. 2018 · 2018
Later among the works it cites.
OpenAI Five
OpenAI. 2018 · 2018
Later among the works it cites.
Predicting human decision-making: From prediction to action
Ariel Rosenfeld and Sarit Kraus. 2018 · 2018
Later among the works it cites.
Reinforcement Learning: An Introduction
Richard S. Sutton and Andrew G. Barto. 2018 · 2018
Later among the works it cites.
Evolving intrinsic motivations for altruistic behavior
Jane X. Wang, Edward Hughes, Chrisantha Fernando, Wojciech M. Czarnecki, Edgar A. Duéñez-Guzmán, and Joel Z. Leibo. 2018 · 2018
Later among the works it cites.
Negotiating Team Formation Using Deep Reinforcement Learning
Yoram Bachrach, Richard Everett, Edward Hughes, Angeliki Lazaridou, Joel Leibo, Marc Lanctot, Mike Johanson, Wojtek Czarnecki, and Thore Graepel. 2019 · 2019
Later among the works it cites.
Superhuman AI for multiplayer poker
Noam Brown and Tuomas Sandholm. 2019 · 2019
Later among the works it cites.
Emergent Coordination Through Competition. In International Conference on Learning Representations
Siqi Liu, Guy Lever, Nicholas Heess, Josh Merel, Saran Tunyasuvunakool, and Thore Graepel. 2019 · 2019
Later among the works it cites.
M3RL: Mind-aware Multi-agent Management Reinforcement Learning. In International Conference on Learning Representations
Tianmin Shu and Yuandong Tian. 2019 · 2019
Later among the works it cites.
Strategic negotiations for extensive-form games
Dave Jonge and Dongmo Zhang. 2020 · 2020
Closest in time.