Fetching the paper…
Reading the bibliography…
Problems of cooperation--in which agents seek ways to jointly improve their welfare--are ubiquitous and important.
Nicolas Anastassacos, Steve Hailes, and Mirco Musolesi · 1902
Earlier work this paper cites.
Joel Z. Leibo, E. Hughes, Marc Lanctot, and Thore Graepel · 1903
Earlier work this paper cites.
Anti-efficient encoding in emergent communication
Rahma Chaabouni, Eugene Kharitonov, Emmanuel Dupoux, and Marco Baroni · 1905
Earlier work this paper cites.
A survey on deep learning based brain computer interface: Recent advances and new frontiers
Xiang Zhang, Lina Yao, Xianzhi Wang, Jessica Monaghan, David Mcalpine, and Yu Zhang · 1905
Earlier work this paper cites.
A neural architecture for designing truthful and efficient auctions
Andrea Tacchetti, DJ Strouse, Marta Garnelo, Thore Graepel, and Yoram Bachrach · 1907
Earlier work this paper cites.
Emergent tool use from multi-agent autocurricula
Bowen Baker, Ingmar Kanitscheider, Todor Markov, Yi Wu, Glenn Powell, Bob McGrew, and Igor Mordatch · 1909
Earlier work this paper cites.
No press diplomacy: Modeling multi-agent gameplay
Philip Paquette, Yuchen Lu, Steven Bocco, Max O. Smith, Satya Ortiz-Gagne, Jonathan K. Kummerfeld, Satinder Singh, Joelle Pineau, and Aaron Courville · 1909
Earlier work this paper cites.
Dota 2 with large scale deep reinforcement learning
Christopher Berner, Greg Brockman, Brooke Chan, Vicki Cheung, Przemysław Dębiak, Christy Dennison, David Farhi, Quirin Fischer, Shariq Hashme, Chris Hesse, et al · 1912
Earlier work this paper cites.
Biases for emergent communication in multi-agent reinforcement learning, 2019
Tom Eccles, Yoram Bachrach, Guy Lever, Angeliki Lazaridou, and Thore Graepel · 1912
Earlier work this paper cites.
Improving policies via search in cooperative partially observable games
Adam Lerer, Hengyuan Hu, Jakob Foerster, and Noam Brown · 1912
Earlier work this paper cites.
The foundations of welfare economics
J. R. Hicks · 1939
Earlier work this paper cites.
Welfare propositions of economics and interpersonal comparisons of utility
Nicholas Kaldor · 1939
Earlier work this paper cites.
The bargaining problem
John F Nash Jr · 1950
Earlier work this paper cites.
Social choice and individual values
Kenneth J Arrow · 1951
Earlier work this paper cites.
A value for n-person games
Lloyd S Shapley · 1953
Earlier work this paper cites.
The theory of committees and elections
Duncan Black · 1958
Earlier work this paper cites.
The burning ships of Hernán Cortés
Winston A Reynolds · 1959
Earlier work this paper cites.
Counterspeculation, auctions, and competitive sealed tenders
William Vickrey · 1961
Earlier work this paper cites.
A taxonomy of 2 × \times 2 games
Anatol Rapoport · 1966
Earlier work this paper cites.
Multipart pricing of public goods
Edward H Clarke · 1971
Earlier work this paper cites.
Dynamic models of segregation
Thomas C Schelling · 1971
Earlier work this paper cites.
Manipulation of voting schemes: a general result
Allan Gibbard · 1973
Earlier work this paper cites.
Incentives in teams
Theodore Groves · 1973
Earlier work this paper cites.
Job market signaling
Michael Spense · 1973
Earlier work this paper cites.
Subjectivity and correlation in randomized strategies
Robert J. Aumann · 1974
Earlier work this paper cites.
Judgment under uncertainty: Heuristics and biases
Amos Tversky and Daniel Kahneman · 1974
Earlier work this paper cites.
Strategy-proofness and Arrow’s conditions: Existence and correspondence theorems for voting procedures and social welfare functions
Mark Allen Satterthwaite · 1975
Earlier work this paper cites.
Social learning theory
A. Bandura · 1977
Earlier work this paper cites.
Characterization of satisfactory mechanisms for the revelation of preferences for public goods
Jerry Green and Jean-Jacques Laffont · 1977
Earlier work this paper cites.
Reliability in communication systems and the evolution of altruism
Amotz Zahavi · 1977
Earlier work this paper cites.
The characterization of implementable choice rules
Kevin Roberts · 1979
Earlier work this paper cites.
The Strategy of Conflict
Thomas C Schelling · 1980
Earlier work this paper cites.
The contract net protocol: High-level communication and control in a distributed problem solver
Reid G Smith · 1980
Earlier work this paper cites.
The evolution of cooperation
Robert Axelrod and William Donald Hamilton · 1981
Earlier work this paper cites.
Strategic information transmission
Vincent P. Crawford and Joel Sobel · 1982
Earlier work this paper cites.
The Rise and Decline of Nations: Economic Growth, Stagflation, and Social Rigidities
Mancur Olson · 1982
Earlier work this paper cites.
Reliance, reputation, and breach of contract
Lewis A Kornhauser · 1983
Earlier work this paper cites.
Eliza — a computer program for the study of natural language communication between man and machine
Joseph Weizenbaum · 1983
Earlier work this paper cites.
The evolution of cooperation
Robert Axelrod · 1984
Earlier work this paper cites.
Does the autistic child have a “theory of mind”
Simon Baron-Cohen, Alan M Leslie, and Uta Frith · 1985
Earlier work this paper cites.
Goals, commitment, and identity
Amartya Sen · 1985
Earlier work this paper cites.
Game theory and political theory: An introduction
Peter C Ordeshook · 1986
Earlier work this paper cites.
Intention, plans, and practical reason
Michael Bratman · 1987
Earlier work this paper cites.
The evolution of individuality
Leo W Buss · 1987
Earlier work this paper cites.
Review of text-to-speech conversion for English
Dennis H Klatt · 1987
Earlier work this paper cites.
Flocks, herds and schools: A distributed behavioral model
Craig W Reynolds · 1987
Earlier work this paper cites.
Communication, coordination and Nash equilibrium
Joseph Farrell · 1988
Earlier work this paper cites.
Communication and interaction in multi-agent planning
Michael Georgeff · 1988
Earlier work this paper cites.
Society of mind
Marvin Minsky · 1988
Earlier work this paper cites.
ALVINN: An autonomous land vehicle in a neural network
Dean A. Pomerleau · 1988
Earlier work this paper cites.
The computational difficulty of manipulating an election
John J Bartholdi, Craig A Tovey, and Michael A Trick · 1989
Earlier work this paper cites.
Reputation and coalitions in medieval trade: evidence on the Maghribi traders
Avner Greif · 1989
Earlier work this paper cites.
The logic of images in international relations
Robert Jervis · 1989
Earlier work this paper cites.
The strategic structure of offer and acceptance: Game theory and the law of contract formation
Avery Katz · 1990
Earlier work this paper cites.
Credible commitments, contract enforcement problems and banks: Intermediation as credibility assurance
Arnoud WA Boot, Anjan V Thakor, and Gregory F Udell · 1991
Earlier work this paper cites.
Grounding in communication
Herbert H. Clark and Susan E. Brennan · 1991
Earlier work this paper cites.
How hard is it to control an election?
John J Bartholdi III, Craig A Tovey, and Michael A Trick · 1992
Earlier work this paper cites.
Institutions and social conflict
Jack Knight · 1992
Earlier work this paper cites.
Regulation of division of labor in insect societies
Gene E Robinson · 1992
Earlier work this paper cites.
On the synthesis of useful social laws for artificial agent societies (preliminary report)
Yoav Shoham and Moshe Tennenholtz · 1992
Earlier work this paper cites.
Meaning and credibility in cheap-talk games
Joseph Farrell · 1993
Earlier work this paper cites.
Why is rent-seeking so costly to growth?
Kevin M. Murphy, Andrei Shleifer, and Robert W. Vishny · 1993
Earlier work this paper cites.
Institutions and credible commitment
Douglass C North · 1993
Earlier work this paper cites.
Strategic alliance structuring: A game theoretic and transaction cost examination of interfirm cooperation
Arvind Parkhe · 1993
Earlier work this paper cites.
The origin of chromosomes I. Selection for linkage
John Maynard Smith and Eörs Száthmary · 1993
Earlier work this paper cites.
A market-oriented programming environment and its application to distributed multicommodity flow problems
Michael P Wellman · 1993
Earlier work this paper cites.
A speech-act-based negotiation protocol: design, implementation, and test use
Man Kit Chang and Carson C Woo · 1994
Earlier work this paper cites.
Coordination, commitment, and enforcement: The case of the merchant guild
Avner Greif, Paul Milgrom, and Barry R Weingast · 1994
Earlier work this paper cites.
Cultural beliefs and the organization of society: A historical and theoretical reflection on collectivist and individualist societies
Avner Greif · 1994
Earlier work this paper cites.
Um-prs: An implementation of the procedural reasoning system for multirobot applications
Jaeho Lee, Marcus J Huber, Edmund H Durfee, and Patrick G Kenny · 1994
Earlier work this paper cites.
The dynamics of informational cascades: The Monday demonstrations in Leipzig, East Germany, 1989-91
Susanne Lohmann · 1994
Earlier work this paper cites.
How might people interact with agents
Donald A Norman · 1994
Earlier work this paper cites.
A Computational Theory of Grounding in Natural Language Conversation
David Rood Traum · 1994
Earlier work this paper cites.
Rules of encounter: designing conventions for automated negotiation among computers
Jeffrey S Rosenschein and Gilad Zlotkin · 1994
Earlier work this paper cites.
An artificial discourse language for collaborative negotiation
Candace L Sidner · 1994
Earlier work this paper cites.
Multiagent systems
Munindar P Singh · 1994
Earlier work this paper cites.
TD-Gammon, a self-teaching backgammon program, achieves master-level play
Gerald Tesauro · 1994
Earlier work this paper cites.
Commitment and observability in games
Kyle Bagwell · 1995
Earlier work this paper cites.
Understanding the functions of norms in social groups through simulation
Rosaria Conte and Cristiano Castelfranchi · 1995
Earlier work this paper cites.
Rationalist explanations for war
James D. Fearon · 1995
Earlier work this paper cites.
Designing and building a negotiating automated agent
Sarit Kraus and Daniel Lehmann · 1995
Earlier work this paper cites.
Interpersonal relations: Mixed-motive interaction
Samuel S Komorita and Craig D Parks · 1995
Earlier work this paper cites.
An integrative model of organizational trust
Roger C Mayer, James H Davis, and F David Schoorman · 1995
Earlier work this paper cites.
Artificial social systems
Yoram Moses and Moshe Tennenholtz · 1995
Earlier work this paper cites.
The construction of social reality
John R Searle · 1995
Earlier work this paper cites.
Issues in automated negotiation and electronic commerce: Extending the contract net framework
Tuomas Sandholm, Victor R Lesser, et al · 1995
Earlier work this paper cites.
On social laws for artificial agent societies: off-line design
Yoav Shoham and Moshe Tennenholtz · 1995
Earlier work this paper cites.
Adaptation and learning in multi-agent systems: Some remarks and a bibliography
Gerhard Weiß · 1995
Earlier work this paper cites.
Fair Division: From cake-cutting to dispute resolution
Steven J Brams and Alan D Taylor · 1996
Earlier work this paper cites.
Coordination techniques for distributed artificial intelligence
Nicholas R Jennings · 1996
Earlier work this paper cites.
The swarm simulation system: A toolkit for building multi-agent simulations
Nelson Minar, Roger Burkhart, Chris Langton, and Manor Askenazi · 1996
Earlier work this paper cites.
Negotiation principles
H Juergen Mueller · 1996
Earlier work this paper cites.
Foundations of Distributed Artificial Intelligence
Greg MP O’Hare and Nicholas R Jennings, editors · 1996
Earlier work this paper cites.
A machine-learning approach to automated negotiation and prospects for electronic commerce
Jim R Oliver · 1996
Earlier work this paper cites.
Limitations of the Vickrey auction in computational multiagent systems
Tuomas W Sandholm · 1996
Earlier work this paper cites.
Trust and commitments
Chris Snijders · 1996
Earlier work this paper cites.
The economics of convention
H Peyton Young · 1996
Earlier work this paper cites.
A formal specification of dMARS
Mark d’Inverno, David Kinny, Michael Luck, and Michael Wooldridge · 1997
Earlier work this paper cites.
Signaling foreign policy interests: Tying hands versus sinking costs
James D Fearon · 1997
Earlier work this paper cites.
Robocup: The robot world cup initiative
Hiroaki Kitano, Minoru Asada, Yasuo Kuniyoshi, Itsuki Noda, and Eiichi Osawa · 1997
Earlier work this paper cites.
Negotiation and cooperation in multi-agent environments
Sarit Kraus · 1997
Earlier work this paper cites.
The major transitions in evolution
John Maynard Smith and Eörs Száthmary · 1997
Earlier work this paper cites.
On the emergence of social conventions: modeling, analysis, and simulations
Yoav Shoham and Moshe Tennenholtz · 1997
Earlier work this paper cites.
The origins of credible commitment to the market
John T Williams, Brian Collins, and Mark I Lichbach · 1997
Earlier work this paper cites.
Social choice theory, game theory, and positive political theory
David Austen-Smith and Jeffrey S Banks · 1998
Earlier work this paper cites.
Learning to teach with a reinforcement learning agent
Joseph Beck · 1998
Earlier work this paper cites.
Modelling social action for AI agents
Cristiano Castelfranchi · 1998
Earlier work this paper cites.
The dynamics of reinforcement learning in cooperative multiagent systems
Caroline Claus and Craig Boutilier · 1998
Earlier work this paper cites.
Autonomous norm acceptance
Rosaria Conte, Cristiano Castelfranchi, and Frank Dignum · 1998
Earlier work this paper cites.
Principles of trust for MAS: Cognitive anatomy, social importance, and quantification
Cristiano Castelfranchi and Rino Falcone · 1998
Earlier work this paper cites.
Commitment problems and the spread of ethnic conflict
James D Fearon · 1998
Earlier work this paper cites.
The belief-desire-intention model of agency
Michael Georgeff, Barney Pell, Martha Pollack, Milind Tambe, and Michael Wooldridge · 1998
Earlier work this paper cites.
Games, threats, and treaties: understanding commitments in international relations
Jon Hovi · 1998
Earlier work this paper cites.
A roadmap of agent research and development
Nicholas R Jennings, Katia Sycara, and Michael Wooldridge · 1998
Earlier work this paper cites.
Designing for human-agent interaction
Michael Lewis · 1998
Earlier work this paper cites.
Determining successful negotiation strategies: An evolutionary approach
Noyda Matos, Carles Sierra, and Nicholas R Jennings · 1998
Earlier work this paper cites.
Evolution of indirect reciprocity by image scoring
Martin A Nowak and Karl Sigmund · 1998
Earlier work this paper cites.
A behavioral approach to the rational choice theory of collective action: Presidential address, American Political Science Association, 1997
Elinor Ostrom · 1998
Earlier work this paper cites.
Deliberative normative agents: Principles and architecture
Cristiano Castelfranchi, Frank Dignum, Catholijn M Jonker, and Jan Treur · 1999
Earlier work this paper cites.
Autonomous agents with norms
Frank Dignum · 1999
Earlier work this paper cites.
Multi-agent systems: an introduction to distributed artificial intelligence
Jacques Ferber · 1999
Earlier work this paper cites.
JAM: A BDI-theoretic mobile agent architecture
Marcus J Huber · 1999
Earlier work this paper cites.
Algorithmic mechanism design
Noam Nisan and Amir Ronen · 1999
Earlier work this paper cites.
An adaptive agent bidding strategy based on stochastic modeling
Sunju Park, Edmund H Durfee, and William P Birmingham · 1999
Earlier work this paper cites.
In the shadow of power: States and strategies in international politics
Robert Powell · 1999
Earlier work this paper cites.
Is imitation learning the route to humanoid robots
Stefan Schaal · 1999
Earlier work this paper cites.
Multiagent systems: a modern approach to distributed artificial intelligence
Gerhard Weiss, editor · 1999
Earlier work this paper cites.
Collaborative multi-robot exploration
Wolfram Burgard, Mark Moors, Dieter Fox, Reid Simmons, and Sebastian Thrun · 2000
Earlier work this paper cites.
Game-theoretic interpretations of commitment
Jack Hirshleifer et al · 2000
Earlier work this paper cites.
A simple adaptive procedure leading to correlated equilibrium
Sergiu Hart and Andreu Mas-Colell · 2000
Earlier work this paper cites.
On agent-based software engineering
Nicholas R Jennings · 2000
Earlier work this paper cites.
Optimal contracts when enforcement is a decision variable
Stefan Krasa and Anne P Villamil · 2000
Earlier work this paper cites.
Algorithms for inverse reinforcement learning
Andrew Y Ng, Stuart J Russell, et al · 2000
Earlier work this paper cites.
Collective action and the evolution of social norms
Elinor Ostrom · 2000
Earlier work this paper cites.
A study of collusion in first-price auctions
Martin Pesendorfer · 2000
Earlier work this paper cites.
Multiagent systems: A survey from a machine learning perspective
Peter Stone and Manuela Veloso · 2000
Earlier work this paper cites.
A cultural perspective in social interface
Yugo Takeuchi, Yasuhiro Katagiri, Clifford Nass, and BJ Fogg · 2000
Earlier work this paper cites.
Languages for negotiation
Michael Wooldridge and Simon Parsons · 2000
Earlier work this paper cites.
Trust management through reputation mechanisms
Giorgos Zacharia and Pattie Maes · 2000
Earlier work this paper cites.
A formal model of open agent societies
Alexander Artikis and Jeremy Pitt · 2001
Earlier work this paper cites.
Incentives for sharing in peer-to-peer networks
Philippe Golle, Kevin Leyton-Brown, Ilya Mironov, and Mark Lillibridge · 2001
Earlier work this paper cites.
Autonomous bidding agents in the trading agent competition
Amy Greenwald and Peter Stone · 2001
Earlier work this paper cites.
Automated negotiation: prospects, methods and challenges
Nicholas R Jennings, Peyman Faratin, Alessio R Lomuscio, Simon Parsons, Carles Sierra, and Michael Wooldridge · 2001
Earlier work this paper cites.
Silly rules improve the capacity of agents to learn stable enforcement and compliance behaviors
Raphael Köster, Dylan Hadfield-Menell, Gillian K Hadfield, and Joel Z Leibo · 2001
Earlier work this paper cites.
Strategic negotiation in multiagent environments
Sarit Kraus · 2001
Earlier work this paper cites.
Bargaining with limited computation: Deliberation equilibrium
Kate Larson and Tuomas Sandholm · 2001
Earlier work this paper cites.
Securities regulation as lobster trap: A credible commitment theory of mandatory disclosure
Edward Rock · 2001
Earlier work this paper cites.
Advances in multi-robot systems
Tamio Arai, Enrico Pagello, Lynne E Parker, et al · 2002
Earlier work this paper cites.
Deep blue
Murray Campbell, A. Joseph Hoane Jr, and Feng-hsiung Hsu · 2002
Earlier work this paper cites.
Complexity of manipulating elections with few candidates
Vincent Conitzer and Tuomas Sandholm · 2002
Earlier work this paper cites.
Vote elicitation: Complexity and strategy-proofness
Vincent Conitzer and Tuomas Sandholm · 2002
Earlier work this paper cites.
From desires, obligations and norms to goals
Frank Dignum, David Kinny, and Liz Sonenberg · 2002
Earlier work this paper cites.
Using similarity criteria to make issue trade-offs in automated negotiations
Peyman Faratin, Carles Sierra, and Nicholas R Jennings · 2002
Earlier work this paper cites.
Desiderata for agent argumentation protocols
Peter McBurney, Simon Parsons, and Michael Wooldridge · 2002
Earlier work this paper cites.
Game theory and decision theory in multi-agent systems
Simon Parsons and Michael Wooldridge · 2002
Earlier work this paper cites.
Trust among strangers in internet transactions: Empirical analysis of eBay’s reputation system
Paul Resnick and Richard Zeckhauser · 2002
Earlier work this paper cites.
Common ground
Robert Stalnaker · 2002
Earlier work this paper cites.
Tractable multiagent planning for epistemic goals
Wiebe van der Hoek and Michael Wooldridge · 2002
Earlier work this paper cites.
The carrot or the stick: Rewards, punishments, and cooperation
James Andreoni, William Harbaugh, and Lise Vesterlund · 2003
Cited alongside, same era.
Adaptive agents and multi-agent systems: adaptation and multi-agent learning
Eduardo Alonso, Daniel Kudenko, and Dimitar Kazakov, editors · 2003
Cited alongside, same era.
Towards a structured design of electronic negotiations
Martin Bichler, Gregory Kersten, and Stefan Strecker · 2003
Cited alongside, same era.
Coordinating multiple agents for workflow-oriented process orchestration
M. Brian Blake · 2003
Cited alongside, same era.
The digitization of word of mouth: Promise and challenges of online feedback mechanisms
Chrysanthos Dellarocas · 2003
Cited alongside, same era.
A reputation system for peer-to-peer networks
Minaxi Gupta, Paul Judge, and Mostafa Ammar · 2003
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Later among the works it cites.
Diverse randomized agents vote to win
Albert Xin Jiang, Leandro Soriano Marcolino, Ariel D Procaccia, Tuomas Sandholm, Nisarg Shah, and Milind Tambe · 2014
Later among the works it cites.
Order within anarchy: The laws of war as an international institution
James D Morrow · 2014
Later among the works it cites.
GloVe: Global vectors for word representation
Jeffrey Pennington, Richard Socher, and Christopher D Manning · 2014
Later among the works it cites.
Social heuristics shape intuitive cooperation
David G Rand, Alexander Peysakhovich, Gordon T Kraft-Todd, George E Newman, Owen Wurzbacher, Martin A Nowak, and Joshua D Greene · 2014
Later among the works it cites.
Review of speech-to-text recognition technology for enhancing learning
Rustam Shadiev, Wu-Yuin Hwang, Nian-Shing Chen, and Yueh-Min Huang · 2014
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Learning to resolve alliance dilemmas in many-player zero-sum games
Edward Hughes, Thomas W. Anthony, Tom Eccles, Joel Z. Leibo, David Balduzzi, and Yoram Bachrach · 2003
Cited alongside, same era.
"Other-play" for zero-shot coordination
Hengyuan Hu, Adam Lerer, Alex Peysakhovich, and Jakob Foerster · 2003
Cited alongside, same era.
Incentives for cooperation in peer-to-peer networks
Kevin Lai, Michal Feldman, Ion Stoica, and John Chuang · 2003
Cited alongside, same era.
Agent technology: Enabling next generation computing (A roadmap for agent based computing)
Michael Luck, Peter McBurney, and Chris Preist · 2003
Cited alongside, same era.
Implementing an untrusted operating system on trusted hardware
David Lie, Chandramohan A Thekkath, and Mark Horowitz · 2003
Cited alongside, same era.
Identity crisis: anonymity vs reputation in P2P systems
Sergio Marti and Hector Garcia-Molina · 2003
Cited alongside, same era.
Later among the works it cites.
Speaking our minds: Why human communication is different, and how language evolved to make it special
Thom Scott-Phillips · 2014
Later among the works it cites.
Reward and punishment in social dilemmas
Paul A M Van Lange, Bettina Rockenbach, and Toshio Yamagishi, editors · 2014
Later among the works it cites.
Designing collective behavior in a termite-inspired robot construction team
Justin Werfel, Kirstin Petersen, and Radhika Nagpal · 2014
Later among the works it cites.
Teacher-student framework: a reinforcement learning approach, 2014
Matthieu Zimmer, Paolo Viappiani, and Paul Weng · 2014
Later among the works it cites.
Multi-agent pathfinding as a combinatorial auction
Ofra Amir, Guni Sharon, and Roni Stern · 2015
Later among the works it cites.
Negotiating with other minds: the role of recursive theory of mind in negotiation with incomplete information
Harmen de Weerd, Rineke Verbrugge, and Bart Verheij · 2015
Later among the works it cites.
World Politics: Interests, Interactions, Institutions: Third International Student Edition
Jeffry A Frieden and David A Lake · 2015
Later among the works it cites.
Perfect prediction equilibrium
Ghislain Fourny, Stéphane Reiche, and Jean-Pierre Dupuy · 2015
Later among the works it cites.
When security games go green: Designing defender strategies to prevent poaching and illegal fishing
Fei Fang, Peter Stone, and Milind Tambe · 2015
Later among the works it cites.
Sapiens: A Brief History of Humankind
Yuval Noah Harari · 2015
Later among the works it cites.
Interplay of approximate planning strategies
Quentin J M Huys, Níall Lally, Paul Faulkner, Neir Eshel, Erich Seifritz, Samuel J Gershman, Peter Dayan, and Jonathan P Roiser · 2015
Later among the works it cites.
Deep recurrent Q-learning for partially observable MDPs
Matthew Hausknecht and Peter Stone · 2015
Later among the works it cites.
Giraffe: Using deep reinforcement learning to play chess
Matthew Lai · 2015
Later among the works it cites.
Human-level control through deep reinforcement learning
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A Rusu, Joel Veness, Marc G Bellemare, Alex Graves, Martin Riedmiller, Andreas K Fidjeland, Georg Ostrovski, Stig Petersen, Charles Beattie, Amir Sadik, Ioannis Antonoglou, Helen King, Dharshan Kumaran, Daan Wierstra, Shane Legg, and Demis Hassabis · 2015
Later among the works it cites.
Improved human–robot team performance through cross-training, an approach inspired by human team training practices
Stefanos Nikolaidis, Przemyslaw Lasota, Ramya Ramakrishnan, and Julie Shah · 2015
Later among the works it cites.
The ease and extent of recursive mindreading, across implicit and explicit tasks
Cathleen O’Grady, Christian Kliesch, Kenny Smith, and Thomas C Scott-Phillips · 2015
Later among the works it cites.
Economic reasoning and artificial intelligence
David Parkes and Michael Wellman · 2015
Later among the works it cites.
Google’s driverless cars run into problem: Cars with drivers
Matt Richtel and Conor Dougherty · 2015
Later among the works it cites.
Research priorities for robust and beneficial artificial intelligence
Stuart Russell, Daniel Dewey, and Max Tegmark · 2015
Later among the works it cites.
A discrete and bounded envy-free cake cutting protocol for any number of agents
Haris Aziz and Simon Mackenzie · 2016
Later among the works it cites.
Concrete problems in AI safety
Dario Amodei, Chris Olah, Jacob Steinhardt, Paul Christiano, John Schulman, and Dan Mané · 2016
Later among the works it cites.
Handbook of computational social choice
Felix Brandt, Vincent Conitzer, Ulle Endriss, Jérôme Lang, and Ariel D Procaccia, editors · 2016
Later among the works it cites.
Incomplete information with communication in voting
Craig Boutilier and Jeffrey S. Rosenschein · 2016
Later among the works it cites.
Blockchains and smart contracts for the internet of things
Konstantinos Christidis and Michael Devetsikiotis · 2016
Later among the works it cites.
Learning to communicate with deep multi-agent reinforcement learning
Jakob Foerster, Ioannis Alexandros Assael, Nando De Freitas, and Shimon Whiteson · 2016
Later among the works it cites.
From institutions to code: Towards automated generation of smart contracts
Christopher K Frantz and Mariusz Nowostawski · 2016
Later among the works it cites.
Homo Deus: A brief history of tomorrow
Yuval Noah Harari · 2016
Later among the works it cites.
The Secret of Our Success: How Culture Is Driving Human Evolution, Domesticating Our Species, and Making Us Smarter
Joseph Henrich · 2016
Later among the works it cites.
Cooperative inverse reinforcement learning
Dylan Hadfield-Menell, Stuart J Russell, Pieter Abbeel, and Anca Dragan · 2016
Later among the works it cites.
Equality of opportunity in supervised learning
Moritz Hardt, Eric Price, and Nathan Srebro · 2016
Later among the works it cites.
Hawk: The blockchain model of cryptography and privacy-preserving smart contracts
Ahmed Kosba, Andrew Miller, Elaine Shi, Zikai Wen, and Charalampos Papamanthou · 2016
Later among the works it cites.
Making smart contracts smarter
Loi Luu, Duc-Hiep Chu, Hrishi Olickel, Prateek Saxena, and Aquinas Hobor · 2016
Later among the works it cites.
Learning to coordinate: Co-evolution and correlated equilibrium
Alejandro Lee-Penagos · 2016
Later among the works it cites.
Multi-agent cooperation and the emergence of (natural) language
Angeliki Lazaridou, Alexander Peysakhovich, and Marco Baroni · 2016
Later among the works it cites.
People do not feel guilty about exploiting machines
Celso De Melo, Stacy Marsella, and Jonathan Gratch · 2016
Later among the works it cites.
Automated mechanism design without money via machine learning
Harikrishna Narasimhan, Shivani Agarwal, and David C. Parkes · 2016
Later among the works it cites.
Abstractive text summarization using sequence-to-sequence RNNs and beyond
Ramesh Nallapati, Bowen Zhou, Cícero Noguiera dos Santos, Çaglar Gülçehre, and Bing Xiang · 2016
Later among the works it cites.
WaveNet: A generative model for raw audio
Aaron van den Oord, Sander Dieleman, Heiga Zen, Karen Simonyan, Oriol Vinyals, Alex Graves, Nal Kalchbrenner, Andrew Senior, and Koray Kavukcuoglu · 2016
Later among the works it cites.
Mastering the game of go with deep neural networks and tree search
David Silver, Aja Huang, Chris J Maddison, Arthur Guez, Laurent Sifre, George van den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, Sander Dieleman, Dominik Grewe, John Nham, Nal Kalchbrenner, Ilya Sutskever, Timothy Lillicrap, Madeleine Leach, Koray Kavukcuoglu, Thore Graepel, and Demis Hassabis · 2016
Later among the works it cites.
Designing the user interface: strategies for effective human-computer interaction
Ben Shneiderman, Catherine Plaisant, Maxine Cohen, Steven Jacobs, Niklas Elmqvist, and Nicholas Diakopoulos · 2016
Later among the works it cites.
Learning multiagent communication with backpropagation
Sainbayar Sukhbaatar, Arthur Szlam, and Rob Fergus · 2016
Later among the works it cites.
Information gathering actions over human internal state
Dorsa Sadigh, S Shankar Sastry, Sanjit A Seshia, and Anca Dragan · 2016
Later among the works it cites.
Repeated inverse reinforcement learning
Kareem Amin, Nan Jiang, and Satinder Singh · 2017
Later among the works it cites.
Continuous adaptation via meta-learning in nonstationary and competitive environments
Maruan Al-Shedivat, Trapit Bansal, Yuri Burda, Ilya Sutskever, Igor Mordatch, and Pieter Abbeel · 2017
Later among the works it cites.
Extravehicular activity operations concepts under communication latency and bandwidth constraints
Kara H. Beaton, Steven P. Chappell, Andrew F. J. Abercromby, Matthew J. Miller, Shannon Kobs Nawotniak, Scott S. Hughes, Allyson Brady, and Darlene S. S. Lim · 2017
Later among the works it cites.
Norms and value based reasoning: justifying compliance and violation
Trevor Bench-Capon and Sanjay Modgil · 2017
Later among the works it cites.
Common ground and development
Manuel Bohn and Bahar Köymen · 2017
Later among the works it cites.
An empirical analysis of smart contracts: platforms, applications, and design patterns
Massimo Bartoletti and Livio Pompianu · 2017
Later among the works it cites.
Observational learning by reinforcement learning
Diana Borsa, Bilal Piot, Rémi Munos, and Olivier Pietquin · 2017
Later among the works it cites.
Deep reinforcement learning from human preferences
Paul F Christiano, Jan Leike, Tom Brown, Miljan Martic, Shane Legg, and Dario Amodei · 2017
Later among the works it cites.
A parametric, resource-bounded generalization of Löb’s theorem, and a robust cooperation criterion for open-source game theory
Andrew Critch · 2017
Later among the works it cites.
Disarmament games
Yuan Deng and Vincent Conitzer · 2017
Later among the works it cites.
Simultaneously learning and advising in multiagent reinforcement learning
Felipe Leno Da Silva, Ruben Glatt, and Anna Helena Reali Costa · 2017
Later among the works it cites.
Negotiating with other minds: the role of recursive theory of mind in negotiation with incomplete information
Harmen de Weerd, Rineke Verbrugge, and Bart Verheij · 2017
Later among the works it cites.
Artificial intelligence & collusion: When computers inhibit competition
Ariel Ezrachi and Maurice E Stucke · 2017
Later among the works it cites.
Learning with opponent-learning awareness
Jakob N Foerster, Richard Y Chen, Maruan Al-Shedivat, Shimon Whiteson, Pieter Abbeel, and Igor Mordatch · 2017
Later among the works it cites.
Gesture, sign, and language: The coming of age of sign language and gesture studies
Susan Goldin-Meadow and Diane Brentari · 2017
Later among the works it cites.
Emergence of language with multi-agent games: Learning to communicate with sequences of symbols
Serhii Havrylov and Ivan Titov · 2017
Later among the works it cites.
Natural language does not emerge’naturally’in multi-agent dialog
Satwik Kottur, José MF Moura, Stefan Lee, and Dhruv Batra · 2017
Later among the works it cites.
The origins of language in teaching
Kevin N. Laland · 2017
Later among the works it cites.
Deal or no deal? end-to-end learning for negotiation dialogues
Mike Lewis, Denis Yarats, Yann N Dauphin, Devi Parikh, and Dhruv Batra · 2017
Later among the works it cites.
Multi-agent reinforcement learning in sequential social dilemmas
Joel Z. Leibo, Vinicius Zambaldi, Marc Lanctot, Janusz Marecki, and Thore Graepel · 2017
Later among the works it cites.
Computational aspects of strategic behaviour in elections with top-truncated ballots
Vijay Menon and Kate Larson · 2017
Later among the works it cites.
Communication-efficient learning of deep networks from decentralized data
H. Brendan McMahan, Eider Moore, Daniel Ramage, Seth Hampson, and Blaise Aguera y Arcas · 2017
Later among the works it cites.
Deepstack: Expert-level artificial intelligence in heads-up no-limit poker
Matej Moravčík, Martin Schmid, Neil Burch, Viliam Lisỳ, Dustin Morrill, Nolan Bard, Trevor Davis, Kevin Waugh, Michael Johanson, and Michael Bowling · 2017
Later among the works it cites.
SecureML: A system for scalable privacy-preserving machine learning
Payman Mohassel and Yupeng Zhang · 2017
Later among the works it cites.
The emergence of altruism as a social norm
María Pereda, Pablo Brañas-Garza, Ismael Rodriguez-Lara, and Angel Sánchez · 2017
Later among the works it cites.
Order without law: Reputation promotes cooperation in a cryptomarket for illegal drugs
Wojtek Przepiorka, Lukas Norbutas, and Rense Corten · 2017
Later among the works it cites.
Locally noisy autonomous agents improve global human coordination in network experiments
Hirokazu Shirado and Nicholas A Christakis · 2017
Later among the works it cites.
Mastering chess and shogi by self-play with a general reinforcement learning algorithm
David Silver, Thomas Hubert, Julian Schrittwieser, Ioannis Antonoglou, Matthew Lai, Arthur Guez, Marc Lanctot, Laurent Sifre, Dharshan Kumaran, Thore Graepel, Timothy Lillicrap, Karen Simonyan, and Demis Hassabis · 2017
Later among the works it cites.
Reinforcement mechanism design
Pingzhong Tang · 2017
Later among the works it cites.
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin · 2017
Later among the works it cites.
A gift from knowledge distillation: Fast optimization, network minimization and transfer learning
Junho Yim, Donggyu Joo, Jihoon Bae, and Junmo Kim · 2017
Later among the works it cites.
Mechanism design for social good
Rediet Abebe and Kira Goldner · 2018
Later among the works it cites.
Occam’s razor is insufficient to infer the preferences of irrational agents
Stuart Armstrong and Sören Mindermann · 2018
Later among the works it cites.
Machine-to-machine communication: An overview of opportunities
Oluwatosin Ahmed Amodu and Mohamed Othman · 2018
Later among the works it cites.
Autonomous agents modelling other agents: A comprehensive survey and open problems
Stefano V Albrecht and Peter Stone · 2018
Later among the works it cites.
Social norms
Cristina Bicchieri, Ryan Muldoon, and Alessandro Sontuoso · 2018
Later among the works it cites.
Emergent communication through negotiation
Kris Cao, Angeliki Lazaridou, Marc Lanctot, Joel Z. Leibo, Karl Tuyls, and Stephen Clark · 2018
Later among the works it cites.
Cooperating with machines
Jacob W Crandall, Mayada Oudah, Fatimah Ishowo-Oloko, Sherief Abdallah, Jean-François Bonnefon, Manuel Cebrian, Azim Shariff, Michael A Goodrich, and Iyad Rahwan · 2018
Later among the works it cites.
Multi-agent common knowledge reinforcement learning
Jakob N. Foerster, Christian A. Schröder de Witt, Gregory Farquhar, Philip H. S. Torr, Wendelin Boehmer, and Shimon Whiteson · 2018
Later among the works it cites.
The biology and evolution of speech: a comparative analysis
W Tecumseh Fitch · 2018
Later among the works it cites.
Deep learning for revenue-optimal auctions with budgets
Zhe Feng, Harikrishna Narisimhan, and David C. Parkes · 2018
Later among the works it cites.
Bayesian action decoder for deep multi-agent reinforcement learning
Jakob N Foerster, Francis Song, Edward Hughes, Neil Burch, Iain Dunning, Shimon Whiteson, Matthew Botvinick, and Michael Bowling · 2018
Later among the works it cites.
On legal contracts, imperative and declarative smart contracts, and blockchain systems
Guido Governatori, Florian Idelberger, Zoran Milosevic, Regis Riveret, Giovanni Sartor, and Xiwei Xu · 2018
Later among the works it cites.
Inequity aversion resolves intertemporal social dilemmas
Edward Hughes, Joel Z. Leibo, Matthew G. Philips, Karl Tuyls, Edgar A. Duéñez-Guzmán, Antonio García Castañeda, Iain Dunning, Tina Zhu, Kevin R. McKee, Raphael Koster, Heather Roff, and Thore Graepel · 2018
Later among the works it cites.
Reward learning from human preferences and demonstrations in Atari
Borja Ibarz, Jan Leike, Tobias Pohlen, Geoffrey Irving, Shane Legg, and Dario Amodei · 2018
Later among the works it cites.
Intrinsic social motivation via causal influence in multi-agent RL
Natasha Jaques, Angeliki Lazaridou, Edward Hughes, Çaglar Gülçehre, Pedro A. Ortega, DJ Strouse, Joel Z. Leibo, and Nando de Freitas · 2018
Later among the works it cites.
Julia Kreutzer, Joshua Uyheng, and Stefan Riezler · 2018
Later among the works it cites.
Stable opponent shaping in differentiable games
Alistair Letcher, Jakob Foerster, David Balduzzi, Tim Rocktäschel, and Shimon Whiteson · 2018
Later among the works it cites.
Emergence of linguistic communication from referential games with symbolic and pixel input
Angeliki Lazaridou, Karl Moritz Hermann, Karl Tuyls, and Stephen Clark · 2018
Later among the works it cites.
Scalable agent alignment via reward modeling: a research direction
Jan Leike, David Krueger, Tom Everitt, Miljan Martic, Vishal Maini, and Shane Legg · 2018
Later among the works it cites.
Twenty-five years of group decision and negotiation: a bibliometric overview
Sigifredo Laengle, Nikunja Mohan Modak, Jose M Merigo, and Gustavo Zurita · 2018
Later among the works it cites.
BARS: a blockchain-based anonymous reputation system for trust management in VANETs
Zhaojun Lu, Qian Wang, Gang Qu, and Zhenglin Liu · 2018
Later among the works it cites.
The building blocks of interpretability
Chris Olah, Arvind Satyanarayan, Ian Johnson, Shan Carter, Ludwig Schubert, Katherine Ye, and Alexander Mordvintsev · 2018
Later among the works it cites.
Learning complex dexterous manipulation with deep reinforcement learning and demonstrations
Aravind Rajeswaran, Vikash Kumar, Abhishek Gupta, Giulia Vezzani, John Schulman, Emanuel Todorov, and Sergey Levine · 2018
Later among the works it cites.
Taxonomy and definitions for terms related to driving automation systems for on-road motor vehicles, 2018
SAE On-Road Automated Vehicle Standards Committee et al · 2018
Later among the works it cites.
Reinforcement learning: An introduction
Richard S Sutton and Andrew G Barto · 2018
Later among the works it cites.
A general reinforcement learning algorithm that masters chess, shogi, and go through self-play
David Silver, Thomas Hubert, Julian Schrittwieser, Ioannis Antonoglou, Matthew Lai, Arthur Guez, Marc Lanctot, Laurent Sifre, Dharshan Kumaran, Thore Graepel, Timothy Lillicrap, Karen Simonyan, and Demis Hassabis · 2018
Later among the works it cites.
Kickstarting deep reinforcement learning
Simon Schmitt, Jonathan J Hudson, Augustin Zidek, Simon Osindero, Carl Doersch, Wojciech M Czarnecki, Joel Z Leibo, Heinrich Kuttler, Andrew Zisserman, Karen Simonyan, and S M Ali Eslami · 2018
Later among the works it cites.
Reward learning from narrated demonstrations
Hsiao-Yu Fish Tung, Adam W Harley, Liang-Kang Huang, and Katerina Fragkiadaki · 2018
Later among the works it cites.
How children solve the two challenges of cooperation
Felix Warneken · 2018
Later among the works it cites.
Design patterns for smart contracts in the ethereum ecosystem
Maximilian Wöhrer and Uwe Zdun · 2018
Later among the works it cites.
Machine learning to strengthen democracy
Ben Armstrong and Kate Larson · 2019
Later among the works it cites.
The evolution of human cooperation
Coren L. Apicella and Joan B. Silk · 2019
Later among the works it cites.
The Hanabi challenge: A new frontier for AI research
Nolan Bard, Jakob N Foerster, Sarath Chandar, Neil Burch, Marc Lanctot, H Francis Song, Emilio Parisotto, Vincent Dumoulin, Subhodeep Moitra, Edward Hughes, et al · 2019
Later among the works it cites.
Extrapolating beyond suboptimal demonstrations via inverse reinforcement learning from observations
Daniel Brown, Wonjoon Goo, Prabhat Nagarajan, and Scott Niekum · 2019
Later among the works it cites.
Learning to understand goal specifications by modelling reward
Dzmitry Bahdanau, Felix Hill, Jan Leike, Edward Hughes, Arian Hosseini, Pushmeet Kohli, and Edward Grefenstette · 2019
Later among the works it cites.
Does machine translation affect international trade? evidence from a large digital platform
Erik Brynjolfsson, Xiang Hui, and Meng Liu · 2019
Later among the works it cites.
Superhuman AI for multiplayer poker
Noam Brown and Tuomas Sandholm · 2019
Later among the works it cites.
Blockchain disruption and smart contracts
Lin William Cong and Zhiguo He · 2019
Later among the works it cites.
The unreasonable fairness of maximum Nash welfare
Ioannis Caragiannis, David Kurokawa, Hervé Moulin, Ariel D Procaccia, Nisarg Shah, and Junxing Wang · 2019
Later among the works it cites.
Cooperative AI workshop, NeurIPS 2020, September 2019
Cooperative AI Workshop · 2019
Later among the works it cites.
On the utility of learning about humans for human-AI coordination
Micah Carroll, Rohin Shah, Mark K Ho, Thomas L Griffiths, Sanjit A Seshia, Pieter Abbeel, and Anca Dragan · 2019
Later among the works it cites.
Trusted AI and the contribution of trust modeling in multiagent systems
Robin Cohen, Mike Schaekermann, Sihao Liu, and Michael Cormier · 2019
Later among the works it cites.
Optimal auctions through deep learning
Paul Duetting, Zhe Feng, Harikrishna Narasimhan, David C. Parkes, and Sai Srivatsa Ravindranath · 2019
Later among the works it cites.
Negotiating at the United Nations: A Practitioner’s Guide
Rebecca W Gaudiosi, Jimena Leiva Roesch, and Ye-Min Wu · 2019
Later among the works it cites.
Human-level performance in 3d multiplayer games with population-based reinforcement learning
Max Jaderberg, Wojciech M Czarnecki, Iain Dunning, Luke Marris, Guy Lever, Antonio Garcia Castañeda, Charles Beattie, Neil C Rabinowitz, Ari S Morcos, Avraham Ruderman, Nicolas Sonnerat, Tim Green, Louise Deason, Joel Z. Leibo, DAvid Silver, Demis Hassabis, Koray Kavukcuoglu, and Thore Graepel · 2019
Later among the works it cites.
Social influence as intrinsic motivation for multi-agent deep reinforcement learning
Natasha Jaques, Angeliki Lazaridou, Edward Hughes, Caglar Gulcehre, Pedro Ortega, DJ Strouse, Joel Z Leibo, and Nando De Freitas · 2019
Later among the works it cites.
Evolution of social norms and correlated equilibria
Bryce Morsky and Erol Akçay · 2019
Later among the works it cites.
Norm emergence in multiagent systems: a viewpoint paper
Andreasa Morris-Martin, Marina De Vos, and Julian Padget · 2019
Later among the works it cites.
A review of deep learning based speech synthesis
Yishuang Ning, Sheng He, Zhiyong Wu, Chunxiao Xing, and Liang-Jie Zhang · 2019
Later among the works it cites.
Speech recognition using deep neural networks: A systematic review
Ali Bou Nassif, Ismail Shahin, Imtinan Attili, Mohammad Azzeh, and Khaled Shaalan · 2019
Later among the works it cites.
Learning to teach in cooperative multiagent reinforcement learning
Shayegan Omidshafiei, Dong-Ki Kim, Miao Liu, Gerald Tesauro, Matthew Riemer, Christopher Amato, Murray Campbell, and Jonathan P How · 2019
Later among the works it cites.
Google’s duplex: Pretending to be human
Daniel E O’Leary · 2019
Later among the works it cites.
Polis: Input crowd, output meaning, 2019
Polis · 2019
Later among the works it cites.
Adversarial robustness through local linearization
Chongli Qin, James Martens, Sven Gowal, Dilip Krishnan, Krishnamurthy Dvijotham, Alhussein Fawzi, Soham De, Robert Stanforth, and Pushmeet Kohli · 2019
Later among the works it cites.
Human compatible: Artificial intelligence and the problem of control
Stuart Russell · 2019
Later among the works it cites.
Finding friend and foe in multi-agent games
Jack Serrino, Max Kleiman-Weiner, David C Parkes, and Josh Tenenbaum · 2019
Later among the works it cites.
Social behavior for autonomous vehicles
Wilko Schwarting, Alyssa Pierson, Javier Alonso-Mora, Sertac Karaman, and Daniela Rus · 2019
Later among the works it cites.
Grandmaster level in StarCraft II using multi-agent reinforcement learning
Oriol Vinyals, Igor Babuschkin, Wojciech M Czarnecki, Michaël Mathieu, Andrew Dudzik, Junyoung Chung, David H Choi, Richard Powell, Timo Ewalds, Petko Georgiev, et al · 2019
Later among the works it cites.
HIBERT: Document level pre-training of hierarchical bidirectional transformers for document summarization
Xingxing Zhang, Furu Wei, and Ming Zhou · 2019
Later among the works it cites.
ReBeL: A general game-playing ai bot that excels at poker and more
Noam Brown and Anton Bakhtin · 2020
Closest in time.
Artificial intelligence & cooperation
Elisa Bertino, Finale Doshi-Velez, Maria Gini, Daniel Lopresti, and David Parkes · 2020
Closest in time.
Artificial intelligence, algorithmic pricing, and collusion
Emilio Calvano, Giacomo Calzolari, Vincenzo Denicolo, and Sergio Pastorello · 2020
Closest in time.
Testing axioms against human reward divisions in cooperative games
Greg d’Eon and Kate Larson · 2020
Closest in time.
Two kinds of cooperative AI challenges: Game play and game design, 2020
James D. Fearon · 2020
Closest in time.
Proof-of-event recording system for autonomous vehicles: A blockchain-based solution
Hao Guo, Wanxin Li, Mark Nejad, and Chien-Chung Shen · 2020
Closest in time.
The normative infrastructure of cooperation, 2020
Gillian Hadfield · 2020
Closest in time.
Reward-rational (implicit) choice: A unifying formalism for reward learning
Hong Jun Jeon, Smitha Milli, and Anca D Dragan · 2020
Closest in time.
Specification gaming: the flip side of ai ingenuity, 2020
Victoria Krakovna, Jonathan Uesato, Vladimir Mikulik, Matthew Rahtz, Tom Everitt, Ramana Kumar, Zac Kenton, Jan Leike, and Shane Legg · 2020
Closest in time.
Gifting in multi-agent reinforcement learning
Andrei Lupu and Doina Precup · 2020
Closest in time.
Learning agent communication under limited bandwidth by message pruning
Hangyu Mao, Zhengchao Zhang, Zhen Xiao, Zhibo Gong, and Yan Ni · 2020
Closest in time.
Agreement among the states to elect the president by national popular vote, 2020
National Popular Vote Inc · 2020
Closest in time.
Learning social learning, 2020
Kamal Ndousse, Douglas Eck, Sergey Levine, and Natasha Jaques · 2020
Closest in time.
The unreasonable effectiveness of deep learning in artificial intelligence
Terrence J Sejnowski · 2020
Closest in time.
Benefits of assistance over reward learning in cooperative AI
Rohin Shah, Pedro Freire, Neel Alex, Rachel Freedman, Dmitrii Krasheninnikov, Lawrence Chan, Michael Dennis, Pieter Abbeel, Anca Dragan, and Stuart Russell · 2020
Closest in time.
Contact tracing apps can help stop coronavirus. but they can hurt privacy
Toby Shevlane, Ben Garfinkel, and Allan Dafoe · 2020
Closest in time.
Vulnerable robots positively shape human conversational dynamics in a human–robot team
Margaret L Traeger, Sarah Strohkorb Sebo, Malte Jung, Brian Scassellati, and Nicholas A Christakis · 2020
Closest in time.
Learning to interactively learn and assist
Mark Woodward, Chelsea Finn, and Karol Hausman · 2020
Closest in time.
Towards playing full MOBA games with deep reinforcement learning
Deheng Ye, Guibin Chen, Wen Zhang, Sheng Chen, Bo Yuan, Bo Liu, Jia Chen, Zhao Liu, Fuhao Qiu, Hongsheng Yu, Yinyuting Yin, Bei Shi, Liang Wang, Tengfei Shi, Qiang Fu, Wei Yang, Lanxiao Huang, and Wei Liu · 2020
Closest in time.
Computationally feasible VCG mechanisms
Noam Nisan and Amir Ronen · 2046
Closest in time.