Fetching the paper…
Reading the bibliography…
Humans can quickly adapt to new partners in collaborative tasks (e.g.
Convention: A philosophical study
David Lewis · 1969
Earlier work this paper cites.
Referring as a collaborative process
Herbert H. Clark and Deanna Wilkes-Gibbs · 1986
Earlier work this paper cites.
Evolutionary principles in self-referential learning
Jürgen Schmidhuber · 1987
Earlier work this paper cites.
The information-processing theory of mind
Herbert A Simon · 1995
Earlier work this paper cites.
Planning, learning and coordination in multiagent decision processes
Craig Boutilier · 1996
Earlier work this paper cites.
Using language
Herbert H Clark · 1996
Earlier work this paper cites.
Multitask learning
Rich Caruana · 1997
Earlier work this paper cites.
Optimal and approximate q-value functions for decentralized pomdps
Frans A Oliehoek, Matthijs TJ Spaan, and Nikos Vlassis · 2008
Earlier work this paper cites.
Maximum entropy inverse reinforcement learning
Brian D. Ziebart, Andrew L. Maas, J. Bagnell, and Anind K. Dey · 2008
Earlier work this paper cites.
Ad hoc autonomous agent teams: Collaboration without pre-coordination
Peter Stone, Gal A Kaminka, Sarit Kraus, Jeffrey S Rosenschein, et al · 2010
Earlier work this paper cites.
Human-robot cross-training: computational formulation, modeling and evaluation of a human team training strategy
Stefanos Nikolaidis and Julie Shah · 2013
Earlier work this paper cites.
Learning in the Rational Speech Acts model
Will Monroe and Christopher Potts · 2015
Earlier work this paper cites.
Pragmatic language interpretation as probabilistic inference
Noah D Goodman and Michael C Frank · 2016
Earlier work this paper cites.
Rational quantitative attribution of beliefs, desires and percepts in human mentalizing
Chris L Baker, Julian Jara-Ettinger, Rebecca Saxe, and Joshua B Tenenbaum · 2017
Cited alongside, same era.
Learning modular neural network policies for multi-task and multi-robot transfer
Coline Devin, Abhishek Gupta, Trevor Darrell, Pieter Abbeel, and Sergey Levine · 2017
Cited alongside, same era.
Model-agnostic meta-learning for fast adaptation of deep networks
Chelsea Finn, Pieter Abbeel, and Sergey Levine · 2017
Cited alongside, same era.
Convention-formation in iterated reference games
Robert XD Hawkins, Mike Frank, and Noah D Goodman · 2017
Cited alongside, same era.
Multi-agent cooperation and the emergence of (natural) language
Angeliki Lazaridou, Alexander Peysakhovich, and Marco Baroni · 2017
Cited alongside, same era.
On the utility of learning about humans for human-ai coordination
Micah Carroll, Rohin Shah, Mark K Ho, Tom Griffiths, Sanjit Seshia, Pieter Abbeel, and Anca Dragan · 2019
Later among the works it cites.
Bayesian action decoder for deep multi-agent reinforcement learning
Jakob N. Foerster, H. Francis Song, Edward Hughes, Neil Burch, Iain Dunning, Shimon Whiteson, Matthew M Botvinick, and Michael H. Bowling · 2019
Later among the works it cites.
Graphical convention formation during visual communication
Robert X. D. Hawkins, Megumi Sano, Noah D. Goodman, and Judith E. Fan · 2019
Later among the works it cites.
Simplified action decoder for deep multi-agent reinforcement learning
Hengyuan Hu and Jakob N Foerster · 2019
Later among the works it cites.
Learning existing social conventions via observationally augmented self-play
Adam Lerer and Alexander Peysakhovich · 2019
Later among the works it cites.
Emergent coordination through competition
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
John Schulman, Filip Wolski, Prafulla Dhariwal, Alec Radford, and Oleg Klimov · 2017
Cited alongside, same era.
Emergent communication through negotiation
Kris Cao, Angeliki Lazaridou, Marc Lanctot, Joel Z. Leibo, Karl Tuyls, and Stephen Clark · 2018
Cited alongside, same era.
Learning with opponent-learning awareness
Jakob Foerster, Richard Y Chen, Maruan Al-Shedivat, Shimon Whiteson, Pieter Abbeel, and Igor Mordatch · 2018
Cited alongside, same era.
Planning, inference and pragmatics in sequential language games
Fereshte Khani, Noah D Goodman, and Percy Liang · 2018
Cited alongside, same era.
Emergence of grounded compositional language in multi-agent populations
Igor Mordatch and Pieter Abbeel · 2018
Cited alongside, same era.
On first-order meta-learning algorithms
Alex Nichol, Joshua Achiam, and John Schulman · 2018
Cited alongside, same era.
Machine theory of mind
Neil C Rabinowitz, Frank Perbet, H Francis Song, Chiyuan Zhang, SM Eslami, and Matthew Botvinick · 2018
Cited alongside, same era.
Siqi Liu, Guy Lever, Josh Merel, Saran Tunyasuvunakool, Nicolas Heess, and Thore Graepel · 2019
Later among the works it cites.
Stable baselines3
Antonin Raffin, Ashley Hill, Maximilian Ernestus, Adam Gleave, Anssi Kanervisto, and Noah Dormann · 2019
Later among the works it cites.
Emergent tool use from multi-agent autocurricula
Bowen Baker, Ingmar Kanitscheider, Todor Markov, Yi Wu, Glenn Powell, Bob McGrew, and Igor Mordatch · 2020
Later among the works it cites.
The Hanabi challenge: A new frontier for AI research
Nolan Bard, Jakob N Foerster, Sarath Chandar, Neil Burch, Marc Lanctot, H Francis Song, Emilio Parisotto, Vincent Dumoulin, Subhodeep Moitra, Edward Hughes, et al · 2020
Later among the works it cites.
Signalling under uncertainty: interpretative alignment without a common prior
Thomas Brochhagen · 2020
Later among the works it cites.
"Other-play" for zero-shot coordination
Hengyuan Hu, Adam Lerer, Alex Peysakhovich, and Jakob Foerster · 2020
Later among the works it cites.
Too many cooks: Coordinating multi-agent collaboration through inverse planning
Rose E Wang, Sarah A Wu, James A Evans, Joshua B Tenenbaum, David C Parkes, and Max Kleiman-Weiner · 2020
Later among the works it cites.