Fetching the paper…
Reading the bibliography…
A desirable goal for autonomous agents is to be able to coordinate on the fly with previously unknown teammates.
Dynamic programming
Bellman, R. 1966 · 1966
Earlier work this paper cites.
New Metrics and Algorithms for Stochastic Goal Recognition Design Problems
Wayllace, C.; Hou, P.; and Yeoh, W. 2017 · 1966
Earlier work this paper cites.
Distributed problem-solving techniques: A survey
Decker, K. S. 1987 · 1987
Earlier work this paper cites.
On team formation
Cohen, P. R.; Levesque, H. J.; and Smith, I. A. 1997 · 1997
Earlier work this paper cites.
The communicative multiagent team decision problem: Analyzing teamwork theories and models
Pynadath, D. V.; and Tambe, M. 2002 · 2002
Earlier work this paper cites.
Decentralized control of cooperative systems: Categorization and complexity analysis
Goldman, C. V.; and Zilberstein, S. 2004 · 2004
Earlier work this paper cites.
A Decision-Theoretic Model of Assistance
Fern, A.; Natarajan, S.; Judah, K.; and Tadepalli, P. 2007 · 2007
Earlier work this paper cites.
Deployed ARMOR protection: the application of a game theoretic model for security at the Los Angeles International Airport
Pita, J.; Jain, M.; Marecki, J.; Ordóñez, F.; Portway, C.; Tambe, M.; Western, C.; Paruchuri, P.; and Kraus, S. 2008 · 2008
Earlier work this paper cites.
Ad hoc autonomous agent teams: Collaboration without pre-coordination
Stone, P.; Kaminka, G. A.; Kraus, S.; and Rosenschein, J. S. 2010 · 2010
Earlier work this paper cites.
Designing robot learners that ask good questions
Cakmak, M.; and Thomaz, A. L. 2012 · 2012
Earlier work this paper cites.
Teaching and leading an ad hoc teammate: Collaboration without pre-coordination
Stone, P.; Kaminka, G. A.; Kraus, S.; Rosenschein, J. S.; and Agmon, N. 2013 · 2013
Earlier work this paper cites.
Teaching on a budget: Agents advising agents in reinforcement learning
Torrey, L.; and Taylor, M. 2013 · 2013
Cited alongside, same era.
Communicating with unknown teammates
Barrett, S.; Agmon, N.; Hazon, N.; Kraus, S.; and Stone, P. 2014 · 2014
Cited alongside, same era.
Goal recognition design
Keren, S.; Gal, A.; and Karpas, E. 2014 · 2014
Cited alongside, same era.
Planning over multi-agent epistemic states: A classical planning approach
Muise, C.; Belle, V.; Felli, P.; McIlraith, S.; Miller, T.; Pearce, A. R.; and Sonenberg, L. 2015 · 2015
Cited alongside, same era.
OpenAI Gym
Brockman, G.; Cheung, V.; Pettersson, L.; Schneider, J.; Schulman, J.; Tang, J.; and Zaremba, W. 2016 · 2016
Cited alongside, same era.
Learning to communicate with deep multi-agent reinforcement learning
Foerster, J.; Assael, I. A.; de Freitas, N.; and Whiteson, S. 2016 · 2016
Cited alongside, same era.
coin-or/Cbc: Version 2.9.9
Forrest, J.; Ralphs, T.; Vigerske, S.; LouHafer; Kristjansson, B.; jpfasano; EdwinStraver; Lubin, M.; Santos, H. G.; rlougee; and Saltzman, M. 2018 · 2018
Later among the works it cites.
Sequential plan recognition: An iterative approach to disambiguating between hypotheses
Mirsky, R.; Stern, R.; Gal, K.; and Kalech, M. 2018 · 2018
Later among the works it cites.
Emergence of grounded compositional language in multi-agent populations
Mordatch, I.; and Abbeel, P. 2018 · 2018
Later among the works it cites.
A survey and critique of multiagent deep reinforcement learning
Hernandez-Leal, P.; Kartal, B.; and Taylor, M. E. 2019 · 2019
Later among the works it cites.
Ad hoc teamwork with behavior switching agents
Ravula, M.; Alkoby, S.; and Stone, P. 2019 · 2019
Later among the works it cites.
Query Content in Sequential One-shot Multi-Agent Limited Inquires when Communicating in Ad Hoc Teamwork
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Reasoning about hypothetical agent behaviours and their parameters
Albrecht, S. V.; and Stone, P. 2017 · 2017
Cited alongside, same era.
Coordinated vs. Decentralized Exploration In Multi-Agent Multi-Armed Bandits
Chakraborty, M.; Chua, K. Y. P.; Das, S.; and Juba, B. 2017 · 2017
Cited alongside, same era.
Autonomous agents modelling other agents: A comprehensive survey and open problems
Albrecht, S. V.; and Stone, P. 2018 · 2018
Cited alongside, same era.
Active reward learning from critiques
Cui, Y.; and Niekum, S. 2018 · 2018
Cited alongside, same era.
Macke, W.; Mirsky, R.; and Stone, P. 2020 · 2020
Later among the works it cites.
A Penny for Your Thoughts: The Value of Communication in Ad Hoc Teamwork
Mirsky, R.; Macke, W.; Wang, A.; Yedidsion, H.; and Stone, P. 2020 · 2020
Later among the works it cites.
Active Goal Recognition
Shvo, M.; and McIlraith, S. A. 2020 · 2020
Later among the works it cites.
Learning to Communicate Proactively in Human-Agent Teaming
van Zoelen, E. M.; Cremers, A.; Dignum, F. P.; van Diggelen, J.; and Peeters, M. M. 2020 · 2020
Later among the works it cites.
Too many cooks: Coordinating multi-agent collaboration through inverse planning
Wang, R. E.; Wu, S. A.; Evans, J. A.; Tenenbaum, J. B.; Parkes, D. C.; and Kleiman-Weiner, M. 2020 · 2020
Later among the works it cites.