Fetching the paper…
Reading the bibliography…
The ad hoc coordination problem is to design an autonomous agent which is able to achieve optimal flexibility and efficiency in a multiagent system with no mechanisms for prior coordination.
Iterative solution of games by fictitious play
Brown, G. (1951) · 1951
Earlier work this paper cites.
Stochastic games
Shapley, L. (1953) · 1953
Earlier work this paper cites.
Dynamic Programming
Bellman, R. (1957) · 1957
Earlier work this paper cites.
Games with incomplete information played by “Bayesian” players. Part I. The basic model
Harsanyi, J. (1967) · 1967
Earlier work this paper cites.
Games with incomplete information played by “Bayesian” players. Part II. Bayesian equilibrium points
Harsanyi, J. (1968) · 1968
Earlier work this paper cites.
Generation of random sequences by human subjects: A critical survey of literature
Wagenaar, W. (1972) · 1972
Earlier work this paper cites.
Perfect Bayesian equilibrium and sequential equilibrium
Fudenberg, D. and Tirole, J. (1991) · 1991
Earlier work this paper cites.
Bayesian learning in normal form games
Jordan, J. (1991) · 1991
Earlier work this paper cites.
Q-learning
Watkins, C. and Dayan, P. (1992) · 1992
Earlier work this paper cites.
Theory of Moves
Brams, S. (1993) · 1993
Cited alongside, same era.
Rational learning leads to Nash equilibrium
Kalai, E. and Lehrer, E. (1993) · 1993
Cited alongside, same era.
The dynamics of reinforcement learning in cooperative multiagent systems
Claus, C. and Boutilier, C. (1998) · 1998
Cited alongside, same era.
Reinforcement learning: An introduction
Sutton, R. and Barto, A. (1998) · 1998
Cited alongside, same era.
A sparse sampling algorithm for near-optimal planning in large Markov decision processes
Kearns, M., Mansour, Y., and Ng, A. (1999) · 1999
Cited alongside, same era.
Rational coordination in multi-agent environments
Gmytrasiewicz, P. and Durfee, E. (2000) · 2000
Cited alongside, same era.
Coordination and adaptation in impromptu teams
Bowling, M. and McCracken, P. (2005) · 2005
Later among the works it cites.
A framework for sequential planning in multiagent settings
Gmytrasiewicz, P. and Doshi, P. (2005) · 2005
Later among the works it cites.
Dynamically formed human-robot teams performing coordinated tasks
Dias, M., Harris, T., Browning, B., Jones, E., Argall, B., Veloso, M., Stentz, A., and Rudnicky, A. (2006) · 2006
Later among the works it cites.
Reaching pareto-optimality in prisoner’s dilemma using conditional joint action learning
Banerjee, D. and Sen, S. (2007) · 2007
Later among the works it cites.
To teach or not to teach? Decision making under uncertainty in ad hoc teams
Stone, P. and Kraus, S. (2010) · 2010
Later among the works it cites.
Empirical evaluation of ad hoc teamwork in the pursuit domain
Barrett, S., Stone, P., and Kraus, S. (2011) · 2011
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Multiagent learning using a variable learning rate
Bowling, M. and Veloso, M. (2002) · 2002
Cited alongside, same era.
Recognizing and predicting agent behavior with case based reasoning
Wendler, J. and Bach, J. (2004) · 2003
Cited alongside, same era.
Learning to play Bayesian games
Dekel, E., Fudenberg, D., and Levine, D. (2004) · 2004
Cited alongside, same era.
Ad hoc autonomous agent teams: Collaboration without pre-coordination
Stone, P., Kaminka, G., Kraus, S., and Rosenschein, J. (2010a)
Cited in the paper.
Leading a best-response teammate in an ad hoc team
Stone, P., Kaminka, G., and Rosenschein, J. (2010b)
Cited in the paper.
Later among the works it cites.
Leading ad hoc agents in joint action settings with multiple teammates
Agmon, N. and Stone, P. (2012) · 2012
Later among the works it cites.
Comparative evaluation of MAL algorithms in a diverse set of ad hoc team problems
Albrecht, S. and Ramamoorthy, S. (2012) · 2012
Later among the works it cites.