Fetching the paper…
Reading the bibliography…
Researchers on artificial intelligence have achieved human-level intelligence in large-scale perfect-information games, but it is still a challenge to achieve (nearly) optimal results (in other words, an approximate Nash Equilibrium) in large-scale imperfect-information games (i.e.
Brown, G.W.: Iterative solution of games by fictitious play. Activity analysis of production and allocation (1951)
1951
Earlier work this paper cites.
Nash, J.: Non-cooperative games. Annals of mathematics pp. 286–295 (1951)
1951
Earlier work this paper cites.
Robinson, J.: An iterative method of solving a game. Annals of Mathematics pp. 296–301 (1951)
1951
Earlier work this paper cites.
Vitter, J.S.: Random sampling with a reservoir. ACM Transactions on Mathematical Software (1985)
1985
Earlier work this paper cites.
Myerson, R.B.: Game Theory:Analysis of Conflict. Harvard University Press (1991)
1991
Earlier work this paper cites.
Sutton, R.S., Barto, A.G.: Reinforcement learning: An introduction, vol. 1 (1998)
1998
Earlier work this paper cites.
Lavi, R.: Algorithmic game theory. Computationally-efficient approximate mechanisms pp. 301–330 (2007)
2007
Earlier work this paper cites.
Sanholm, T.: The state of solving large incomplete-information games, and application to poker. AI Magazine 31
2010
Earlier work this paper cites.
Browne, C.B., Powley, E., Whitehouse, D., Lucas, S.M., Cowling, P.I., Rohlfshagen, P., Tavener, S., Perez, D., Samothrakis, S., Colton, S.: A survey of monte carlo tree search methods. IEEE Transactions on Computational Intelligence and AI in Games 4
2012
Cited alongside, same era.
Bosansky, B., Kiekintveld, C., Lisy, V., Pechoucek, M.: An exact double-oracle algorithm for zero-sum extensive-form games with imperfect information. Journal of Artificial Intelligence Research pp. 829–866 (2014)
2014
Cited alongside, same era.
Heinrich, J., Lanctot, M., Silver, D.: Fictitious self-play in extensive-form games. In: Proceedings of the 32nd International Conference on Machine Learning. (2015)
2015
Cited alongside, same era.
Heinrich, J., Silver, D.: Smooth uct search in computer poker. In: Twenty-Fourth International Joint Conference on Artificial Intelligence (2015)
2015
Cited alongside, same era.
Silver, D., Huang, A., Maddison, C.J., Guez, A., Sifre, L., Van Den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V., Lanctot, M., et al.: Mastering the game of go with deep neural networks and tree search. nature 529
2016
Later among the works it cites.
Sukhbaatar, S., Fergus, R., et al.: Learning multiagent communication with backpropagation. In: Advances in Neural Information Processing Systems. pp. 2244–2252 (2016)
2016
Later among the works it cites.
Brown, N., Sandholm, T.: Libratus: The superhuman ai for no-limit poker. In: IJCAI. pp. 5226–5228 (2017)
2017
Later among the works it cites.
Moravčík, M., Schmid, M., Burch, N., Lisỳ, V., Morrill, D., Bard, N., Davis, T., Waugh, K., Johanson, M., Bowling, M.: Deepstack: Expert-level artificial intelligence in heads-up no-limit poker. Science 356
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Lisỳ, V., Lanctot, M., Bowling, M.: Online monte carlo counterfactual regret minimization for search in imperfect information games. In: Proceedings of the 2015 International Conference on Autonomous Agents and Multiagent Systems. pp. 27–36. International Foundation for Autonomous Agents and Multiagent Systems (2015)
2015
Cited alongside, same era.
Mnih, V., Kavukcuoglu, K., Silver, D., Rusu, A.A., Veness, J., Bellemare, M.G., Graves, A., Riedmiller, M., Fidjeland, A.K., Ostrovski, G., Petersen, S., Beattie, C., Sadik, A., Antonoglou, I., King, H., Kumaran, D., Wierstra, D., Legg, S., Hassabis, D.: Human-level control through deep reinforcement learning. Nature pp. 529–533 (2015)
2015
Cited alongside, same era.
2016
Cited alongside, same era.
Mnih, V., Badia, A.P., Mirza, M., Graves, A., Lillicrap, T.P., Harley, T., Silver, D., Kavukcuoglu, K.: Asynchronous methods for deep reinforcement learning. In: Proceedings of Machine Learning Research (2016)
2016
Cited alongside, same era.
2017
Later among the works it cites.
Silver, D., Schrittwieser, J., Simonyan, K., Antonoglou, I., Huang, A., Guez, A., Hubert, T., Baker, L., Lai, M., Bolton, A., et al.: Mastering the game of go without human knowledge. Nature 550
2017
Later among the works it cites.
Vinyals, O., Babuschkin, I., Chung, J., Mathieu, M., Jaderberg, M., Czarnecki, W.M., Dudzik, A., Huang, A., Georgiev, P., Powell, R., Ewalds, T., Horgan, D., Kroiss, M., Danihelka, I., Agapiou, J., Oh, J., Dalibard, V., Choi, D., Sifre, L., Sulsky, Y., Vezhnevets, S., Molloy, J., Cai, T., Budden, D., Paine, T., Gulcehre, C., Wang, Z., Pfaff, T., Pohlen, T., Yogatama, D., Cohen, J., McKinney, K., Smith, O., Schaul, T., Lillicrap, T., Apps, C., Kavukcuoglu, K., Hassabis, D., Silver, D.: AlphaStar: Mastering the Real-Time Strategy Game StarCraft II. https://deepmind.com/blog/alphastar-mastering-real-time-strategy-game-starcraft-ii/ (2019)
2019
Closest in time.