Fetching the paper…
Reading the bibliography…
Recent years have witnessed the great breakthrough of deep reinforcement learning (DRL) in various perfect and imperfect information games.
V. Mnih, A. P. Badia, M. Mirza, A. Graves, T. Lillicrap, T. Harley, D. Silver, and K. Kavukcuoglu, “Asynchronous methods for deep reinforcement learning,” in International Conference on Machine Learning (ICML) , 2016, pp. 1928–1937
1937
Earlier work this paper cites.
D. Carmel and S. Markovitch, “Model-based learning of interaction strategies in multi-agent systems,” Journal of Experimental & Theoretical Artificial Intelligence , vol. 10, no. 3, pp. 309–332, 1998
1998
Earlier work this paper cites.
M. Fagan and P. Cunningham, “Case-based plan recognition in computer games,” in International Conference on Case-Based Reasoning , 2003, pp. 161–170
2003
Earlier work this paper cites.
M. Zinkevich, M. Johanson, M. Bowling, and C. Piccione, “Regret minimization in games with incomplete information,” Advances in Neural Information Processing Systems (NeurIPS) , vol. 20, pp. 1729–1736, 2007
2007
Earlier work this paper cites.
F. Schadd, S. Bakkes, and P. Spronck, “Opponent modeling in real-time strategy games.” in GAMEON , 2007, pp. 61–70
2007
Earlier work this paper cites.
M. Johanson, K. Waugh, M. Bowling, and M. Zinkevich, “Accelerating best response calculation in large extensive games,” in International Joint Conferences on Artificial Intelligence (IJCAI) , 2011
2011
Earlier work this paper cites.
H. Zhang, G. Gao, W. Li, C. Zhong, W. Yu, and C. Wang, “Botzone: A game playing system for artificial intelligence education,” in International Conference on Frontiers in Education: Computer Science and Computer Engineering (FECS) , 2012, p. 1
2012
Earlier work this paper cites.
N. Sweeney and D. Sinclair, “Applying reinforcement learning to poker,” in Computer Poker Symposium , 2012
2012
Earlier work this paper cites.
2012
Earlier work this paper cites.
L. F. Teófilo, N. Passos, L. P. Reis, and H. L. Cardoso, “Adapting strategies to opponent models in incomplete information games: a reinforcement learning approach for poker,” in International Conference on Autonomous and Intelligent Systems , 2012, pp. 220–227
2012
Earlier work this paper cites.
T. W. Neller and M. Lanctot, “An introduction to counterfactual regret minimization,” in Educational Advances in Artificial Intelligence (EAAI) , vol. 11, 2013
2013
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski et al. , “Human-level control through deep reinforcement learning,” Nature , vol. 518, no. 7540, pp. 529–533, 2015
2015
Earlier work this paper cites.
N. Mizukami and Y. Tsuruoka, “Building a computer mahjong player based on monte carlo simulation and opponent models,” in IEEE Conference on Computational Intelligence and Games (CIG) , 2015, pp. 275–283
2015
Earlier work this paper cites.
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot et al. , “Mastering the game of go with deep neural networks and tree search,” Nature , vol. 529, no. 7587, pp. 484–489, 2016
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
H. He, J. Boyd-Graber, K. Kwok, and H. Daumé III, “Opponent modeling in deep reinforcement learning,” in International Conference on Machine Learning (ICML) , 2016, pp. 1804–1813
2016
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , 2016, pp. 770–778
2016
Cited alongside, same era.
D. Silver, J. Schrittwieser, K. Simonyan, I. Antonoglou, A. Huang, A. Guez, T. Hubert, L. Baker, M. Lai, A. Bolton et al. , “Mastering the game of go without human knowledge,” Nature , vol. 550, no. 7676, pp. 354–359, 2017
2017
Cited alongside, same era.
M. Moravčík, M. Schmid, N. Burch, V. Lisỳ, D. Morrill, N. Bard, T. Davis, K. Waugh, M. Johanson, and M. Bowling, “Deepstack: Expert-level artificial intelligence in heads-up no-limit poker,” Science , vol. 356, no. 6337, pp. 508–513, 2017
2017
Cited alongside, same era.
S. J. Knegt, M. M. Drugan, and M. A. Wiering, “Opponent modelling in the game of tron using reinforcement learning.” in International Conference on Agents and Artificial Intelligence (ICAART) , 2018, pp. 29–40
2018
Later among the works it cites.
——, “Superhuman ai for multiplayer poker,” Science , vol. 365, no. 6456, pp. 885–890, 2019
2019
Later among the works it cites.
O. Vinyals, I. Babuschkin, W. M. Czarnecki, M. Mathieu, A. Dudzik, J. Chung, D. H. Choi, R. Powell, T. Ewalds, P. Georgiev et al. , “Grandmaster level in starcraft II using multi-agent reinforcement learning,” Nature , vol. 575, no. 7782, pp. 350–354, 2019
2019
Later among the works it cites.
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
H. Zhou, Y. Zhou, H. Zhang, H. Huang, and W. Li, “Botzone: A competitive and interactive platform for game ai education,” in Proceedings of the ACM Turing 50th Celebration Conference-China , 2017, pp. 1–5
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
D. Silver, T. Hubert, J. Schrittwieser, I. Antonoglou, M. Lai, A. Guez, M. Lanctot, L. Sifre, D. Kumaran, T. Graepel et al. , “A general reinforcement learning algorithm that masters chess, shogi, and go through self-play,” Science , vol. 362, no. 6419, pp. 1140–1144, 2018
2018
Cited alongside, same era.
N. Brown and T. Sandholm, “Superhuman ai for heads-up no-limit poker: Libratus beats top professionals,” Science , vol. 359, no. 6374, pp. 418–424, 2018
2018
Cited alongside, same era.
2018
Cited alongside, same era.
Q. Jiang, K. Li, B. Du, H. Chen, and H. Fang, “Deltadou: Expert-level doudizhu ai through self-play.” in International Joint Conferences on Artificial Intelligence (IJCAI) , 2019, pp. 1265–1271
2019
Later among the works it cites.
D.-W. Kim, S. Park, and S.-i. Yang, “Mastering fighting game using deep reinforcement learning with self-play,” in IEEE Conference on Games (CoG) , 2020, pp. 576–583
2020
Later among the works it cites.
2020
Later among the works it cites.
Y. You, L. Li, B. Guo, W. Wang, and C. Lu, “Combinatorial q-learning for dou di zhu,” in AAAI Conference on Artificial Intelligence and Interactive Digital Entertainment , vol. 16, no. 1, 2020, pp. 301–307
2020
Later among the works it cites.
D. Ye, Z. Liu, M. Sun, B. Shi, P. Zhao, H. Wu, H. Yu, S. Yang, X. Wu, Q. Guo et al. , “Mastering complex control in moba games with deep reinforcement learning.” in AAAI Conference on Artificial Intelligence (AAAI) , 2020
2020
Later among the works it cites.
2020
Later among the works it cites.
B. Bouzy, A. Rimbaud, and V. Ventos, “Recursive monte carlo search for bridge card play,” in IEEE Conference on Games (CoG) , 2020, pp. 229–236
2020
Later among the works it cites.
S. Ariyurek, A. Betin-Can, and E. Surer, “Enhancing the monte carlo tree search algorithm for video game testing,” in IEEE Conference on Games (CoG) , 2020, pp. 25–32
2020
Later among the works it cites.
T. Cazenave, “Improving model and search for computer go,” in IEEE Conference on Games (CoG) , 2021, pp. 1–8
2021
Later among the works it cites.
2021
Later among the works it cites.
J. Zhou, “Design and application of tibetan long chess using monte carlo algorithm and artificial intelligence,” in Journal of Physics: Conference Series , vol. 1952, no. 4, 2021, p. 042104
2021
Later among the works it cites.