Evaluating state-space abstractions in extensive-form games
Johanson, M., Burch, N., Valenzano, R., and Bowling, M. (2013) · 2013
Later among the works it cites.
Further Limit Hold ’em: Exploring the Model Poker Game
Newall, P. (2013) · 2013
Later among the works it cites.
Exponential reservoir sampling for streaming language models
Osborne, M., Lall, A., and Van Durme, B. (2014) · 2014
Later among the works it cites.
Tactex’13: a champion adaptive power trading agent
Urieli, D. and Stone, P. (2014) · 2014
Later among the works it cites.
Heads-up limit hold’em poker is solved
Bowling, M., Burch, N., Johanson, M., and Tammelin, O. (2015) · 2015
Later among the works it cites.
Training deep convolutional neural networks to play go
Clark, C. and Storkey, A. (2015) · 2015
Later among the works it cites.
Fictitious self-play in extensive-form games
Heinrich, J., Lanctot, M., and Silver, D. (2015) · 2015
Later among the works it cites.
Smooth UCT search in computer poker
Heinrich, J. and Silver, D. (2015) · 2015
Later among the works it cites.
Continuous control with deep reinforcement learning
Original
Lillicrap, T. P., Hunt, J. J., Pritzel, A., Heess, N., Erez, T., Tassa, Y., Silver, D., and Wierstra, D. (2015) · 2015
Later among the works it cites.
Move evaluation in go using deep convolutional neural networks
Maddison, C. J., Huang, A., Sutskever, I., and Silver, D. (2015) · 2015
Later among the works it cites.
Human-level control through deep reinforcement learning
Mnih, V., Kavukcuoglu, K., Silver, D., Rusu, A. A., Veness, J., Bellemare, M. G., Graves, A., Riedmiller, M., Fidjeland, A. K., Ostrovski, G., et al. (2015) · 2015
Later among the works it cites.
Solving games with functional regret estimation
Waugh, K., Morrill, D., Bagnell, J. A., and Bowling, M. (2015) · 2015
Later among the works it cites.
Poker-cnn: A pattern learning strategy for making draws and bets in poker games using convolutional networks
Yakovenko, N., Cao, L., Raffel, C., and Fan, J. (2016) · 2016
Closest in time.