Fetching the paper…
Reading the bibliography…
Large language models (LLMs) such as ChatGPT and GPT-4 have recently demonstrated their remarkable abilities of communicating with human users.
Deep blue
Campbell, M., Hoane Jr, A. J., and Hsu, F.-h · 2002
Earlier work this paper cites.
Probabilistic robotics (intelligent robotics and autonomous agents) , 2005
Thrun, S., Burgard, W., and Fox, D · 2005
Earlier work this paper cites.
Deep residual learning for image recognition
He, K., Zhang, X., Ren, S., and Sun, J · 2016
Earlier work this paper cites.
Mastering the game of go with deep neural networks and tree search
Silver, D., Huang, A., Maddison, C. J., Guez, A., Sifre, L., Van Den Driessche, G., Schrittwieser, J., Antonoglou, I., Panneershelvam, V., Lanctot, M., et al · 2016
Earlier work this paper cites.
Deepstack: Expert-level artificial intelligence in heads-up no-limit poker
Moravčík, M., Schmid, M., Burch, N., Lisỳ, V., Morrill, D., Bard, N., Davis, T., Waugh, K., Johanson, M., and Bowling, M · 2017
Earlier work this paper cites.
Superhuman ai for heads-up no-limit poker: Libratus beats top professionals
Brown, N. and Sandholm, T · 2018
Earlier work this paper cites.
Recurrent world models facilitate policy evolution
Ha, D. and Schmidhuber, J · 2018
Cited alongside, same era.
Alphastar: An evolutionary computation perspective
Arulkumaran, K., Cully, A., and Togelius, J · 2019
Cited alongside, same era.
Dota 2 with large scale deep reinforcement learning
Berner, C., Brockman, G., Chan, B., Cheung, V., Dębiak, P., Dennison, C., Farhi, D., Fischer, Q., Hashme, S., Hesse, C., et al · 2019
Cited alongside, same era.
Nail: A general interactive fiction agent
Hausknecht, M., Loynd, R., Yang, G., Swaminathan, A., and Williams, J. D · 2019
Cited alongside, same era.
Graph constrained reinforcement learning for natural language action spaces
Ammanabrolu, P. and Hausknecht, M · 2020
Cited alongside, same era.
Interactive fiction game playing as multi-paragraph reading comprehension with reinforcement learning
Guo, X., Yu, M., Gao, Y., Gan, C., Campbell, M., and Chang, S · 2020
Later among the works it cites.
Interactive fiction games: A colossal adventure
Hausknecht, M., Ammanabrolu, P., Côté, M.-A., and Yuan, X · 2020
Later among the works it cites.
On the dangers of stochastic parrots: Can language models be too big?
Bender, E. M., Gebru, T., McMillan-Major, A., and Shmitchell, S · 2021
Later among the works it cites.
Deep learning, reinforcement learning, and world models
Matsuo, Y., LeCun, Y., Sahani, M., Precup, D., Silver, D., Sugiyama, M., Uchibe, E., and Morimoto, J · 2022
Later among the works it cites.
Sparks of artificial general intelligence: Early experiments with gpt-4
Bubeck, S., Chandrasekaran, V., Eldan, R., Gehrke, J., Horvitz, E., Kamar, E., Lee, P., Lee, Y. T., Li, Y., Lundberg, S., et al · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Climbing towards nlu: On meaning, form, and understanding in the age of data
Bender, E. M. and Koller, A · 2020
Cited alongside, same era.
OpenAI · 2023
Closest in time.