Fetching the paper…
Reading the bibliography…
In order to train a computer agent to play a text-based computer game, we must represent each hidden state of the game.
P. J. Huber, “Robust estimation of a location parameter,” Ann. Math. Statist. , vol. 35, no. 1, pp. 73–101, 03 1964. [Online]. Available: https://doi.org/10.1214/aoms/1177703732
1964
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,” Neural computation , vol. 9, pp. 1735–80, 12 1997
1997
Earlier work this paper cites.
Infocom, Zork I Manual , 2001, http://infodoc.plover.net/manuals/zork1.pdf
2001
Earlier work this paper cites.
T. Mikolov, I. Sutskever, K. Chen, G. S. Corrado, and J. Dean, “Distributed representations of words and phrases and their compositionality,” in Advances in Neural Information Processing Systems 26 , C. J. C. Burges, L. Bottou, M. Welling, Z. Ghahramani, and K. Q. Weinberger, Eds. Curran Associates, Inc., 2013, pp. 3111–3119
2013
Earlier work this paper cites.
Y. Kim, “Convolutional neural networks for sentence classification,” in Proceedings of the 2014 Conference on Empirical Methods in Natural Language Processing (EMNLP) . Association for Computational Linguistics, 2014, pp. 1746–1751. [Online]. Available: http://aclweb.org/anthology/D14-1181
2014
Earlier work this paper cites.
D. Chen and C. Manning, “A fast and accurate dependency parser using neural networks,” in Proceedings of the 2014 Conference on Empirical Methods in Natural Language Processing (EMNLP) . Doha, Qatar: Association for Computational Linguistics, Oct. 2014, pp. 740–750. [Online]. Available: https://www.aclweb.org/anthology/D14-1082
2014
Earlier work this paper cites.
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, S. Petersen, C. Beattie, A. Sadik, I. Antonoglou, H. King, D. Kumaran, D. Wierstra, S. Legg, and D. Hassabis, “Human-level control through deep reinforcement learning,” Nature , vol. 518, pp. 529 EP –, 02 2015. [Online]. Available: https://doi.org/10.1038/nature14236
2015
Earlier work this paper cites.
K. Narasimhan, T. Kulkarni, and R. Barzilay, “Language understanding for text-based games using deep reinforcement learning,” in Proceedings of the 2015 Conference on Empirical Methods in Natural Language Processing . Association for Computational Linguistics, 2015, pp. 1–11. [Online]. Available: http://aclweb.org/anthology/D15-1001
2015
Cited alongside, same era.
T. Luong, H. Pham, and C. D. Manning, “Effective approaches to attention-based neural machine translation,” in Proceedings of the 2015 Conference on Empirical Methods in Natural Language Processing . Association for Computational Linguistics, 2015, pp. 1412–1421. [Online]. Available: http://aclweb.org/anthology/D15-1166
2015
Cited alongside, same era.
J. He, J. Chen, X. He, J. Gao, L. Li, L. Deng, and M. Ostendorf, “Deep reinforcement learning with a natural language action space,” in Proceedings of the 54th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . Association for Computational Linguistics, 2016, pp. 1621–1630. [Online]. Available: http://aclweb.org/anthology/P16-1153
2016
Cited alongside, same era.
B. Kostka, J. Kwiecieli, J. Kowalski, and P. Rychlikowski, “Text-based adventures of the golovin AI agent,” in CIG . IEEE, 2017, pp. 181–188
2017
Later among the works it cites.
J. Gehring, M. Auli, D. Grangier, D. Yarats, and Y. N. Dauphin, “Convolutional sequence to sequence learning,” in Proceedings of the 34th International Conference on Machine Learning , ser. Proceedings of Machine Learning Research, D. Precup and Y. W. Teh, Eds., vol. 70. International Convention Centre, Sydney, Australia: PMLR, 06–11 Aug 2017, pp. 1243–1252. [Online]. Available: http://proceedings.mlr.press/v70/gehring17a.html
2017
Later among the works it cites.
T. Zahavy, M. Haroush, N. Merlis, D. J. Mankowitz, and S. Mannor, “Learn what not to learn: Action elimination with deep reinforcement learning,” in Advances in Neural Information Processing Systems 31 , S. Bengio, H. Wallach, H. Larochelle, K. Grauman, N. Cesa-Bianchi, and R. Garnett, Eds. Curran Associates, Inc., 2018, pp. 3562–3573
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
T. Schaul, J. Quan, I. Antonoglou, and D. Silver, “Prioritized experience replay,” in International Conference on Learning Representations , Puerto Rico, 2016
2016
Cited alongside, same era.
J. Li, W. Monroe, A. Ritter, D. Jurafsky, M. Galley, and J. Gao, “Deep reinforcement learning for dialogue generation,” in Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing . Austin, Texas: Association for Computational Linguistics, Nov. 2016, pp. 1192–1202. [Online]. Available: https://www.aclweb.org/anthology/D16-1127
2016
Cited alongside, same era.
N. Fulda, D. Ricks, B. Murdoch, and D. Wingate, “What can you do with a rock? affordance extraction via word embeddings.” in IJCAI , C. Sierra, Ed. ijcai.org, 2017, pp. 1039–1045. [Online]. Available: http://dblp.uni-trier.de/db/conf/ijcai/ijcai2017.html#FuldaRMW17
2017
Cited alongside, same era.
M. Norton, Zork Transcript , http://steel.lcc.gatech.edu/~marleigh/zork/transcript.html
Cited in the paper.
2018
Later among the works it cites.
2018
Later among the works it cites.
2018
Later among the works it cites.