Fetching the paper…
Reading the bibliography…
Humans observe and interact with the world to acquire knowledge.
Transfer in deep reinforcement learning using knowledge graphs
Ammanabrolu, P. and Riedl, M. (2019b) · 1908
Earlier work this paper cites.
Yin, X. and May, J. (2019) · 1908
Earlier work this paper cites.
Q-learning
Watkins, C. J. C. H. and Dayan, P. (1992) · 1992
Earlier work this paper cites.
Planning and acting in partially observable stochastic domains
Kaelbling, L. P., Littman, M. L., and Cassandra, A. R. (1998) · 1998
Earlier work this paper cites.
Seeking Meaning: A Process Approach to Library and Information Services
Kuhlthau, C. (2004) · 2004
Earlier work this paper cites.
Formal theory of creativity, fun, and intrinsic motivation (1990–2010)
Schmidhuber, J. (2010) · 2010
Earlier work this paper cites.
Semantic parsing on freebase from question-answer pairs
Berant, J., Chou, A., Frostig, R., and Liang, P. (2013) · 2013
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, D. P. and Ba, J. (2014) · 2014
Earlier work this paper cites.
Human-level control through deep reinforcement learning
Mnih, V., Kavukcuoglu, K., Silver, D., Rusu, A. A., Veness, J., Bellemare, M. G., Graves, A., Riedmiller, M., Fidjeland, A. K., Ostrovski, G., et al. (2015) · 2015
Earlier work this paper cites.
Language understanding for text-based games using deep reinforcement learning
Narasimhan, K., Kulkarni, T., and Barzilay, R. (2015) · 2015
Earlier work this paper cites.
Srivastava, R. K., Greff, K., and Schmidhuber, J. (2015) · 2015
Earlier work this paper cites.
Towards ai-complete question answering: A set of prerequisite toy tasks
Weston, J., Bordes, A., Chopra, S., Rush, A. M., van Merriënboer, B., Joulin, A., and Mikolov, T. (2015) · 2015
Earlier work this paper cites.
Towards information-seeking agents
Bachman, P., Sordoni, A., and Trischler, A. (2016) · 2016
Earlier work this paper cites.
MS MARCO: A human generated machine reading comprehension dataset
Nguyen, T., Rosenberg, M., Song, X., Gao, J., Tiwary, S., Majumder, R., and Deng, L. (2016) · 2016
Earlier work this paper cites.
Squad: 100, 000+ questions for machine comprehension of text
Rajpurkar, P., Zhang, J., Lopyrev, K., and Liang, P. (2016) · 2016
Earlier work this paper cites.
Newsqa: A machine comprehension dataset
Trischler, A., Wang, T., Yuan, X., Harris, J., Sordoni, A., Bachman, P., and Suleman, K. (2016) · 2016
Earlier work this paper cites.
Machine comprehension using match-lstm and answer pointer
Wang, S. and Jiang, J. (2016) · 2016
Cited alongside, same era.
Home: a household multimodal environment
Brodeur, S., Perez, E., Anand, A., Golemo, F., Celotti, L., Strub, F., Rouat, J., Larochelle, H., and Courville, A. C. (2017) · 2017
Cited alongside, same era.
Reading wikipedia to answer open-domain questions
Chen, D., Fisch, A., Weston, J., and Bordes, A. (2017) · 2017
Cited alongside, same era.
Das, A., Datta, S., Gkioxari, G., Lee, S., Parikh, D., and Batra, D. (2017) · 2017
Cited alongside, same era.
Searchqa: A new q&a dataset augmented with context from a search engine
Quac : Question answering in context
Choi, E., He, H., Iyyer, M., Yatskar, M., Yih, W., Choi, Y., Liang, P., and Zettlemoyer, L. (2018) · 2018
Later among the works it cites.
Think you have solved question answering? try arc, the AI2 reasoning challenge
Clark, P., Cowhey, I., Etzioni, O., Khot, T., Sabharwal, A., Schoenick, C., and Tafjord, O. (2018) · 2018
Later among the works it cites.
Quantifying generalization in reinforcement learning
Cobbe, K., Klimov, O., Hesse, C., Kim, T., and Schulman, J. (2018) · 2018
Later among the works it cites.
Textworld: A learning environment for text-based games
Côté, M.-A., Kádár, A., Yuan, X., Kybartas, B., Barnes, T., Fine, E., Moore, J., Hausknecht, M., Asri, L. E., Adada, M., Tay, W., and Trischler, A. (2018) · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Dunn, M., Sagun, L., Higgins, M., Güney, V. U., Cirik, V., and Cho, K. (2017) · 2017
Cited alongside, same era.
Noisy networks for exploration
Fortunato, M., Azar, M. G., Piot, B., Menick, J., Osband, I., Graves, A., Mnih, V., Munos, R., Hassabis, D., Pietquin, O., Blundell, C., and Legg, S. (2017) · 2017
Cited alongside, same era.
IQA: visual question answering in interactive environments
Gordon, D., Kembhavi, A., Rastegari, M., Redmon, J., Fox, D., and Farhadi, A. (2017) · 2017
Cited alongside, same era.
Rainbow: Combining improvements in deep reinforcement learning
Hessel, M., Modayil, J., van Hasselt, H., Schaul, T., Ostrovski, G., Dabney, W., Horgan, D., Piot, B., Azar, M. G., and Silver, D. (2017) · 2017
Cited alongside, same era.
Adversarial examples for evaluating reading comprehension systems
Jia, R. and Liang, P. (2017) · 2017
Cited alongside, same era.
Triviaqa: A large scale distantly supervised challenge dataset for reading comprehension
Joshi, M., Choi, E., Weld, D. S., and Zettlemoyer, L. (2017) · 2017
Cited alongside, same era.
Automatic differentiation in pytorch
Paszke, A., Gross, S., Chintala, S., Chanan, G., Yang, E., DeVito, Z., Lin, Z., Desmaison, A., Antiga, L., and Lerer, A. (2017) · 2017
Cited alongside, same era.
World of bits: An open-domain platform for web-based agents
Shi, T., Karpathy, A., Fan, L., Hernandez, J., and Liang, P. (2017) · 2017
Cited alongside, same era.
Das, R., Munkhdalai, T., Yuan, X., Trischler, A., and McCallum, A. (2018) · 2018
Later among the works it cites.
BERT: pre-training of deep bidirectional transformers for language understanding
Devlin, J., Chang, M., Lee, K., and Toutanova, K. (2018) · 2018
Later among the works it cites.
Procedural level generation improves generality of deep reinforcement learning
Justesen, N., Torrado, R. R., Bontrager, P., Khalifa, A., Togelius, J., and Risi, S. (2018) · 2018
Later among the works it cites.
Advances in pre-training distributed word representations
Mikolov, T., Grave, E., Bojanowski, P., Puhrsch, C., and Joulin, A. (2018) · 2018
Later among the works it cites.
Coqa: A conversational question answering challenge
Reddy, S., Chen, D., and Manning, C. D. (2018) · 2018
Later among the works it cites.
Does it care what you asked? understanding importance of verbs in deep learning qa system
Rychalska, B., Basaj, D., Biecek, P., and Wroblewska, A. (2018) · 2018
Later among the works it cites.
The web as a knowledge-base for answering complex questions
Talmor, A. and Berant, J. (2018) · 2018
Later among the works it cites.
Hotpotqa: A dataset for diverse, explainable multi-hop question answering
Yang, Z., Qi, P., Zhang, S., Bengio, Y., Cohen, W. W., Salakhutdinov, R., and Manning, C. D. (2018) · 2018
Later among the works it cites.
Qanet: Combining local convolution with global self-attention for reading comprehension
Yu, A. W., Dohan, D., Luong, M.-T., Zhao, R., Chen, K., Norouzi, M., and Le, Q. V. (2018) · 2018
Later among the works it cites.
Counting to explore and generalize in text-based games
Yuan, X., Côté, M., Sordoni, A., Laroche, R., des Combes, R. T., Hausknecht, M. J., and Trischler, A. (2018) · 2018
Later among the works it cites.
Nail: A general interactive fiction agent
Hausknecht, M., Loynd, R., Yang, G., Swaminathan, A., and Williams, J. D. (2019) · 2019
Closest in time.
Natural questions: a benchmark for question answering research
Kwiatkowski, T., Palomaki, J., Redfield, O., Collins, M., Parikh, A., Alberti, C., Epstein, D., Polosukhin, I., Kelcey, M., Devlin, J., Lee, K., Toutanova, K. N., Jones, L., Chang, M.-W., Dai, A., Uszkoreit, J., Le, Q., and Petrov, S. (2019) · 2019
Closest in time.