Fetching the paper…
Reading the bibliography…
Given a simple request like Put a washed apple in the kitchen fridge, humans can reason in purely abstract terms by imagining action sequences and scoring their likelihood of success, prototypicality, and efficiency, all without moving a muscle.
Distilbert, a distilled version of bert: smaller, faster, cheaper and lighter
Sanh, V., Debut, L., Chaumond, J., and Wolf, T. (2019) · 1910
Earlier work this paper cites.
Speech understanding systems: A summary of results of the five-year research effort
Reddy, D. R. et al. (1977) · 1977
Earlier work this paper cites.
Pddl-the planning domain definition language
McDermott, D., Ghallab, M., Howe, A., Knoblock, C., Ram, A., Veloso, M., Weld, D., and Wilkins, D. (1998) · 1998
Earlier work this paper cites.
Unnatural language processing: Bridging the gap between synthetic and natural language data
Marzoev, A., Madden, S., Kaashoek, M. F., Cafarella, M., and Andreas, J. (2020) · 2004
Earlier work this paper cites.
The Fast Downward planning system
Helmert, M. (2006) · 2006
Earlier work this paper cites.
Walk the talk: Connecting language, knowledge, and action in route instructions
MacMahon, M., Stankiewicz, B., and Kuipers, B. (2006) · 2006
Earlier work this paper cites.
Hierarchical task and motion planning in the now
Kaelbling, L. P. and Lozano-Pérez, T. (2011) · 2011
Earlier work this paper cites.
A reduction of imitation learning and structured prediction to no-regret online learning
Ross, S., Gordon, G., and Bagnell, D. (2011) · 2011
Earlier work this paper cites.
Mujoco: A physics engine for model-based control
Todorov, E., Erez, T., and Tassa, Y. (2012) · 2012
Earlier work this paper cites.
Learning phrase representations using RNN encoder–decoder for statistical machine translation
Cho, K., van Merriënboer, B., Gulcehre, C., Bahdanau, D., Bougares, F., Schwenk, H., and Bengio, Y. (2014) · 2014
Earlier work this paper cites.
Thinking in words: language as an embodied medium of thought
Dove, G. (2014) · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, D. P. and Ba, J. (2014) · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
Lin, T.-Y., Maire, M., Belongie, S., Hays, J., Perona, P., Ramanan, D., Dollár, P., and Zitnick, C. L. (2014) · 2014
Earlier work this paper cites.
VQA: Visual Question Answering
Antol, S., Agrawal, A., Lu, J., Mitchell, M., Batra, D., Zitnick, C. L., and Parikh, D. (2015) · 2015
Earlier work this paper cites.
Deep recurrent q-learning for partially observable mdps
Hausknecht, M. and Stone, P. (2015) · 2015
Earlier work this paper cites.
Deep convolutional inverse graphics network
Kulkarni, T. D., Whitney, W. F., Kohli, P., and Tenenbaum, J. (2015) · 2015
Earlier work this paper cites.
Ba, L. J., Kiros, J. R., and Hinton, G. E. (2016) · 2016
Earlier work this paper cites.
Pointing the unknown words
Gulcehre, C., Ahn, S., Nallapati, R., Zhou, B., and Bengio, Y. (2016) · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
He, K., Zhang, X., Ren, S., and Sun, J. (2016) · 2016
Cited alongside, same era.
Densecap: Fully convolutional localization networks for dense captioning
Johnson, J., Karpathy, A., and Fei-Fei, L. (2016) · 2016
Cited alongside, same era.
Using the output embedding to improve language models
Press, O. and Wolf, L. (2016) · 2016
Cited alongside, same era.
Classical planning in deep latent space: Bridging the subsymbolic-symbolic boundary
Asai, M. and Fukunaga, A. (2017) · 2017
Cited alongside, same era.
Mask r-cnn
He, K., Gkioxari, G., Dollár, P., and Girshick, R. (2017) · 2017
Cited alongside, same era.
Modeling relationships in referential expressions with compositional modular networks
Grounding language for transfer in deep reinforcement learning
Narasimhan, K., Barzilay, R., and Jaakkola, T. (2018) · 2018
Later among the works it cites.
Counting to explore and generalize in text-based games
Yuan, X., Côté, M.-A., Sordoni, A., Laroche, R., Combes, R. T. d., Hausknecht, M., and Trischler, A. (2018) · 2018
Later among the works it cites.
BabyAI: First steps towards grounded language learning with a human in the loop
Chevalier-Boisvert, M., Bahdanau, D., Lahlou, S., Willems, L., Saharia, C., Nguyen, T. H., and Bengio, Y. (2019) · 2019
Later among the works it cites.
Hierarchical decision making by generating and following natural language instructions
Hu, H., Yarats, D., Gong, Q., Tian, Y., and Lewis, M. (2019) · 2019
Later among the works it cites.
Language as an abstraction for hierarchical deep reinforcement learning
Jiang, Y., Gu, S. S., Murphy, K. P., and Finn, C. (2019) · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Hu, R., Rohrbach, M., Andreas, J., Darrell, T., and Saenko, K. (2017) · 2017
Cited alongside, same era.
Clevr: A diagnostic dataset for compositional language and elementary visual reasoning
Johnson, J., Hariharan, B., van der Maaten, L., Fei-Fei, L., Zitnick, C. L., and Girshick, R. (2017) · 2017
Cited alongside, same era.
Ai2-thor: An interactive 3d environment for visual ai
Kolve, E., Mottaghi, R., Han, W., VanderBilt, E., Weihs, L., Herrasti, A., Gordon, D., Zhu, Y., Gupta, A., and Farhadi, A. (2017) · 2017
Cited alongside, same era.
Sharma, S., Asri, L. E., Schulz, H., and Zumer, J. (2017) · 2017
Cited alongside, same era.
Attention is all you need
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, L. u., and Polosukhin, I. (2017) · 2017
Cited alongside, same era.
Learning to see physics via visual de-animation
Wu, J., Lu, E., Kohli, P., Freeman, B., and Tenenbaum, J. (2017) · 2017
Cited alongside, same era.
Visual semantic planning using deep successor representations
Zhu, Y., Gordon, D., Kolve, E., Fox, D., Fei-Fei, L., Gupta, A., Mottaghi, R., and Farhadi, A. (2017) · 2017
Cited alongside, same era.
Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks
Lu, J., Batra, D., Parikh, D., and Lee, S. (2019) · 2019
Later among the works it cites.
Language is power: Representing states using natural language in reinforcement learning
Schwartz, E., Tennenholtz, G., Tessler, C., and Mannor, S. (2019) · 2019
Later among the works it cites.
Videobert: A joint model for video and language representation learning
Sun, C., Myers, A., Vondrick, C., Murphy, K., and Schmid, C. (2019) · 2019
Later among the works it cites.
Reinforced cross-modal matching and self-supervised imitation learning for vision-language navigation
Wang, X., Huang, Q., Celikyilmaz, A., Gao, J., Shen, D., Wang, Y.-F., Wang, W. Y., and Zhang, L. (2019) · 2019
Later among the works it cites.
Learning dynamic belief graphs to generalize on text-based games
Adhikari, A., Yuan, X., Côté, M.-A., Zelinka, M., Rondeau, M.-A., Laroche, R., Poupart, P., Tang, J., Trischler, A., and Hamilton, W. L. (2020) · 2020
Closest in time.
Graph constrained reinforcement learning for natural language action spaces
Ammanabrolu, P. and Hausknecht, M. (2020) · 2020
Closest in time.
Linguistic distributional knowledge and sensorimotor grounding both contribute to semantic category production
Banks, B., Wingfield, C., and Connell, L. (2020) · 2020
Closest in time.
Experience Grounds Language
Bisk, Y., Holtzman, A., Thomason, J., Andreas, J., Bengio, Y., Chai, J., Lapata, M., Lazaridou, A., May, J., Nisnevich, A., Pinto, N., and Turian, J. (2020) · 2020
Closest in time.
Interactive fiction games: A colossal adventure
Hausknecht, M. J., Ammanabrolu, P., Côté, M.-A., and Yuan, X. (2020) · 2020
Closest in time.
The nethack learning environment
Küttler, H., Nardelli, N., Miller, A. H., Raileanu, R., Selvatici, M., Grefenstette, E., and Rocktäschel, T. (2020) · 2020
Closest in time.
Alfred: A benchmark for interpreting grounded instructions for everyday tasks
Shridhar, M., Thomason, J., Gordon, D., Bisk, Y., Han, W., Mottaghi, R., Zettlemoyer, L., and Fox, D. (2020) · 2020
Closest in time.
Unbiased scene graph generation from biased training
Tang, K., Niu, Y., Huang, J., Shi, J., and Zhang, H. (2020) · 2020
Closest in time.
RTFM: Generalising to novel environment dynamics via reading
Zhong, V., Rocktäschel, T., and Grefenstette, E. (2020) · 2020
Closest in time.