Fetching the paper…
Reading the bibliography…
The Touchdown dataset (Chen et al., 2019) provides instructions by human annotators for navigation through New York City streets and for resolving spatial descriptions at a given location.
The streetlearn environment and dataset
Mirowski, P., Banki-Horvath, A., Anderson, K., Teplyashin, D., Hermann, K. M., Malinowski, M., Grimes, M. K., Simonyan, K., Kavukcuoglu, K., Zisserman, A., and Hadsell, R. (2019) · 1903
Earlier work this paper cites.
The HCRC Map Task corpus: Natural dialogue for speech recognition
Thompson, H. S., Anderson, A., Bard, E. G., Doherty-Sneddon, G., Newlands, A., and Sotillo, C. (1993) · 1993
Earlier work this paper cites.
Walk the talk: Connecting language, knowledge, action in route instructions
Macmahon, M., Stankiewicz, B., and Kuipers, B. (2006) · 2006
Earlier work this paper cites.
Hogwild: A lock-free approach to parallelizing stochastic gradient descent
Recht, B., Ré, C., Wright, S. J., and Niu, F. (2011) · 2011
Earlier work this paper cites.
Matterport3d: Learning from RGB-D data in indoor environments
Chang, A., Dai, A., Funkhouser, T., Halber, M., Niessner, M., Savva, M., Song, S., Zeng, A., and Zhang, Y. (2017) · 2017
Earlier work this paper cites.
Grounded language learning in a simulated 3d world
Hermann, K. M., Hill, F., Green, S., Wang, F., Faulkner, R., Soyer, H., Szepesvari, D., Czarnecki, W. M., Jaderberg, M., Teplyashin, D., Wainwright, M., Apps, C., Hassabis, D., and Blunsom, P. (2017) · 2017
Earlier work this paper cites.
Understanding grounded language learning agents
Hill, F., Hermann, K. M., Blunsom, P., and Clark, S. (2017) · 2017
Earlier work this paper cites.
Mapping instructions and visual observations to actions with reinforcement learning
Misra, D., Langford, J., and Artzi, Y. (2017) · 2017
Cited alongside, same era.
Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments
Anderson, P., Wu, Q., Teney, D., Bruce, J., Johnson, M., Sünderhauf, N., Reid, I., Gould, S., and van den Hengel, A. (2018) · 2018
Cited alongside, same era.
Following formulaic map instructions in a street simulation environment
Cirik, V., Zhang, Y., and Baldridge, J. (2018) · 2018
Cited alongside, same era.
Talk the walk: Navigating new york city through grounded dialogue
de Vries, H., Shuster, K., Batra, D., Parikh, D., Weston, J., and Kiela, D. (2018) · 2018
Cited alongside, same era.
Learning to navigate in cities without a map
Mirowski, P., Grimes, M. K., Malinowski, M., Hermann, K. M., Anderson, K., Teplyashin, D., Simonyan, K., Kavukcuoglu, K., Zisserman, A., and Hadsell, R. (2018) · 2018
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Devlin, J., Chang, M.-W., Lee, K., and Toutanova, K. (2019) · 2019
Later among the works it cites.
Effective and general evaluation for instruction conditioned navigation using dynamic time warping
Ilharco, G., Jain, V., Ku, A., Ie, E., and Baldridge, J. (2019) · 2019
Later among the works it cites.
Valan: Vision and language agent navigation
Lansing, L., Jain, V., Mehta, H., Huang, H., and Ie, E. (2019) · 2019
Later among the works it cites.
Reinforced cross-modal matching and self-supervised imitation learning for vision-language navigation
Wang, X., Huang, Q., Çelikyilmaz, A., Gao, J., Shen, D., Wang, Y., Wang, W. Y., and Zhang, L. (2019) · 2019
Later among the works it cites.
Navigation agents for the visually impaired: A sidewalk simulator and experiments
Weiss, M., Chamorro, S., Girgis, R., Luck, M., Kahou, S. E., Cohen, J. P., Nowrouzezahrai, D., Precup, D., Golemo, F., and Pal, C. (2019) · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Mapping instructions to actions in 3D environments with visual goal prediction
Misra, D., Bennett, A., Blukis, V., Niklasson, E., Shatkhin, M., and Artzi, Y. (2018) · 2018
Cited alongside, same era.
Touchdown: Natural language navigation and spatial reasoning in visual street environments
Chen, H., Suhr, A., Misra, D., and Artzi, Y. (2019) · 2019
Cited alongside, same era.
Vision-language navigation with self-supervised auxiliary reasoning tasks
Zhu, F., Zhu, Y., Chang, X., and Liang, X. (2019) · 2019
Later among the works it cites.
Learning to follow directions in street view
Hermann, K. M., Malinowski, M., Mirowski, P., Banki-Horvath, A., Anderson, K., and Hadsell, R. (2020) · 2020
Closest in time.