Fetching the paper…
Reading the bibliography…
The role of robots in society keeps expanding, bringing with it the necessity of interacting and communicating with humans.
Neural computation 9
Hochreiter, S., Schmidhuber, J.: Long short-term memory · 1997
Earlier work this paper cites.
Proceedings of the IEEE 86
LeCun, Y., Bottou, L., Bengio, Y., Haffner, P., et al.: Gradient-based learning applied to document recognition · 1998
Earlier work this paper cites.
In: C. Freksa, D.M. Mark (eds.) Spatial Information Theory. Cognitive and Computational Foundations of Geographic Information Science, pp. 51–64 (1999)
Tversky, B., Lee, P.U.: Pictorial and verbal tools for conveying routes · 1999
Earlier work this paper cites.
In: Spatial Information Theory (2001)
Michon, P.E., Denis, M.: When and why are visual landmarks used in giving directions? · 2001
Earlier work this paper cites.
In: International Conference on Spatial Information Theory, pp. 362–374. Springer (2003)
Tom, A., Denis, M.: Referring to landmark or street information in route directions: What difference does it make? · 2003
Earlier work this paper cites.
Applied Cognitive Psychology: The Official Journal of the Society for Applied Research in Memory and Cognition 18
Tom, A., Denis, M.: Language and spatial cognition: Comparing the roles of landmarks and street names in route instructions · 2004
Earlier work this paper cites.
In: IEEE Conference on Compter Vision and Pattern Recognition (2005)
Chopra, S.: Learning a similarity metric discriminatively, with application to face verification · 2005
Earlier work this paper cites.
Journal of Visual Languages and Computing 16
Klippel, A., Tappe, H., Kulik, L., Lee, P.U.: Wayfinding choremes—a language for modeling conceptual route knowledge · 2005
Earlier work this paper cites.
In: Spatial Information Theory (2005)
Klippel, A., Winter, S.: Structural salience of landmarks for route directions · 2005
Earlier work this paper cites.
IEEE Transactions on Intelligent Transportation Systems 8
Millonig, A., Schechtner, K.: Developing landmark-based pedestrian-navigation systems · 2007
Earlier work this paper cites.
In: ACM SIGGRAPH (2008)
Grabler, F., Agrawala, M., Sumner, R.W., Pauly, M.: Automatic generation of tourist maps · 2008
Earlier work this paper cites.
In: IEEE conference on computer vision and pattern recognition (2009)
Deng, J., Dong, W., Socher, R., Li, L.J., Li, K., Fei-Fei, L.: Imagenet: A large-scale hierarchical image database · 2009
Earlier work this paper cites.
In: Proceedings of the 48th Annual Meeting of the Association for Computational Linguistics, pp. 806–814. Association for Computational Linguistics (2010)
Vogel, A., Jurafsky, D.: Learning to follow navigational directions · 2010
Earlier work this paper cites.
In: Twenty-Fifth AAAI Conference on Artificial Intelligence (2011)
Chen, D.L., Mooney, R.J.: Learning to interpret natural language navigation instructions from observations · 2011
Earlier work this paper cites.
Cognition 121
Hölscher, C., Tenbrink, T., Wiener, J.M.: Would you follow your own route description? cognitive strategies in urban route planning · 2011
Earlier work this paper cites.
Spatial Cognition & Computation 12
Ishikawa, T., Nakamura, U.: Landmark selection in the environment: Relationships with object characteristics and sense of direction · 2012
Earlier work this paper cites.
In: Advances in neural information processing systems, pp. 3111–3119 (2013)
Mikolov, T., Sutskever, I., Chen, K., Corrado, G.S., Dean, J.: Distributed representations of words and phrases and their compositionality · 2013
Earlier work this paper cites.
arXiv preprint arXiv:1409.0473 (2014)
Bahdanau, D., Cho, K., Bengio, Y.: Neural machine translation by jointly learning to align and translate · 2014
Earlier work this paper cites.
Graves, A., Wayne, G., Danihelka, I.: Neural turing machines · 2014
Earlier work this paper cites.
In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp. 3710–3717 (2014)
Khosla, A., An An, B., Lim, J.J., Torralba, A.: Looking beyond the visible scene · 2014
Earlier work this paper cites.
In: Proceedings of the 2014 conference on empirical methods in natural language processing (EMNLP), pp. 1532–1543 (2014)
Pennington, J., Socher, R., Manning, C.: Glove: Global vectors for word representation · 2014
Earlier work this paper cites.
In: Advances in neural information processing systems, pp. 3104–3112 (2014)
Sutskever, I., Vinyals, O., Le, Q.V.: Sequence to sequence learning with neural networks · 2014
Earlier work this paper cites.
In: Proceedings of the 2nd ACM SIGSPATIAL International Workshop on Interacting with Maps, pp. 8–14. ACM (2014)
Weissenberg, J., Gygli, M., Riemenschneider, H., Van Gool, L.: Navigation using special buildings as signposts · 2014
Earlier work this paper cites.
In: IEEE International Conference on Robotics and Automation (ICRA) (2015)
Boularias, A., Duvallet, F., Oh, J., Stentz, A.: Grounding spatial relations for outdoor robot navigation · 2015
Earlier work this paper cites.
In: Proceedings of the IEEE conference on computer vision and pattern recognition, pp. 2625–2634 (2015)
Donahue, J., Anne Hendricks, L., Guadarrama, S., Rohrbach, M., Venugopalan, S., Saenko, K., Darrell, T.: Long-term recurrent convolutional networks for visual recognition and description · 2015
Earlier work this paper cites.
In: Proceedings of the IEEE conference on computer vision and pattern recognition (2015)
Karpathy, A., Fei-Fei, L.: Deep visual-semantic alignments for generating image descriptions · 2015
Earlier work this paper cites.
arXiv preprint arXiv:1511.02793 (2015)
Mansimov, E., Parisotto, E., Ba, J.L., Salakhutdinov, R.: Generating images from captions with attention · 2015
Earlier work this paper cites.
In: International conference on machine learning, pp. 2048–2057 (2015)
Xu, K., Ba, J., Kiros, R., Cho, K., Courville, A., Salakhudinov, R., Zemel, R., Bengio, Y.: Show, attend and tell: Neural image caption generation with visual attention · 2015
Earlier work this paper cites.
In: Proceedings of the IEEE international conference on computer vision, pp. 19–27 (2015)
Zhu, Y., Kiros, R., Zemel, R., Salakhutdinov, R., Urtasun, R., Torralba, A., Fidler, S.: Aligning books and movies: Towards story-like visual explanations by watching movies and reading books · 2015
Cited alongside, same era.
arXiv preprint arXiv:1601.01705 (2016)
Andreas, J., Rohrbach, M., Darrell, T., Klein, D.: Learning to compose neural networks for question answering · 2016
Cited alongside, same era.
MIT Press (2016)
Goodfellow, I., Bengio, Y., Courville, A.: Deep Learning · 2016
Cited alongside, same era.
arXiv preprint arXiv:1603.08983 (2016)
Graves, A.: Adaptive computation time for recurrent neural networks · 2016
Cited alongside, same era.
Nature 538
Graves, A., Wayne, G., Reynolds, M., Harley, T., Danihelka, I., Grabska-Barwińska, A., Colmenarejo, S.G., Grefenstette, E., Ramalho, T., Agapiou, J., et al.: Hybrid computing using a neural network with dynamic external memory · 2016
In: NIPS (2018)
Fried, D., Hu, R., Cirik, V., Rohrbach, A., Andreas, J., Morency, L.P., Berg-Kirkpatrick, T., Saenko, K., Klein, D., Darrell, T.: Speaker-follower models for vision-and-language navigation · 2018
Later among the works it cites.
In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (2018)
Gordon, D., Kembhavi, A., Rastegari, M., Redmon, J., Fox, D., Farhadi, A.: Iqa: Visual question answering in interactive environments · 2018
Later among the works it cites.
In: European Conference on Computer Vision (ECCV) (2018)
Hecker, S., Dai, D., Van Gool, L.: End-to-end learning of driving models with surround-view cameras and route planners · 2018
Later among the works it cites.
In: Proceedings of the European conference on computer vision (ECCV) (2018)
Hu, R., Andreas, J., Darrell, T., Saenko, K.: Explainable neural computation via stack neural module networks · 2018
Later among the works it cites.
arXiv preprint arXiv:1803.03067 (2018)
Hudson, D.A., Manning, C.D.: Compositional attention networks for machine reasoning · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (2016)
Gygli, M., Song, Y., Cao, L.: Video2gif: Automatic generation of animated gifs from video · 2016
Cited alongside, same era.
In: Proceedings of the IEEE conference on computer vision and pattern recognition (2016)
He, K., Zhang, X., Ren, S., Sun, J.: Deep residual learning for image recognition · 2016
Cited alongside, same era.
In: European Conference on Computer Vision (ECCV) (2016)
Weyand, T., Kostrikov, I., Philbin, J.: Planet - photo geolocation with convolutional neural networks · 2016
Cited alongside, same era.
arXiv preprint arXiv:1609.08144 (2016)
Wu, Y., Schuster, M., Chen, Z., Le, Q.V., Norouzi, M., Macherey, W., Krikun, M., Cao, Y., Gao, Q., Macherey, K., et al.: Google’s neural machine translation system: Bridging the gap between human and machine translation · 2016
Cited alongside, same era.
In: Proceedings of the IEEE conference on computer vision and pattern recognition, pp. 21–29 (2016)
Yang, Z., He, X., Gao, J., Deng, L., Smola, A.: Stacked attention networks for image question answering · 2016
Cited alongside, same era.
International Journal of Computer Vision 123
Agrawal, A., Lu, J., Antol, S., Mitchell, M., Zitnick, C.L., Parikh, D., Batra, D.: Vqa: Visual question answering · 2017
Cited alongside, same era.
In: Proceedings of the IEEE International Conference on Computer Vision (2017)
Anne Hendricks, L., Wang, O., Shechtman, E., Sivic, J., Darrell, T., Russell, B.: Localizing moments in video with natural language · 2017
Cited alongside, same era.
In: Advances in Neural Information Processing Systems, pp. 773–782 (2018)
Kumar, A., Gupta, S., Fouhey, D., Levine, S., Malik, J.: Visual memory for robust path following · 2018
Later among the works it cites.
arXiv preprint arXiv:1803.04376 (2018)
Luo, R., Price, B., Cohen, S., Shakhnarovich, G.: Discriminability objective for training descriptive captions · 2018
Later among the works it cites.
In: NIPS (2018)
Mirowski, P., Grimes, M., Malinowski, M., Hermann, K.M., Anderson, K., Teplyashin, D., Simonyan, K., kavukcuoglu, k., Zisserman, A., Hadsell, R.: Learning to navigate in cities without a map · 2018
Later among the works it cites.
In: 2018 IEEE Winter Conference on Applications of Computer Vision (WACV), pp. 1861–1870. IEEE (2018)
Vasudevan, A.B., Dai, D., Van Gool, L.: Object referring in visual scene with spoken language · 2018
Later among the works it cites.
arXiv preprint arXiv:1807.03367 (2018)
de Vries, H., Shuster, K., Batra, D., Parikh, D., Weston, J., Kiela, D.: Talk the walk: Navigating new york city through grounded dialogue · 2018
Later among the works it cites.
In: ECCV (2018)
Wang, X., Xiong, W., Wang, H., Yang Wang, W.: Look before you leap: Bridging model-free and model-based reinforcement learning for planned-ahead vision-and-language navigation · 2018
Later among the works it cites.
CoRR (2018)
Zang, X., Pokle, A., Vázquez, M., Chen, K., Niebles, J.C., Soto, A., Savarese, S.: Translating navigation instructions in natural language to a high-level plan for behavioral robot navigation · 2018
Later among the works it cites.
Applied Sciences 8
Zhu, X., Li, L., Liu, J., Peng, H., Niu, X.: Captioning transformer with stacked attention modules · 2018
Later among the works it cites.
In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (2019)
Chen, H., Shur, A., Misra, D., Snavely, N., Artzi, Y.: Touchdown: Natural language navigation and spatial reasoning in visual street environments · 2019
Closest in time.
In: Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP), pp. 2088–2098 (2019)
Deruyttere, T., Vandenhende, S., Grujicic, D., Van Gool, L., Moens, M.F.: Talk2car: Taking control of your self-driving car · 2019
Closest in time.
International Journal of Computer Vision (2019)
Gupta, S., Tolani, V., Davidson, J., Levine, S., Sukthankar, R., Malik, J.: Cognitive mapping and planning for visual navigation · 2019
Closest in time.
In: arXiv-1903.10995 (2019)
Hecker, S., Dai, D., Van Gool, L.: Learning accurate, comfortable and human-like driving · 2019
Closest in time.
arXiv e-prints (2019)
Hermann, K.M., Malinowski, M., Mirowski, P., Banki-Horvath, A., Anderson, K., Hadsell, R.: Learning To Follow Directions in Street View · 2019
Closest in time.
arXiv preprint arXiv:1905.04405 (2019)
Hu, R., Rohrbach, A., Darrell, T., Saenko, K.: Language-conditioned graph networks for relational reasoning · 2019
Closest in time.
In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp. 6741–6749 (2019)
Ke, L., Li, X., Bisk, Y., Holtzman, A., Gan, Z., Liu, J., Gao, J., Choi, Y., Srinivasa, S.: Tactical rewind: Self-correction via backtracking in vision-and-language navigation · 2019
Closest in time.
In: The IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2019)
Kim, J., Misu, T., Chen, Y.T., Tawari, A., Canny, J.: Grounding human-to-vehicle advice for self-driving vehicles · 2019
Closest in time.
arXiv preprint arXiv:1901.03035 (2019)
Ma, C.Y., Lu, J., Wu, Z., AlRegib, G., Kira, Z., Socher, R., Xiong, C.: Self-monitoring navigation agent via auxiliary progress estimation · 2019
Closest in time.
In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp. 6732–6740 (2019)
Ma, C.Y., Wu, Z., AlRegib, G., Xiong, C., Kira, Z.: The regretful agent: Heuristic-aided navigation through progress estimation · 2019
Closest in time.
In: The IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2019)
Nguyen, K., Dey, D., Brockett, C., Dolan, B.: Vision-based navigation with language-based assistance via imitation learning with indirect intervention · 2019
Closest in time.
In: The IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2019)
Thoma, J., Paudel, D.P., Chhatkuli, A., Probst, T., Gool, L.V.: Mapping, localization and path planning for image-based navigation using visual features and map · 2019
Closest in time.
In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp. 6629–6638 (2019)
Wang, X., Huang, Q., Celikyilmaz, A., Gao, J., Shen, D., Wang, Y.F., Yang Wang, W., Zhang, L.: Reinforced cross-modal matching and self-supervised imitation learning for vision-language navigation · 2019
Closest in time.
In: The IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (2019)
Wortsman, M., Ehsani, K., Rastegari, M., Farhadi, A., Mottaghi, R.: Learning to learn how to learn: Self-adaptive visual navigation using meta-learning · 2019
Closest in time.