Fetching the paper…
Reading the bibliography…
Vision-and-Language Navigation (VLN) requires an agent to follow natural-language instructions, explore the given environments, and reach the desired target locations.
Bleu: a method for automatic evaluation of machine translation
K. Papineni, S. Roukos, et al · 2002
Earlier work this paper cites.
Rouge: A package for automatic evaluation of summaries
C. Lin · 2004
Earlier work this paper cites.
Walk the talk: connecting language, knowledge, and action in route instructions
M. MacMahon, B. Stankiewicz, and B. Kuipers · 2006
Earlier work this paper cites.
Generalizing from several related classification tasks to a new unlabeled sample
G. Blanchard, G. Lee, and C. Scott · 2011
Earlier work this paper cites.
Domain generalization via invariant feature representation
K. Muandet, D. Balduzzi, and B. Schölkopf · 2013
Earlier work this paper cites.
Generative adversarial nets
I. Goodfellow, J. Pouget-Abadie, et al · 2014
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
D. Bahdanau, K. Cho, and Y. Bengio · 2015
Earlier work this paper cites.
Faster r-cnn: Towards real-time object detection with region proposal networks
S. Ren, K. He, et al · 2015
Earlier work this paper cites.
Imagenet large scale visual recognition challenge
O. Russakovsky, J. Deng, et al · 2015
Earlier work this paper cites.
Deep residual learning for image recognition
K. He, X. Zhang, et al · 2016
Earlier work this paper cites.
Matterport3d: Learning from rgb-d data in indoor environments
A. Chang, A. Dai, et al · 2017
Earlier work this paper cites.
Visual genome: Connecting language and vision using crowdsourced dense image annotations
R. Krishna, Y. Zhu, et al · 2017
Earlier work this paper cites.
Bottom-up and top-down attention for image captioning and visual question answering
P. Anderson, X. He, et al · 2018
Cited alongside, same era.
Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments
P. Anderson, Q. Wu, et al · 2018
Cited alongside, same era.
Mapping navigation instructions to continuous control actions with position-visitation prediction
V. Blukis, D. Misra, et al · 2018
Cited alongside, same era.
Embodied question answering
A. Das, S. Datta, et al · 2018
Cited alongside, same era.
Speaker-follower models for vision-and-language navigation
D. Fried, R. Hu, et al · 2018
Cited alongside, same era.
Conditional adversarial domain adaptation
M. Long, Z. Cao, et al · 2018
Cited alongside, same era.
Touchdown: Natural language navigation and spatial reasoning in visual street environments
H. Chen, A. Suhr, et al · 2019
Later among the works it cites.
Are you looking? grounding to multiple modalities in vision-and-language navigation
R. Hu, D. Fried, et al · 2019
Later among the works it cites.
Transferable representation learning in vision-and-language navigation
H. Huang, V. Jain, et al · 2019
Later among the works it cites.
Stay on the path: Instruction fidelity in vision-and-language navigation
V. Jain, G. Magalhaes, et al · 2019
Later among the works it cites.
Tactical rewind: Self-correction via backtracking in vision-and-language navigation
L. Ke, X. Li, et al · 2019
Later among the works it cites.
Self-monitoring navigation agent via auxiliary progress estimation
C. Ma, J. Lu, et al · 2019
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Learning to navigate in cities without a map
P. Mirowski, M. Grimes, et al · 2018
Cited alongside, same era.
Residual parameter transfer for deep domain adaptation
A. Rozantsev, M. Salzmann, and P. Fua · 2018
Cited alongside, same era.
Look before you leap: Bridging model-free and model-based reinforcement learning for planned-ahead vision-and-language navigation
X. Wang, W. Xiong, et al · 2018
Cited alongside, same era.
Chasing ghosts: Instruction following as bayesian state tracking
P. Anderson, A. Shrivastava, et al · 2019
Cited alongside, same era.
Domain generalization by solving jigsaw puzzles
F. M. Carlucci, A. D’Innocente, et al · 2019
Cited alongside, same era.
Joint domain alignment and discriminative feature learning for unsupervised deep domain adaptation
C. Chen, Z. Chen, et al · 2019
Cited alongside, same era.
Later among the works it cites.
The regretful agent: Heuristic-aided navigation through progress estimation
C. Ma, Z. Wu, et al · 2019
Later among the works it cites.
Learning to navigate unseen environments: Back translation with environmental dropout
H. Tan, L. Yu, and M. Bansal · 2019
Later among the works it cites.
Shifting the baseline: Single modality performance on visual navigation & qa
J. Thomason, D. Gordan, and Y. Bisk · 2019
Later among the works it cites.
Vision-and-dialog navigation
J. Thomason, M. Murray, et al · 2019
Later among the works it cites.
Reinforced cross-modal matching and self-supervised imitation learning for vision-language navigation
X. Wang, Q. Huang, et al · 2019
Later among the works it cites.
Multi-target embodied question answering
L. Yu, X. Chen, et al · 2019
Later among the works it cites.