Fetching the paper…
Reading the bibliography…
Vision-and-Language Navigation (VLN) is a challenging task in the field of artificial intelligence.
Self-monitoring Navigation Agent via Auxiliary Progress Estimation
Ma, C.-Y.; Lu, J.; Wu, Z.; AlRegib, G.; Kira, Z.; Socher, R.; and Xiong, C. 2019a · 1901
Earlier work this paper cites.
Learning to Navigate Unseen Environments: Back Translation with Environmental Dropout
Tan, H.; Yu, L.; and Bansal, M. 2019 · 1904
Earlier work this paper cites.
Stay on the Path: Instruction Fidelity in Vision-and-language Navigation
Jain, V.; Magalhaes, G.; Ku, A.; Vaswani, A.; Ie, E.; and Baldridge, J. 2019 · 1905
Earlier work this paper cites.
Are You Looking? Grounding to Multiple Modalities in Vision-and-language Navigation
Hu, R.; Fried, D.; Rohrbach, A.; Klein, D.; Darrell, T.; and Saenko, K. 2019 · 1906
Earlier work this paper cites.
Vilbert: Pretraining Task-agnostic Visiolinguistic Representations for Vision-and-language Tasks
Lu, J.; Batra, D.; Parikh, D.; and Lee, S. 2019 · 1908
Earlier work this paper cites.
LXMERT: Learning Cross-modality Encoder Representations from Transformers
Tan, H.; and Bansal, M. 2019 · 1908
Earlier work this paper cites.
Robust Navigation with Language Pretraining and Stochastic Sampling
Li, X.; Li, C.; Xia, Q.; Bisk, Y.; Celikyilmaz, A.; Gao, J.; Smith, N.; and Choi, Y. 2019 · 1909
Earlier work this paper cites.
Asynchronous Methods for Deep Reinforcement Learning
Mnih, V.; Badia, A. P.; Mirza, M.; Graves, A.; Lillicrap, T.; Harley, T.; Silver, D.; and Kavukcuoglu, K. 2016 · 1937
Earlier work this paper cites.
Procedures as a Representation for Data in a Computer Program for Understanding Natural Language
Winograd, T. 1971 · 1971
Earlier work this paper cites.
Neural Network Ensembles
Hansen, L. K.; and Salamon, P. 1990 · 1990
Earlier work this paper cites.
Random Decision Forests
Ho, T. K. 1995 · 1995
Earlier work this paper cites.
Bagging Predictors
Breiman, L. 1996 · 1996
Earlier work this paper cites.
A Decision-theoretic Generalization of On-line Learning and an Application to Boosting
Freund, Y.; and Schapire, R. E. 1997 · 1997
Earlier work this paper cites.
Sub-instruction Aware Vision-and-language Navigation
Hong, Y.; Rodriguez-Opazo, C.; Wu, Q.; and Gould, S. 2020b · 2004
Earlier work this paper cites.
Language and Visual Entity Relationship Graph for Agent Navigation
Hong, Y.; Rodriguez-Opazo, C.; Qi, Y.; Wu, Q.; and Gould, S. 2020a · 2010
Cited alongside, same era.
Room-across-room: Multilingual Vision-and-language Navigation with Dense Spatiotemporal Grounding
Ku, A.; Anderson, P.; Patel, R.; Ie, E.; and Baldridge, J. 2020 · 2010
Cited alongside, same era.
LSTM Neural Networks for Language Modeling
Sundermeyer, M.; Schlüter, R.; and Ney, H. 2012 · 2012
Cited alongside, same era.
Horizontal and Vertical Ensemble with Deep Representation for Classification
Xie, J.; Xu, B.; and Chuang, Z. 2013 · 2013
Cited alongside, same era.
Effective Approaches to Attention-based Neural Machine Translation
Attention is All You Need
Vaswani, A.; Shazeer, N.; Parmar, N.; Uszkoreit, J.; Jones, L.; Gomez, A. N.; Kaiser, Ł.; and Polosukhin, I. 2017 · 2017
Later among the works it cites.
Vision-and-language Navigation: Interpreting Visually-grounded Navigation Instructions in Real Environments
Anderson, P.; Wu, Q.; Teney, D.; Bruce, J.; Johnson, M.; Sünderhauf, N.; Reid, I.; Gould, S.; and Van Den Hengel, A. 2018 · 2018
Later among the works it cites.
Speaker-follower Models for Vision-and-language Navigation
Fried, D.; Hu, R.; Cirik, V.; Rohrbach, A.; Andreas, J.; Morency, L.-P.; Berg-Kirkpatrick, T.; Saenko, K.; Klein, D.; and Darrell, T. 2018 · 2018
Later among the works it cites.
Building Generalizable Agents with a Realistic and Rich 3D Environment
Wu, Y.; Wu, Y.; Gkioxari, G.; and Tian, Y. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Luong, M.-T.; Pham, H.; and Manning, C. D. 2015 · 2015
Cited alongside, same era.
Temporal Ensembling for Semi-supervised Learning
Laine, S.; and Aila, T. 2016 · 2016
Cited alongside, same era.
Sgdr: Stochastic Gradient Descent with Warm Restarts
Loshchilov, I.; and Hutter, F. 2016 · 2016
Cited alongside, same era.
Boosted Convolutional Neural Networks
Moghimi, M.; Belongie, S. J.; Saberian, M. J.; Yang, J.; Vasconcelos, N.; and Li, L.-J. 2016 · 2016
Cited alongside, same era.
Home: A Household Multimodal Environment
Brodeur, S.; Perez, E.; Anand, A.; Golemo, F.; Celotti, L.; Strub, F.; Rouat, J.; Larochelle, H.; and Courville, A. 2017 · 2017
Cited alongside, same era.
Matterport3D: Learning from RGB-D Data in Indoor Environments
Chang, A.; Dai, A.; Funkhouser, T.; Halber, M.; Niessner, M.; Savva, M.; Song, S.; Zeng, A.; and Zhang, Y. 2017 · 2017
Cited alongside, same era.
Self-ensembling for Visual Domain Adaptation
French, G.; Mackiewicz, M.; and Fisher, M. 2017 · 2017
Cited alongside, same era.
Snapshot Ensembles: Train 1, Get M for Free
Huang, G.; Li, Y.; Pleiss, G.; Liu, Z.; Hopcroft, J. E.; and Weinberger, K. Q. 2017 · 2017
Cited alongside, same era.
Yan, C.; Misra, D.; Bennnett, A.; Walsman, A.; Bisk, Y.; and Artzi, Y. 2018 · 2018
Later among the works it cites.
Towards Learning a Generic Agent for Vision-and-language Navigation via Pre-training
Hao, W.; Li, C.; Li, X.; Carin, L.; and Gao, J. 2020 · 2020
Later among the works it cites.
Beyond the Nav-graph: Vision-and-language Navigation in Continuous Environments
Krantz, J.; Wijmans, E.; Majumdar, A.; Batra, D.; and Lee, S. 2020 · 2020
Later among the works it cites.
Oscar: Object-semantics aligned pre-training for vision-language tasks
Li, X.; Yin, X.; Li, C.; Zhang, P.; Hu, X.; Zhang, L.; Wang, L.; Hu, H.; Dong, L.; Wei, F.; et al. 2020 · 2020
Later among the works it cites.
Improving Vision-and-language Navigation with Image-text Pairs from the Web
Majumdar, A.; Shrivastava, A.; Lee, S.; Anderson, P.; Parikh, D.; and Batra, D. 2020 · 2020
Later among the works it cites.
Vision-Language Navigation With Self-Supervised Auxiliary Reasoning Tasks
Zhu, F.; Zhu, Y.; Chang, X.; and Liang, X. 2020 · 2020
Later among the works it cites.
Topological Planning with Transformers for Vision-and-Language Navigation
Chen, K.; Chen, J. K.; Chuang, J.; Vázquez, M.; and Savarese, S. 2021 · 2021
Closest in time.
VLN BERT: A Recurrent Vision-and-Language BERT for Navigation
Hong, Y.; Wu, Q.; Qi, Y.; Rodriguez-Opazo, C.; and Gould, S. 2021 · 2021
Closest in time.
Vision-Language Navigation with Random Environmental Mixup
Liu, C.; Zhu, F.; Chang, X.; Liang, X.; and Shen, Y.-D. 2021 · 2021
Closest in time.
Structured Scene Memory for Vision-Language Navigation
Wang, H.; Wang, W.; Liang, W.; Xiong, C.; and Shen, J. 2021 · 2021
Closest in time.