Fetching the paper…
Reading the bibliography…
Vision Language Navigation in Continuous Environments (VLN-CE) represents a frontier in embodied AI, demanding agents to navigate freely in unbounded 3D spaces solely guided by natural language instructions.
E. C. Tolman, “Cognitive maps in rats and men.” Psychological review , vol. 55, no. 4, p. 189, 1948
1948
Earlier work this paper cites.
R. W. Komorowski, C. G. Garcia, A. Wilson, S. Hattori, M. W. Howard, and H. Eichenbaum, “Ventral hippocampal neurons are shaped by experience to represent behaviorally relevant contexts,” Journal of Neuroscience , vol. 33, no. 18, pp. 8079–8087, 2013
2013
Earlier work this paper cites.
A. R. Preston and H. Eichenbaum, “Interplay of hippocampus and prefrontal cortex in memory,” Current biology , vol. 23, no. 17, pp. R764–R773, 2013
2013
Earlier work this paper cites.
2017
Earlier work this paper cites.
2018
Earlier work this paper cites.
P. Anderson, Q. Wu, D. Teney, J. Bruce, M. Johnson, N. Sünderhauf, I. Reid, S. Gould, and A. Van Den Hengel, “Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments,” in Proceedings of the IEEE conference on computer vision and pattern recognition , 2018, pp. 3674–3683
2018
Earlier work this paper cites.
H. Chen, A. Suhr, D. Misra, N. Snavely, and Y. Artzi, “Touchdown: Natural language navigation and spatial reasoning in visual street environments,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2019, pp. 12 538–12 547
2019
Earlier work this paper cites.
M. Savva, A. Kadian, O. Maksymets, Y. Zhao, E. Wijmans, B. Jain, J. Straub, J. Liu, V. Koltun, J. Malik, et al. , “Habitat: A platform for embodied ai research,” in Proceedings of the IEEE/CVF international conference on computer vision , 2019, pp. 9339–9347
2019
Earlier work this paper cites.
X. Wang, Q. Huang, A. Celikyilmaz, J. Gao, D. Shen, Y.-F. Wang, W. Y. Wang, and L. Zhang, “Reinforced cross-modal matching and self-supervised imitation learning for vision-language navigation,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2019, pp. 6629–6638
2019
Earlier work this paper cites.
J. Krantz, E. Wijmans, A. Majumdar, D. Batra, and S. Lee, “Beyond the nav-graph: Vision-and-language navigation in continuous environments,” in Computer Vision–ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part XXVIII 16 . Springer, 2020, pp. 104–120
2020
Earlier work this paper cites.
Y. Qi, Q. Wu, P. Anderson, X. Wang, W. Y. Wang, C. Shen, and A. v. d. Hengel, “Reverie: Remote embodied visual referring expression in real indoor environments,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 9982–9991
2020
Earlier work this paper cites.
X. Wang, Q. Huang, A. Celikyilmaz, J. Gao, D. Shen, and L. Zhang, “Vision-language navigation policy learning and adaptation,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 43, pp. 4205–4216, 2020
2020
Cited alongside, same era.
2020
Cited alongside, same era.
S. Chen, P.-L. Guhur, C. Schmid, and I. Laptev, “History aware multimodal transformer for vision-and-language navigation,” Advances in neural information processing systems , vol. 34, pp. 5834–5847, 2021
2021
Cited alongside, same era.
C. Gao, J. Chen, S. Liu, L. Wang, Q. Zhang, and Q. Wu, “Room-and-object aware knowledge reasoning for remote embodied referring expression,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2021, pp. 3064–3073
2021
Cited alongside, same era.
J. Krantz and S. Lee, “Sim-2-sim transfer for vision-and-language navigation in continuous environments,” in European Conference on Computer Vision . Springer, 2022, pp. 588–603
2022
Later among the works it cites.
2022
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
P.-L. Guhur, M. Tapaswi, S. Chen, I. Laptev, and C. Schmid, “Airbert: In-domain pretraining for vision-and-language navigation,” in 2021 IEEE/CVF International Conference on Computer Vision (ICCV) , 2021, pp. 1614–1623
2021
Cited alongside, same era.
M. Z. Irshad, C.-Y. Ma, and Z. Kira, “Hierarchical cross-modal agent for robotics vision-and-language navigation,” in Proceedings of the IEEE International Conference on Robotics and Automation (ICRA) , 2021
2021
Cited alongside, same era.
J. Krantz, A. Gokaslan, D. Batra, S. Lee, and O. Maksymets, “Waypoint models for instruction-guided navigation in continuous environments,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2021, pp. 15 162–15 171
2021
Cited alongside, same era.
2021
Cited alongside, same era.
J. Chen, J. Luo, Y. Pan, Y. Li, T. Yao, H. Chao, and T. Mei, “Boosting vision-and-language navigation with direction guiding and backtracing,” ACM Transactions on Multimedia Computing, Communications and Applications , vol. 19, pp. 1 – 16, 2022
2022
Cited alongside, same era.
Y. Hong, Z. Wang, Q. Wu, and S. Gould, “Bridging the gap between learning in discrete and continuous environments for vision-and-language navigation,” in 2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2022, pp. 15 418–15 428
2022
Cited alongside, same era.
M. Z. Irshad, N. Chowdhury Mithun, Z. Seymour, H.-P. Chiu, S. Samarasekera, and R. Kumar, “Semantically-aware spatio-temporal reasoning agent for vision-and-language navigation in continuous environments,” in 2022 26th International Conference on Pattern Recognition (ICPR) , 2022, pp. 4065–4071
2022
Cited alongside, same era.
2023
Later among the works it cites.
D. Shah, M. R. Equi, B. Osiński, F. Xia, B. Ichter, and S. Levine, “Navigation with large language models: Semantic guesswork as a heuristic for planning,” in Conference on Robot Learning . PMLR, 2023, pp. 2683–2699
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.