Fetching the paper…
Reading the bibliography…
Navigation presents a significant challenge for persons with visual impairments (PVI).
D. L. Rudman and M. Durdle, “Living with fear: The lived experience of community mobility among older adults with low vision,” Journal of aging and physical activity , vol. 17, no. 1, pp. 106–122, 2008
2008
Earlier work this paper cites.
N. A. Giudice and G. E. Legge, “Blind navigation and the role of technology,” The engineering handbook of smart technology for aging, disability, and independence , pp. 479–500, 2008
2008
Earlier work this paper cites.
M. Swobodzinski and M. Raubal, “An indoor routing algorithm for the blind: development and comparison to a routing algorithm for the sighted,” International Journal of Geographical Information Science , vol. 23, no. 10, pp. 1315–1343, 2009
2009
Earlier work this paper cites.
M. Y. Wang, J. Rousseau, H. Boisjoly, H. Schmaltz, M.-J. Kergoat, S. Moghadaszadeh, F. Djafari, and E. E. Freeman, “Activity limitation due to a fear of falling in older adults with eye disease,” Investigative ophthalmology & visual science , vol. 53, no. 13, pp. 7967–7972, 2012
2012
Earlier work this paper cites.
R. V. Jawale, M. V. Kadam, R. S. Gaikawad, and L. S. Kondaka, “Ultrasonic navigation based blind aid for the visually impaired,” in 2017 IEEE International Conference on Power, Control, Signals and Instrumentation Engineering (ICPCSI) . IEEE, 2017, pp. 923–928
2017
Earlier work this paper cites.
D. Fried, R. Hu, V. Cirik, A. Rohrbach, J. Andreas, L.-P. Morency, T. Berg-Kirkpatrick, K. Saenko, D. Klein, and T. Darrell, “Speaker-follower models for vision-and-language navigation,” Advances in neural information processing systems , vol. 31, 2018
2018
Earlier work this paper cites.
X. Wang, W. Xiong, H. Wang, and W. Y. Wang, “Look before you leap: Bridging model-free and model-based reinforcement learning for planned-ahead vision-and-language navigation,” in Proceedings of the European Conference on Computer Vision (ECCV) , 2018, pp. 37–53
2018
Earlier work this paper cites.
W. Jeamwatthanachai, M. Wald, and G. Wills, “Indoor navigation by blind people: Behaviors and challenges in unfamiliar spaces and buildings,” British Journal of Visual Impairment , vol. 37, no. 2, pp. 140–153, 2019
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
X. Li, A. Yuan, and X. Lu, “Vision-to-language tasks based on attributes and attention mechanism,” IEEE transactions on cybernetics , vol. 51, no. 2, pp. 913–926, 2019
2019
Earlier work this paper cites.
D. Ahmetovic, J. Guerreiro, E. Ohn-Bar, K. M. Kitani, and C. Asakawa, “Impact of expertise on interaction preferences for navigation assistance of visually impaired individuals,” in Proceedings of the 16th International Web for All Conference , 2019, pp. 1–9
2019
Earlier work this paper cites.
Z. Başgöze, J. Gualtieri, M. T. Sachs, and E. A. Cooper, “Navigational aid use by individuals with visual impairments,” in Journal on technology and persons with disabilities:… Annual International Technology and Persons with Disabilities Conference , vol. 8. NIH Public Access, 2020, p. 22
2020
Earlier work this paper cites.
N. A. Giudice, B. A. Guenther, T. M. Kaplan, S. M. Anderson, R. J. Knuesel, and J. F. Cioffi, “Use of an indoor navigation system by sighted and blind travelers: Performance similarities across visual status and age,” ACM Transactions on Accessible Computing (TACCESS) , vol. 13, no. 3, pp. 1–27, 2020
2020
Earlier work this paper cites.
M. M. Islam, M. S. Sadi, and T. Bräunl, “Automated walking guide to enhance the mobility of visually impaired people,” IEEE Transactions on Medical Robotics and Bionics , vol. 2, no. 3, pp. 485–496, 2020
2020
Earlier work this paper cites.
H. Wang, Q. Wu, and C. Shen, “Soft expert reward learning for vision-and-language navigation,” in Computer Vision–ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part IX 16 . Springer, 2020, pp. 126–141
2020
Earlier work this paper cites.
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark et al. , “Learning transferable visual models from natural language supervision,” in International conference on machine learning . PMLR, 2021, pp. 8748–8763
2021
Earlier work this paper cites.
P. Slade, A. Tambe, and M. J. Kochenderfer, “Multimodal sensing and intuitive steering assistance improve navigation and mobility for people with impaired vision,” Science Robotics , vol. 6, no. 59, 10 2021
2021
Earlier work this paper cites.
L. Jin, H. Zhang, and C. Ye, “A wearable robotic device for assistive navigation and object manipulation,” in 2021 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) . IEEE, 2021, pp. 765–770
2021
Earlier work this paper cites.
Y. Bouteraa, “Design and development of a wearable assistive device integrating a fuzzy decision support system for blind and visually impaired people,” Micromachines , vol. 12, no. 9, p. 1082, 2021
2021
Earlier work this paper cites.
A. B. Vasudevan, D. Dai, and L. Van Gool, “Talk2nav: Long-range vision-and-language navigation with dual attention and spatial memory,” International Journal of Computer Vision , vol. 129, pp. 246–266, 2021
2021
Cited alongside, same era.
2021
Cited alongside, same era.
G. Li, J. Xu, Z. Li, C. Chen, and Z. Kan, “Sensing and navigation of wearable assistance cognitive systems for the visually impaired,” IEEE Transactions on Cognitive and Developmental Systems , vol. 15, no. 1, pp. 122–133, 2022
2022
Cited alongside, same era.
2022
Cited alongside, same era.
X. Jiang, Y. Dong, L. Wang, F. Zheng, Q. Shang, G. Li, Z. Jin, and W. Jiao, “Self-planning Code Generation with Large Language Models,” ACM Transactions on Software Engineering and Methodology , 6 2024
2024
Closest in time.
Y. Cui, S. Huang, J. Zhong, Z. Liu, Y. Wang, C. Sun, B. Li, X. Wang, and A. Khajepour, “Drivellm: Charting the Path Toward Full Autonomous Driving With Large Language Models,” IEEE Transactions on Intelligent Vehicles , vol. 9, no. 1, pp. 1450–1464, 1 2024
2024
Closest in time.
S. Liu, A. Hasan, K. Hong, R. Wang, P. Chang, Z. Mizrachi, J. Lin, D. L. McPherson, W. A. Rogers, and K. Driggs-Campbell, “Dragon: A dialogue-based robot for assistive navigation with visual language grounding,” IEEE Robotics and Automation Letters , 2024
2024
Closest in time.
S. Chen, X. Chen, C. Zhang, M. Li, G. Yu, H. Fei, H. Zhu, J. Fan, and T. Chen, “Ll3da: Visual interactive instruction tuning for omni-3d understanding reasoning and planning,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, pp. 26 428–26 438
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
W. Huang, P. Abbeel, D. Pathak, and I. Mordatch, “Language models as zero-shot planners: Extracting actionable knowledge for embodied agents,” in International conference on machine learning . PMLR, 2022, pp. 9118–9147
2022
Cited alongside, same era.
2023
Cited alongside, same era.
J. Li, D. Li, S. Savarese, and S. Hoi, “Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models,” in International conference on machine learning . PMLR, 2023, pp. 19 730–19 742
2023
Cited alongside, same era.
K. Rana, J. Haviland, S. Garg, J. Abou-Chakra, I. Reid, and N. Suenderhauf, “Sayplan: Grounding large language models using 3d scene graphs for scalable robot task planning,” in 7th Annual Conference on Robot Learning , 2023
2023
Cited alongside, same era.
D. Shah, B. Osiński, S. Levine et al. , “Lm-nav: Robotic navigation with large pre-trained models of language, vision, and action,” in Conference on robot learning . PMLR, 2023, pp. 492–504
2023
Cited alongside, same era.
C. Huang, O. Mees, A. Zeng, and W. Burgard, “Visual language maps for robot navigation,” in 2023 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2023, pp. 10 608–10 615
2023
Cited alongside, same era.
2023
Cited alongside, same era.
Y. Hong, H. Zhen, P. Chen, S. Zheng, Y. Du, Z. Chen, and C. Gan, “3d-llm: Injecting the 3d world into large language models,” Advances in Neural Information Processing Systems , vol. 36, pp. 20 482–20 494, 2023
2023
Cited alongside, same era.
2024
Closest in time.
S. Shafique, W. Setti, C. Campus, S. Zanchi, A. Del Bue, and M. Gori, “How path integration abilities of blind people change in different exploration conditions,” Frontiers in Neuroscience , vol. 18, p. 1375225, 2024
2024
Closest in time.
G. Zhou, Y. Hong, and Q. Wu, “Navgpt: Explicit reasoning in vision-and-language navigation with large language models,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 38, no. 7, 2024, pp. 7641–7649
2024
Closest in time.
X. Liu, B. Wang, and Z. Li, “Vision-based wearable steering assistance for people with impaired vision in jogging,” in 2024 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2024, pp. 15 270–15 275
2024
Closest in time.
P. Sermanet, T. Ding, J. Zhao, F. Xia, D. Dwibedi, K. Gopalakrishnan, C. Chan, G. Dulac-Arnold, S. Maddineni, N. J. Joshi et al. , “Robovqa: Multimodal long-horizon reasoning for robotics,” in 2024 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2024, pp. 645–652
2024
Closest in time.
A. Rajvanshi, K. Sikka, X. Lin, B. Lee, H.-P. Chiu, and A. Velasquez, “Saynav: Grounding large language models for dynamic planning to navigation in new environments,” in Proceedings of the International Conference on Automated Planning and Scheduling , vol. 34, 2024, pp. 464–474
2024
Closest in time.
J. Chen, B. Lin, R. Xu, Z. Chai, X. Liang, and K.-Y. Wong, “Mapgpt: Map-guided prompting with adaptive path planning for vision-and-language navigation,” in Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , 2024, pp. 9796–9810
2024
Closest in time.
V. S. Dorbala, J. F. Mullen, and D. Manocha, “ Can an Embodied Agent Find Your “Cat-shaped Mug”?
2024
Closest in time.
2024
Closest in time.
R. Schumann, W. Zhu, W. Feng, T.-J. Fu, S. Riezler, and W. Y. Wang, “Velma: Verbalization embodiment of llm agents for vision and language navigation in street view,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 38, no. 17, 2024, pp. 18 924–18 933
2024
Closest in time.
OpenAI, “Hello gpt-4,” https://openai.com/index/hello-gpt-4o/ , accessed: Sept. 14, 2024
2024
Closest in time.
M. AI, “Llama 3.2: Revolutionizing edge ai and vision with open, customizable models,” https://ai.meta.com/blog/llama-3-2-connect-2024-vision-edge-mobile-devices/ , September 2024
2024
Closest in time.
2024
Closest in time.
Anthropic, “Introducing claude 3.5 sonnet,” https://www.anthropic.com/news/claude-3-5-sonnet , June 2024, accessed: February 26, 2025
2025
Closest in time.
2025
Closest in time.