Fetching the paper…
Reading the bibliography…
This paper presents UnderwaterVLA, a novel framework for autonomous underwater navigation that integrates multimodal foundation models with embodied intelligence systems.
E. Bovio, D. Cecchi, and F. Baralli, “Autonomous underwater vehicles for scientific and naval operations,” Annual Reviews in Control , vol. 30, no. 2, pp. 117–130, 2006
2006
Earlier work this paper cites.
M. Stojanovic, “Underwater acoustic communications: Design considerations on the physical layer,” in 2008 Fifth Annual Conference on Wireless on Demand Network Systems and Services . IEEE, 2008, pp. 1–10
2008
Earlier work this paper cites.
T. I. Fossen, Handbook of marine craft hydrodynamics and motion control . John Wiley & Sons, 2011
2011
Earlier work this paper cites.
G. Antonelli, “Underwater robots, volume 96 of springer tracts in advanced robotics,” 2014
2014
Earlier work this paper cites.
J. J. Leonard and A. Bahr, “Autonomous underwater vehicle navigation,” Springer handbook of ocean engineering , pp. 341–358, 2016
2016
Earlier work this paper cites.
Y. Cong, C. Gu, T. Zhang, and Y. Gao, “Underwater robot sensing technology: A survey,” Fundamental Research , vol. 1, no. 3, pp. 337–345, 2021
2021
Earlier work this paper cites.
J. McConnell, I. Collado-Gonzalez, and B. Englot, “Perception for underwater robots,” Current Robotics Reports , vol. 3, no. 4, pp. 177–186, 2022
2022
Earlier work this paper cites.
B. Zhang, D. Ji, S. Liu, X. Zhu, and W. Xu, “Autonomous underwater vehicle navigation: A review,” Ocean Engineering , vol. 273, p. 113861, 2023
2023
Earlier work this paper cites.
W. Lu, K. Cheng, and M. Hu, “Reinforcement learning for autonomous underwater vehicles via data-informed domain randomization,” Applied Sciences , vol. 13, no. 3, p. 1723, 2023
2023
Cited alongside, same era.
Y. Qiao, J. Yin, W. Wang, F. Duarte, J. Yang, and C. Ratti, “Survey of deep learning for autonomous surface vehicles in marine environments,” IEEE Transactions on Intelligent Transportation Systems , vol. 24, no. 4, pp. 3678–3701, 2023
2023
Cited alongside, same era.
C. Lynch, A. Wahid, J. Tompson, T. Ding, J. Betker, R. Baruch, T. Armstrong, and P. Florence, “Interactive language: Talking to robots in real time,” IEEE Robotics and Automation Letters , 2023
2023
Cited alongside, same era.
P. Ding, H. Zhao, W. Zhang, W. Song, M. Zhang, S. Huang, N. Yang, and D. Wang, “Quar-vla: Vision-language-action model for quadruped robots,” in European Conference on Computer Vision . Springer, 2024, pp. 352–367
2024
Cited alongside, same era.
Y. Fan, P. Ding, S. Bai, X. Tong, Y. Zhu, H. Lu, F. Dai, W. Zhao, Y. Liu, S. Huang et al. , “Long-vla: Unleashing long-horizon capability of vision language action model for robot manipulation,” arXiv e-prints , pp. arXiv–2508, 2025
2025
Closest in time.
S. T. Myers, “The vla sky survey (vlass) and beyond: Lessons, challenges, and future surveys,” in 2025 United States National Committee of URSI National Radio Science Meeting (USNC-URSI NRSM) , 2025, pp. 283–283
2025
Closest in time.
O. Sautenkov, Y. Yaqoot, A. Lykov, M. A. Mustafa, G. Tadevosyan, A. Akhmetkazy, M. A. Cabrera, M. Martynov, S. Karaf, and D. Tsetserukou, “Uav-vla: Vision-language-action system for large scale aerial mission generation,” in 2025 20th ACM/IEEE International Conference on Human-Robot Interaction (HRI) . IEEE, 2025, pp. 1588–1592
2025
Closest in time.
Y. Li, “Vision-language-action models: Foundations, techniques and applications,” in 2025 17th International Conference on Intelligent Human-Machine Systems and Cybernetics (IHMSC) , 2025, pp. 174–177
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
P. White, R. Jensen, J. Bradfield, C. Thompson, and D. Copeland, “New horizons vla experiment,” in 2024 IEEE Aerospace Conference . IEEE, 2024, pp. 1–10
2024
Cited alongside, same era.
2025
Cited alongside, same era.
W. Biao, L. Ruilong, W. Fang, and C. Weicheng, “Research status and development trends of deep-sea unmanned equipment control system,” Journal of Unmanned Underwater Systems , vol. 33, no. 3, pp. 390–399, 2025
2025
Cited alongside, same era.
2025
Closest in time.
J. Wen, Y. Zhu, M. Zhu, Z. Tang, J. Li, Z. Zhou, X. Liu, C. Shen, Y. Peng, and F. Feng, “Diffusionvla: Scaling robot foundation models via unified diffusion and autoregression,” in Forty-second International Conference on Machine Learning , 2025
2025
Closest in time.
M. Elmezain, L. S. Saoud, A. Sultan, M. Heshmat, L. Seneviratne, and I. Hussain, “Advancing underwater vision: a survey of deep learning models for underwater object recognition and tracking,” IEEE Access , 2025
2025
Closest in time.
H. Yin, H. Wei, X. Xu, W. Guo, J. Zhou, and J. Lu, “GC-VLN: Instruction as graph constraints for training-free vision-and-language navigation,” arXiv preprint , 2025, accepted to CoRL 2025. [Online]. Available: https://doi.org/10.48550/arXiv.2509.10454
2025
Closest in time.