Fetching the paper…
Reading the bibliography…
Navigating and understanding complex environments over extended periods of time is a significant challenge for robots.
A. Das, S. Datta, G. Gkioxari, S. Lee, D. Parikh, and D. Batra, “Embodied Question Answering,” Conference on Computer Vision and Pattern Recognition (CVPR) , 2018
2018
Earlier work this paper cites.
A. Das, G. Gkioxari, S. Lee, D. Parikh, and D. Batra, “Neural Modular Control for Embodied Question Answering,” Conference on Robot Learning (CoRL) , 2018
2018
Earlier work this paper cites.
P. Anderson, Q. Wu, D. Teney, J. Bruce, M. Johnson, N. Sünderhauf, I. Reid, S. Gould, and A. Van Den Hengel, “Vision-and-Language Navigation: Interpreting Visually-Grounded Navigation Instructions in Real Environments,” Conference on Computer Vision and Pattern Recognition (CVPR) , 2018
2018
Earlier work this paper cites.
N. Savinov, A. Dosovitskiy, and V. Koltun, “Semi-Parametric Topological Memory for Navigation,” International Conference on Learning Representations (ICLR) , 2018
2018
Earlier work this paper cites.
J. Thomason, D. Gordon, and Y. Bisk, “Shifting the Baseline: Single Modality Performance on Visual Navigation & QA ,” North American Association for Computational Linguistics (NAACL) , 2019
2019
Earlier work this paper cites.
E. Wijmans, S. Datta, O. Maksymets, A. Das, G. Gkioxari, S. Lee, I. Essa, D. Parikh, and D. Batra, “Embodied Question Answering in Photorealistic Environments with Point Cloud Perception,” Conference on Computer Vision and Pattern Recognition (CVPR) , 2019
2019
Earlier work this paper cites.
B. Eysenbach, R. R. Salakhutdinov, and S. Levine, “Search on the Replay Buffer: Bridging Planning and Reinforcement Learning,” Conference on Neural Information Processing Systems (NeurIPS) , 2019
2019
Earlier work this paper cites.
K. Fang, A. Toshev, L. Fei-Fei, and S. Savarese, “Scene Memory Transformer for Embodied Agents in Long-Horizon Tasks,” Conference on Computer Vision and Pattern Recognition (CVPR) , 2019
2019
Earlier work this paper cites.
J. Thomason, M. Murray, M. Cakmak, and L. Zettlemoyer, “Vision-and-Dialog Navigation,” Conference on Robot Learning (CoRL) , 2020
2020
Earlier work this paper cites.
D. S. Chaplot, D. P. Gandhi, A. Gupta, and R. R. Salakhutdinov, “Object Goal Navigation Using Goal-Oriented Semantic Exploration,” Conference on Neural Information Processing Systems (NeurIPS) , 2020
2020
Earlier work this paper cites.
J. Wald, H. Dhamo, N. Navab, and F. Tombari, “Learning 3D Semantic Scene Graphs from 3D Indoor Reconstructions,” Conference on Computer Vision and Pattern Recognition (CVPR) , 2020
2020
Earlier work this paper cites.
K. Grauman, A. Westbury, E. Byrne, Z. Chavis, A. Furnari, R. Girdhar, J. Hamburger, H. Jiang, M. Liu, X. Liu et al. , “Ego4D: Around the World in 3,000 Hours of Egocentric Video,” Conference on Computer Vision and Pattern Recognition (CVPR) , 2022
2022
Earlier work this paper cites.
Z. Gan, L. Li, C. Li, L. Wang, Z. Liu, J. Gao et al. , “Vision-Language Pre-Training: Basics, Recent Advances, and Future Trends,” Foundations and Trends in Computer Graphics and Vision , 2022
2022
Earlier work this paper cites.
A. Majumdar, G. Aggarwal, B. Devnani, J. Hoffman, and D. Batra, “ZSON: Zero-Shot Object-Goal Navigation Using Multimodal Goal Embeddings,” Conference on Neural Information Processing Systems (NeurIPS) , 2022
2022
Earlier work this paper cites.
R. Ramrakhya, E. Undersander, D. Batra, and A. Das, “Habitat-Web: Learning Embodied Object-Search Strategies from Human Demonstrations at Scale,” Conference on Computer Vision and Pattern Recognition (CVPR) , 2022
2022
Earlier work this paper cites.
D. Shah, B. Osiński, S. Levine et al. , “LM-Nav: Robotic Navigation with Large Pre-Trained Models of Language, Vision, and Action,” Conference on Robot Learning (CoRL) , 2022
2022
Earlier work this paper cites.
V. S. Dorbala, G. Sigurdsson, R. Piramuthu, J. Thomason, and G. S. Sukhatme, “CLIP-Nav: Using CLIP for Zero-Shot Vision-and-Language Navigation,” Workshop on Language and Robot Learning (LangRob) @ CoRL , 2022
2022
Earlier work this paper cites.
X. Li, D. Guo, H. Liu, and F. Sun, “Embodied Semantic Scene Graph Generation,” Conference on Robot Learning (CoRL) , 2022
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
J. Wei, X. Wang, D. Schuurmans, M. Bosma, F. Xia, E. Chi, Q. V. Le, D. Zhou et al. , “Chain-of-Thought Prompting Elicits Reasoning in Large Language Models,” Conference on Neural Information Processing Systems (NeurIPS) , 2022
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2022
Cited alongside, same era.
W. Huang, F. Xia, T. Xiao, H. Chan, J. Liang, P. Florence, A. Zeng, J. Tompson, I. Mordatch, Y. Chebotar et al. , “Inner Monologue: Embodied Reasoning through Planning with Language Models,” Conference on Robot Learning (CoRL) , 2022
2022
J. Liang, W. Huang, F. Xia, P. Xu, K. Hausman, B. Ichter, P. Florence, and A. Zeng, “Code as policies: Language model programs for embodied control,” International Conference on Robotics and Automation (ICRA) , 2023
2023
Later among the works it cites.
A. Radford, J. W. Kim, T. Xu, G. Brockman, C. McLeavey, and I. Sutskever, “Robust Speech Recognition via Large-Scale Weak Supervision,” in International Conference on Machine Learning (ICML) , 2023
2023
Later among the works it cites.
P. Sermanet, T. Ding, J. Zhao, F. Xia, D. Dwibedi, K. Gopalakrishnan, C. Chan, G. Dulac-Arnold, S. Maddineni, N. J. Joshi, P. Florence, W. Han, R. Baruch, Y. Lu, S. Mirchandani, P. Xu, P. Sanketi, K. Hausman, I. Shafran, B. Ichter, and Y. Cao, “RoboVQA: Multimodal Long-Horizon Reasoning for Robotics,” International Conference on Robotics and Automation (ICRA) , 2024
2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
B. Chen, F. Xia, B. Ichter, K. Rao, K. Gopalakrishnan, M. S. Ryoo, A. Stone, and D. Kappler, “Open-Vocabulary Queryable Scene Representations for Real World Planning,” International Conference on Robotics and Automation (ICRA) , 2023
2023
Cited alongside, same era.
J. Krantz, S. Banerjee, W. Zhu, J. Corso, P. Anderson, S. Lee, and J. Thomason, “Iterative Vision-and-Language Navigation,” Conference on Computer Vision and Pattern Recognition (CVPR) , 2023
2023
Cited alongside, same era.
S. Y. Gadre, M. Wortsman, G. Ilharco, L. Schmidt, and S. Song, “Cows on Pasture: Baselines and Benchmarks for Language-Driven Zero-Shot Object Navigation,” Conference on Computer Vision and Pattern Recognition (CVPR) , 2023
2023
Cited alongside, same era.
D. Shah, M. R. Equi, B. Osiński, F. Xia, B. Ichter, and S. Levine, “Navigation with Large Language Models: Semantic Guesswork as a Heuristic for Planning,” Conference on Robot Learning (CoRL) , 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
S. Yao, D. Yu, J. Zhao, I. Shafran, T. Griffiths, Y. Cao, and K. Narasimhan, “Tree of Thoughts: Deliberate Problem Solving with Large Language Models,” Conference on Neural Information Processing Systems (NeurIPS) , 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2024
Closest in time.
A. Majumdar, A. Ajay, X. Zhang, P. Putta, S. Yenamandra, M. Henaff, S. Silwal, P. Mcvay, O. Maksymets, S. Arnaud et al. , “OpenEQA: Embodied Question Answering in the Era of Foundation Models,” Conference on Computer Vision and Pattern Recognition (CVPR) , 2024
2024
Closest in time.
2024
Closest in time.
G. Zhou, Y. Hong, and Q. Wu, “NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models,” AAAI Conference on Artificial Intelligence , 2024
2024
Closest in time.
2024
Closest in time.
Z. Hu, F. Lucchetti, C. Schlesinger, Y. Saxena, A. Freeman, S. Modak, A. Guha, and J. Biswas, “Deploying and Evaluating LLMs to Program Service Mobile Robots,” IEEE Robotics and Automation Letters (RA-L) , 2024
2024
Closest in time.
J. Lin, H. Yin, W. Ping, P. Molchanov, M. Shoeybi, and S. Han, “VILA: On Pre-Training for Visual Language Models,” Conference on Computer Vision and Pattern Recognition (CVPR) , 2024
2024
Closest in time.
S. Lee, A. Shakir, D. Koenig, and J. Lipp, “Open source strikes bread - new fluffy embeddings model,” 2024. [Online]. Available: https://www.mixedbread.ai/blog/mxbai-embed-large-v1
2024
Closest in time.
A. Zhang, C. Eranki, C. Zhang, J.-H. Park, R. Hong, P. Kalyani, L. Kalyanaraman, A. Gamare, A. Bagad, M. Esteva et al. , “Towards Robust Robot 3D Perception in Urban Environments: The UT Campus Object Dataset,” IEEE Transactions on Robotics (T-RO) , 2024
2024
Closest in time.
Clearpath Robotics, “Husky UGV - Outdoor Field Research Robot,” 2024. [Online]. Available: https://clearpathrobotics.com/husky-unmanned-ground-vehicle-robot/
2024
Closest in time.
Mistral AI, “Codestral model card,” 2024. [Online]. Available: https://huggingface.co/mistralai/Codestral-22B-v0.1
2024
Closest in time.
Cohere for AI, “Command-r model card,” 2024. [Online]. Available: https://huggingface.co/CohereForAI/c4ai-command-r-v01
2024
Closest in time.
2024
Closest in time.
Segway Robotics, “Nova Carter - Complete Robotics Development Platform,” 2024. [Online]. Available: https://robotics.segway.com/nova-carter/
2024
Closest in time.
NVIDIA, “NVIDIA NIM,” 2024. [Online]. Available: https://docs.nvidia.com/nim/index.html
2024
Closest in time.