Fetching the paper…
Reading the bibliography…
Navigation in unfamiliar environments presents a major challenge for robots: while mapping and planning techniques can be used to build up a representation of the world, quickly discovering a path to a desired goal in unfamiliar settings with such methods often requires lengthy mapping and exploration.
A frontier-based approach for autonomous exploration
B. Yamauchi · 1997
Earlier work this paper cites.
Vision for mobile robot navigation: A survey
G. N. DeSouza and A. C. Kak · 2002
Earlier work this paper cites.
Vision-and-language navigation: Interpreting visually-grounded navigation instructions in real environments
P. Anderson, Q. Wu, D. Teney, J. Bruce, M. Johnson, N. Sünderhauf, I. Reid, S. Gould, and A. van den Hengel · 2018
Earlier work this paper cites.
Semi-parametric topological memory for navigation
N. Savinov, A. Dosovitskiy, and V. Koltun · 2018
Earlier work this paper cites.
Deep visual MPC-policy learning for navigation
N. Hirose, F. Xia, R. Martín-Martín, A. Sadeghian, and S. Savarese · 2019
Earlier work this paper cites.
Improving vision-and-language navigation with image-text pairs from the web
A. Majumdar, A. Shrivastava, S. Lee, P. Anderson, D. Parikh, and D. Batra · 2020
Earlier work this paper cites.
DD-PPO: Learning Near-Perfect PointGoal Navigators from 2.5 Billion Frames
E. Wijmans, A. Kadian, A. Morcos, S. Lee, I. Essa, D. Parikh, M. Savva, and D. Batra · 2020
Earlier work this paper cites.
Semantic curiosity for active visual learning
D. S. Chaplot, H. Jiang, S. Gupta, and A. Gupta · 2020
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, et al · 2021
Earlier work this paper cites.
Cliport: What and where pathways for robotic manipulation
M. Shridhar, L. Manuelli, and D. Fox · 2021
Earlier work this paper cites.
Rapid exploration for open-world navigation with latent goal models
D. Shah, B. Eysenbach, N. Rhinehart, and S. Levine · 2021
Earlier work this paper cites.
GNM: A General Navigation Model to Drive Any Robot
D. Shah, A. Sridhar, A. Bhorkar, N. Hirose, and S. Levine · 2022
Earlier work this paper cites.
Poni: Potential functions for objectgoal navigation with interaction-free learning
S. K. Ramakrishnan, D. S. Chaplot, Z. Al-Halah, J. Malik, and K. Grauman · 2022
Earlier work this paper cites.
Pali: A jointly-scaled multilingual language-image model
X. Chen, X. Wang, S. Changpinyo, A. Piergiovanni, P. Padlewski, D. Salz, S. Goodman, A. Grycner, B. Mustafa, L. Beyer, et al · 2022
Earlier work this paper cites.
Zson: Zero-shot object-goal navigation using multimodal goal embeddings
A. Majumdar, G. Aggarwal, B. Devnani, J. Hoffman, and D. Batra · 2022
Cited alongside, same era.
Open-vocabulary queryable scene representations for real world planning
B. Chen, F. Xia, B. Ichter, K. Rao, K. Gopalakrishnan, M. S. Ryoo, A. Stone, and D. Kappler · 2022
Cited alongside, same era.
Visual language maps for robot navigation
C. Huang, O. Mees, A. Zeng, and W. Burgard · 2022
Cited alongside, same era.
LM-nav: Robotic navigation with large pre-trained models of language, vision, and action
D. Shah, B. Osinski, B. Ichter, and S. Levine · 2022
Cited alongside, same era.
Progprompt: Generating situated robot task plans using large language models, 2022
I. Singh, V. Blukis, A. Mousavian, A. Goyal, D. Xu, J. Tremblay, D. Fox, J. Thomason, and A. Garg · 2022
Conceptfusion: Open-set multimodal 3d mapping
K. Jatavallabhula, A. Kuwajerwala, Q. Gu, M. Omama, T. Chen, S. Li, G. Iyer, S. Saryazdi, N. Keetha, A. Tewari, J. Tenenbaum, C. de Melo, M. Krishna, L. Paull, F. Shkurti, and A. Torralba · 2023
Closest in time.
Can an Embodied Agent Find Your ”Cat-shaped Mug”? LLM-Based Zero-Shot Object Navigation, 2023
V. S. Dorbala, J. F. J. Mullen, and D. Manocha · 2023
Closest in time.
Grounding language with visual affordances over unstructured data, 2023
O. Mees, J. Borja-Diaz, and W. Burgard · 2023
Closest in time.
Translating natural language to planning goals with large-language models
Y. Xie, C. Yu, T. Zhu, J. Bai, Z. Gong, and H. Soh · 2023
Closest in time.
Task and motion planning with large language models for object rearrangement, 2023
Y. Ding, X. Zhang, C. Paxton, and S. Zhang · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Language models as zero-shot planners: Extracting actionable knowledge for embodied agents
W. Huang, P. Abbeel, D. Pathak, and I. Mordatch · 2022
Cited alongside, same era.
Do as i can, not as i say: Grounding language in robotic affordances
B. Ichter, A. Brohan, Y. Chebotar, C. Finn, K. Hausman, A. Herzog, D. Ho, J. Ibarz, A. Irpan, E. Jang, R. Julian, D. Kalashnikov, S. Levine, Y. Lu, C. Parada, K. Rao, P. Sermanet, A. T. Toshev, V. Vanhoucke, F. Xia, T. Xiao, P. Xu, M. Yan, N. Brown, M. Ahn, O. Cortes, N. Sievers, C. Tan, S. Xu, D. Reyes, J. Rettinghouse, J. Quiambao, P. Pastor, L. Luu, K.-H. Lee, Y. Kuang, S. Jesmonth, K. Jeffrey, R. J. Ruano, J. Hsu, K. Gopalakrishnan, B. David, A. Zeng, and C. K. Fu · 2022
Cited alongside, same era.
Vima: General robot manipulation with multimodal prompts
Y. Jiang, A. Gupta, Z. Zhang, G. Wang, Y. Dou, Y. Chen, L. Fei-Fei, A. Anandkumar, Y. Zhu, and L. Fan · 2022
Cited alongside, same era.
Chain of thought prompting elicits reasoning in large language models
J. Wei, X. Wang, D. Schuurmans, M. Bosma, B. Ichter, F. Xia, E. H. Chi, Q. V. Le, and D. Zhou · 2022
Cited alongside, same era.
Viking: Vision-based kilometer-scale navigation with geographic hints
D. Shah and S. Levine · 2022
Cited alongside, same era.
Detecting twenty-thousand classes using image-level supervision
X. Zhou, R. Girdhar, A. Joulin, P. Krähenbühl, and I. Misra · 2022
Cited alongside, same era.
Habitat challenge 2022
K. Yadav, S. K. Ramakrishnan, J. Turner, A. Gokaslan, O. Maksymets, R. Jain, R. Ramrakhya, A. X. Chang, A. Clegg, M. Savva, E. Undersander, D. S. Chaplot, and D. Batra · 2022
Cited alongside, same era.
Text2motion: From natural language instructions to feasible plans, 2023
K. Lin, C. Agia, T. Migimatsu, M. Pavone, and J. Bohg · 2023
Closest in time.
Large language models still can’t plan (a benchmark for llms on planning and reasoning about change), 2023
K. Valmeekam, A. Olmo, S. Sreedharan, and S. Kambhampati · 2023
Closest in time.
Grounded decoding: Guiding text generation with grounded models for robot control, 2023
W. Huang, F. Xia, D. Shah, D. Driess, A. Zeng, Y. Lu, P. Florence, I. Mordatch, S. Levine, K. Hausman, and B. Ichter · 2023
Closest in time.
Language is not all you need: Aligning perception with language models, 2023
S. Huang, L. Dong, W. Wang, Y. Hao, S. Singhal, S. Ma, T. Lv, L. Cui, O. K. Mohammed, B. Patra, Q. Liu, K. Aggarwal, Z. Chi, J. Bjorck, V. Chaudhary, S. Som, X. Song, and F. Wei · 2023
Closest in time.
Palm-e: An embodied multimodal language model
D. Driess, F. Xia, M. S. M. Sajjadi, C. Lynch, A. Chowdhery, B. Ichter, A. Wahid, J. Tompson, Q. Vuong, T. Yu, W. Huang, Y. Chebotar, P. Sermanet, D. Duckworth, S. Levine, V. Vanhoucke, K. Hausman, M. Toussaint, K. Greff, A. Zeng, I. Mordatch, and P. Florence · 2023
Closest in time.
Navigating to objects in the real world
T. Gervet, S. Chintala, D. Batra, J. Malik, and D. S. Chaplot · 2023
Closest in time.
L3mvn: Leveraging large language models for visual target navigation, 2023
B. Yu, H. Kasaei, and M. Cao · 2023
Closest in time.
NoMaD: Goal Masked Diffusion Policies for Navigation and Exploration
A. Sridhar, D. Shah, C. Glossop, and S. Levine · 2023
Closest in time.
Ovrl-v2: A simple state-of-art baseline for imagenav and objectnav, 2023
K. Yadav, A. Majumdar, R. Ramrakhya, N. Yokoyama, A. Baevski, Z. Kira, O. Maksymets, and D. Batra · 2023
Closest in time.