Fetching the paper…
Reading the bibliography…
Humans can flexibly interpret and compose different goal specifications, such as language instructions, spatial coordinates, or visual references, when navigating to a destination.
1903
Earlier work this paper cites.
H. Xu, Y. Gao, F. Yu, and T. Darrell, “End-to-end learning of driving models from large-scale video datasets,” in
2017
Earlier work this paper cites.
N. Savinov, A. Dosovitskiy, and V. Koltun, “Semi-parametric topological memory for navigation,”
2018
Earlier work this paper cites.
G. Kahn, A. Villaflor, B. Ding, P. Abbeel, and S. Levine, “Self-supervised deep reinforcement learning with generalized computation graphs for robot navigation,” in
2018
Earlier work this paper cites.
N. Hirose, A. Sadeghian, M. Vázquez, P. Goebel, and S. Savarese, “Gonet: A semi-supervised deep learning approach for traversability estimation,” in
2018
Earlier work this paper cites.
N. Hirose, F. Xia, R. Martín-Martín, A. Sadeghian, and S. Savarese, “Deep visual mpc-policy learning for navigation,”
2019
Earlier work this paper cites.
2021
Earlier work this paper cites.
2022
Earlier work this paper cites.
D. Shah and S. Levine, “Viking: Vision-based kilometer-scale navigation with geographic hints,”
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
O. Mees, L. Hermann, and W. Burgard, “What matters in language conditioned robotic imitation learning over unstructured data,”
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
T. Niwa, S. Taguchi, and N. Hirose, “Spatio-temporal graph localization networks for image-based navigation,” in
2022
Earlier work this paper cites.
M. Minderer, A. Gritsenko, A. Stone, M. Neumann, D. Weissenborn, A. Dosovitskiy, A. Mahendran, A. Arnab, M. Dehghani, Z. Shen,
2022
Earlier work this paper cites.
N. Hirose and K. Tahara, “Depth360: Self-supervised learning for monocular depth estimation using learnable camera distortion model,” in
2022
Earlier work this paper cites.
S. Triest, M. Sivaprakasam, S. J. Wang, W. Wang, A. M. Johnson, and S. Scherer, “Tartandrive: A large-scale dataset for learning off-road dynamics models,” in
2022
Cited alongside, same era.
A. Shaban, X. Meng, J. Lee, B. Boots, and D. Fox, “Semantic terrain classification for off-road autonomous driving,” in
2022
Cited alongside, same era.
H. Karnan, A. Nair, X. Xiao, G. Warnell, S. Pirk, A. Toshev, J. Hart, J. Biswas, and P. Stone, “Socially compliant navigation dataset (scand): A large-scale dataset of demonstrations for social navigation,”
2022
Cited alongside, same era.
D. Shah, A. Sridhar, A. Bhorkar, N. Hirose, and S. Levine, “Gnm: A general navigation model to drive any robot,” in
2023
Cited alongside, same era.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Z. Xu, H.-T. L. Chiang, Z. Fu, M. G. Jacob, T. Zhang, T.-W. E. Lee, W. Yu, C. Schenck, D. Rendleman, D. Shah,
2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2023
Cited alongside, same era.
S. Y. Gadre, M. Wortsman, G. Ilharco, L. Schmidt, and S. Song, “Cows on pasture: Baselines and benchmarks for language-driven zero-shot object navigation,” in
2023
Cited alongside, same era.
D. Shah, B. Osiński, S. Levine,
2023
Cited alongside, same era.
D. Driess, F. Xia, M. S. Sajjadi, C. Lynch, A. Chowdhery, B. Ichter, A. Wahid, J. Tompson, Q. Vuong, T. Yu,
2023
Cited alongside, same era.
V. Myers, A. He, K. Fang, H. Walke, P. Hansen-Estruch, C.-A. Cheng, M. Jalobeanu, A. Kolobov, A. Dragan, and S. Levine, “Goal representations for instruction following: A semi-supervised language interface to control,” 2023
2023
Cited alongside, same era.
N. Hirose, D. Shah, A. Sridhar, and S. Levine, “Sacson: Scalable autonomous control for social navigation,”
2023
Cited alongside, same era.
J. Lin, H. Yin, W. Ping, P. Molchanov, M. Shoeybi, and S. Han, “Vila: On pre-training for visual language models,” in
2024
Cited alongside, same era.
2024
Cited alongside, same era.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
FrodoBots, “Frodobots-2k,” 2024. [Online]. Available:
2024
Later among the works it cites.
S. Belkhale and D. Sadigh, “Minivla: A better vla with a smaller footprint,” 2024. [Online]. Available:
2024
Later among the works it cites.
M. Deitke, C. Clark, S. Lee, R. Tripathi, Y. Yang, J. S. Park, M. Salehi, N. Muennighoff, K. Lo, L. Soldaini,
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
“EarthRover Zero,”
2025
Closest in time.