Fetching the paper…
Reading the bibliography…
Mobile robots exploring indoor environments increasingly rely on vision-language models to perceive high-level semantic cues in camera images, such as object categories.
High resolution maps from wide angle sonar
H. Moravec and A. Elfes · 1985
Earlier work this paper cites.
A Frontier-Based Approach for Autonomous Exploration
B. Yamauchi · 1997
Earlier work this paper cites.
Efficient Global Optimization of Expensive Black-Box Functions
D.R. Jones, M. Schonlau, and W.J. Welch · 1998
Earlier work this paper cites.
Scene Analysis using Latent Dirichlet Allocation
F. Endres, C. Plagemann, C. Stachniss, and W. Burgard · 2009
Earlier work this paper cites.
Information-Theoretic Regret Bounds for Gaussian Process Optimization in the Bandit Setting
N. Srinivas, A. Krause, S.M. Kakade, and M.W. Seeger · 2012
Earlier work this paper cites.
Springer Handbook of Robotics, 2nd edition
C. Stachniss, J. Leonard, and S. Thrun · 2016
Earlier work this paper cites.
Matterport3d: Learning from rgb-d data in indoor environments
A. Chang, A. Dai, T. Funkhouser, M. Halber, M. Niessner, M. Savva, S. Song, A. Zeng, and Y. Zhang · 2017
Earlier work this paper cites.
Habitat: A Platform for Embodied AI Research
M. Savva, A. Kadian, O. Maksymets, Y. Zhao, E. Wijmans, B. Jain, J. Straub, J. Liu, V. Koltun, J. Malik, D. Parikh, and D. Batra · 2019
Earlier work this paper cites.
ObjectNav Revisited: On Evaluation of Embodied Agents Navigating to Objects
D. Batra, A. Gokaslan, A. Kembhavi, O. Maksymets, R. Mottaghi, M. Savva, A. Toshev, and E. Wijmans · 2020
Earlier work this paper cites.
Object Goal Navigation using Goal-Oriented Semantic Exploration
D.S. Chaplot, D. Gandhi, A. Gupta, and R. Salakhutdinov · 2020
Earlier work this paper cites.
DD-PPO: Learning Near-Perfect PointGoal Navigators from 2.5 Billion Frames
E. Wijmans, A. Kadian, A. Morcos, S. Lee, I. Essa, D. Parikh, M. Savva, and D. Batra · 2020
Earlier work this paper cites.
The power of scale for parameter-efficient prompt tuning
B. Lester, R. Al-Rfou, and N. Constant · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision
A. Radford, J.W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, G. Krueger, and I. Sutskever · 2021
Cited alongside, same era.
Habitat-Matterport 3D Dataset (HM3D): 1000 Large-scale 3D Environments for Embodied AI
S.K. Ramakrishnan, A. Gokaslan, E. Wijmans, O. Maksymets, A. Clegg, J.M. Turner, E. Undersander, W. Galuba, A. Westbury, A.X. Chang, M. Savva, Y. Zhao, and D. Batra · 2021
Cited alongside, same era.
Auxiliary Tasks and Exploration Enable ObjectGoal Navigation
J. Ye, D. Batra, A. Das, and E. Wijmans · 2021
Cited alongside, same era.
Hydra: A real-time spatial perception system for 3D scene graph construction and optimization
N. Hughes, Y. Chang, and L. Carlone · 2022
Cited alongside, same era.
Learning-Augmented Model-Based Planning for Visual Exploration
Y. Li, A. Debnath, G.J. Stein, and J. Kosecka · 2023
Later among the works it cites.
Grounding dino: Marrying dino with grounded pre-training for open-set object detection
S. Liu, Z. Zeng, T. Ren, F. Li, H. Zhang, J. Yang, C. Li, J. Yang, H. Su, J. Zhu, et al · 2023
Later among the works it cites.
YOLOv7: Trainable Bag-of-Freebies Sets New State-of-the-Art for Real-Time Object Detectors
C.Y. Wang, A. Bochkovskiy, and H.Y.M. Liao · 2023
Later among the works it cites.
Offline visual representation learning for embodied navigation
K. Yadav, R. Ramrakhya, A. Majumdar, V.P. Berges, S. Kuhar, D. Batra, A. Baevski, and O. Maksymets · 2023
Later among the works it cites.
L3MVN: Leveraging Large Language Models for Visual Target Navigation
B. Yu, H. Kasaei, and M. Cao · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Stubborn: A Strong Baseline for Indoor Object Navigation
H. Luo, A. Yue, Z.W. Hong, and P. Agrawal · 2022
Cited alongside, same era.
Zson: Zero-shot object-goal navigation using multimodal goal embeddings
A. Majumdar, G. Aggarwal, B. Devnani, J. Hoffman, and D. Batra · 2022
Cited alongside, same era.
How To Not Train Your Dragon: Training-free Embodied Object Goal Navigation with Semantic Frontiers
J. Chen, G. Li, S. Kumar, B. Ghanem, and F. Yu · 2023
Cited alongside, same era.
CoWs on Pasture: Baselines and Benchmarks for Language-Driven Zero-Shot Object Navigation
S.Y. Gadre, M. Wortsman, G. Ilharco, L. Schmidt, and S. Song · 2023
Cited alongside, same era.
Visual Language Maps for Robot Navigation
C. Huang, O. Mees, A. Zeng, and W. Burgard · 2023
Cited alongside, same era.
BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models
J. Li, D. Li, S. Savarese, and S. Hoi · 2023
Cited alongside, same era.
C. Zhang, D. Han, Y. Qiao, J.U. Kim, S.H. Bae, S. Lee, and C.S. Hong · 2023
Later among the works it cites.
ESC: exploration with soft commonsense constraints for zero-shot object navigation
K. Zhou, K. Zheng, C. Pryor, Y. Shen, H. Jin, L. Getoor, and X.E. Wang · 2023
Later among the works it cites.
J. Achiam, S. Adler, et al · 2024
Later among the works it cites.
Hierarchical Open-Vocabulary 3D Scene Graphs for Language-Grounded Robot Navigation
A. Werby, C. Huang, M. Büchner, A. Valada, and W. Burgard · 2024
Later among the works it cites.
VLFM: Vision-Language Frontier Maps for Zero-Shot Semantic Navigation
N. Yokoyama, S. Ha, D. Batra, J. Wang, and B. Bucher · 2024
Later among the works it cites.
TriHelper: Zero-Shot Object Navigation with Dynamic Assistance
L. Zhang, Q. Zhang, H. Wang, E. Xiao, Z. Jiang, H. Chen, and R. Xu · 2024
Later among the works it cites.