Fetching the paper…
Reading the bibliography…
Visual target navigation is a critical capability for autonomous robots operating in unknown environments, particularly in human-robot interaction scenarios.
J. A. Sethian, “A fast marching level set method for monotonically advancing fronts.,” Proceedings of the National Academy of Sciences
1996
Earlier work this paper cites.
M. Daszykowski and B. Walczak, “Density-Based Clustering Methods,” in Comprehensive Chemometrics
2009
Earlier work this paper cites.
A. V. Segal, D. Haehnel, and S. Thrun, “Generalized-ICP,” in Robotics: Science and Systems
2010
Earlier work this paper cites.
D. Puig, M. A. Garcia, and L. Wu, “A new global optimization strategy for coordinated multi-robot exploration: Development and comparative evaluation,” Robotics and Autonomous Systems
2011
Earlier work this paper cites.
M. Juliá, A. Gil, and O. Reinoso, “A comparison of path planning strategies for autonomous exploration and mapping of unknown environments,” Autonomous Robots
2012
Earlier work this paper cites.
A. Visser and J. D. Hoog, “Discussion of multi-robot exploration in communication-limited environments,” in 2013 IEEE International Conference on Robotics and Automation
2013
Earlier work this paper cites.
Y. Zhu, R. Mottaghi, E. Kolve, J. J. Lim, A. Gupta, L. Fei-Fei, and A. Farhadi, “Target-driven visual navigation in indoor scenes using deep reinforcement learning,” in Proceedings - IEEE International Conference on Robotics and Automation
2017
Earlier work this paper cites.
F. Xia, A. R. Zamir, Z. He, A. Sax, J. Malik, and S. Savarese, “Gibson Env: Real-World Perception for Embodied Agents,” in 2018 IEEE/CVF Conference on Computer Vision and Pattern Recognition
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
P. Anderson, A. Chang, D. S. Chaplot, A. Dosovitskiy, S. Gupta, V. Koltun, J. Kosecka, J. Malik, R. Mottaghi, M. Savva, and A. R. Zamir, “On Evaluation of Embodied Navigation Agents,” arXiv
2018
Earlier work this paper cites.
M. Savva, A. Kadian, O. Maksymets, Y. Zhao, E. Wijmans, B. Jain, J. Straub, J. Liu, V. Koltun, J. Malik, D. Parikh, and D. Batra, “Habitat: A Platform for Embodied AI Research,” in 2019 IEEE/CVF International Conference on Computer Vision (ICCV)
2019
Earlier work this paper cites.
D. S. Chaplot, D. Gandhi, A. Gupta, and R. Salakhutdinov, “Object goal navigation using goal-oriented semantic exploration,” Advances in Neural Information Processing Systems
2020
Earlier work this paper cites.
R. Druon, Y. Yoshiyasu, A. Kanezaki, and A. Watt, “Visual object search by learning spatial context,” IEEE Robotics and Automation Letters
2020
Earlier work this paper cites.
D. S. Chaplot, R. Salakhutdinov, A. Gupta, and S. Gupta, “Neural topological SLAM for visual navigation,” in Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition
2020
Earlier work this paper cites.
D. S. Chaplot, D. Gandhi, S. Gupta, A. Gupta, and R. Salakhutdinov, “Learning To Explore Using Active Neural Slam,” in 8th International Conference on Learning Representations, ICLR 2020
2020
Earlier work this paper cites.
A. Szot, A. Clegg, E. Undersander, E. Wijmans, Y. Zhao, J. Turner, N. Maestre, M. Mukadam, D. Chaplot, O. Maksymets, A. Gokaslan, V. Vondrus, S. Dharur, F. Meier, W. Galuba, A. Chang, Z. Kira, V. Koltun, J. Malik, M. Savva, and D. Batra, “Habitat 2.0: Training Home Assistants to Rearrange their Habitat,” Advances in Neural Information Processing Systems
2021
Earlier work this paper cites.
S. K. Ramakrishnan, A. Gokaslan, E. Wijmans, O. Maksymets, A. Clegg, J. Turner, E. Undersander, W. Galuba, A. Westbury, A. X. Chang, M. Savva, Y. Zhao, and D. Batra, “Habitat-Matterport 3D Dataset (HM3D): 1000 Large-scale 3D Environments for Embodied AI,” ArXiv
2021
Earlier work this paper cites.
J. Ye, D. Batra, A. Das, and E. Wijmans, “Auxiliary Tasks and Exploration Enable ObjectGoal Navigation,” Proceedings of the IEEE International Conference on Computer Vision
2021
Cited alongside, same era.
O. Maksymets, V. Cartillier, A. Gokaslan, E. Wijmans, W. Galuba, S. Lee, and D. Batra, “THDA: Treasure Hunt Data Augmentation for Semantic Navigation,” Proceedings of the IEEE International Conference on Computer Vision
2021
Cited alongside, same era.
Y. Liang, B. Chen, and S. Song, “SSCNav: Confidence-Aware Semantic Scene Completion for Visual Semantic Navigation,” in Proceedings - IEEE International Conference on Robotics and Automation
2021
Cited alongside, same era.
A. Khandelwal, L. Weihs, R. Mottaghi, and A. Kembhavi, “Simple but Effective: CLIP Embeddings for Embodied AI,” in 2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)
2022
Cited alongside, same era.
C. H. Song, B. M. Sadler, J. Wu, W. L. Chao, C. Washington, and Y. Su, “LLM-Planner: Few-Shot Grounded Planning for Embodied Agents with Large Language Models,” Proceedings of the IEEE International Conference on Computer Vision
2023
Closest in time.
B. Yu, H. Kasaei, and M. Cao, “L3MVN: Leveraging Large Language Models for Visual Target Navigation,” IEEE International Conference on Intelligent Robots and Systems
2023
Closest in time.
C. Yu, X. Yang, J. Gao, J. Chen, Y. Li, J. Liu, Y. Xiang, R. Huang, H. Yang, Y. Wu, and Y. Wang, “Asynchronous Multi-Agent Reinforcement Learning for Efficient Real-Time Multi-Robot Cooperative Exploration,” Proceedings of the International Joint Conference on Autonomous Agents and Multiagent Systems, AAMAS
2023
Closest in time.
A. Shtedritski, C. Rupprecht, and A. Vedaldi, “What does clip know about a red circle? visual prompt engineering for vlms,” in Proceedings of the IEEE/CVF International Conference on Computer Vision
2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
S. K. Ramakrishnan, D. S. Chaplot, Z. Al-Halah, J. Malik, and K. Grauman, “PONI: Potential Functions for ObjectGoal Navigation with Interaction-free Learning,” Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition
2022
Cited alongside, same era.
Y. Lyu, Y. Shi, and X. Zhang, “Improving Target-driven Visual Navigation with Attention on 3D Spatial Relationships,” Neural Processing Letters
2022
Cited alongside, same era.
X. Liu, D. Guo, H. Liu, and F. Sun, “Multi-Agent Embodied Visual Semantic Navigation with Scene Prior Knowledge,” IEEE Robotics and Automation Letters
2022
Cited alongside, same era.
W. Huang, P. Abbeel, D. Pathak, and I. Mordatch, “Language Models as Zero-Shot Planners: Extracting Actionable Knowledge for Embodied Agents,” Proceedings of Machine Learning Research
2022
Cited alongside, same era.
R. Ramrakhya, E. Undersander, D. Batra, and A. Das, “Habitat-Web: Learning Embodied Object-Search Strategies from Human Demonstrations at Scale,” Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition
2022
Cited alongside, same era.
A. Majumdar, G. Aggarwal, B. Devnani, J. Hoffman, and D. Batra, “ZSON: Zero-Shot Object-Goal Navigation using Multimodal Goal Embeddings,” Advances in Neural Information Processing Systems
2022
Cited alongside, same era.
Z. Al-Halah, S. K. Ramakrishnan, and K. Grauman, “Zero Experience Required: Plug & Play Modular Transfer Learning for Semantic Visual Navigation,” Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition
2022
Cited alongside, same era.
K. Ye, S. Dong, Q. Fan, H. Wang, L. Yi, F. Xia, J. Wang, and B. Chen, “Multi-Robot Active Mapping via Neural Bipartite Graph Matching,” Proceedings of the IEEE Computer Society Conference on Computer Vision and Pattern Recognition
2022
Cited alongside, same era.
Closest in time.
G. Jocher, A. Chaurasia, and J. Qiu, “Ultralytics yolov8,” 2023
2023
Closest in time.
2023
Closest in time.
K. Koide, S. Oishi, M. Yokozuka, and A. Banno, “General, single-shot, target-less, and automatic lidar-camera extrinsic calibration toolbox,” in 2023 IEEE International Conference on Robotics and Automation (ICRA)
2023
Closest in time.
Technical Report
OpenAI, “Gpt-4o: Advancing cost-efficient intelligence,” 2024 · 2024
Closest in time.
G. Zhou, Y. Hong, and Q. Wu, “NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models,” Proceedings of the AAAI Conference on Artificial Intelligence
2024
Closest in time.
S. Hong, M. Zhuge, J. Chen, X. Zheng, Y. Cheng, C. Zhang, J. Wang, Z. Wang, S. K. S. Yau, Z. Lin, L. Zhou, C. Ran, L. Xiao, C. Wu, and J. Schmidhuber, “Metagpt: Meta Programming for a Multi-Agent Collaborative Framework,” 12th International Conference on Learning Representations, ICLR 2024
2024
Closest in time.
H. Zhang, W. Du, J. Shan, Q. Zhou, Y. Du, J. B. Tenenbaum, T. Shu, and C. Gan, “Building Cooperative Embodied Agents Modularly With Large Language Models,” 12th International Conference on Learning Representations, ICLR 2024
2024
Closest in time.
Z. Mandi, S. Jain, and S. Song, “RoCo: Dialectic Multi-Robot Collaboration with Large Language Models,” in Proceedings - IEEE International Conference on Robotics and Automation
2024
Closest in time.
K. Fang, F. Liu, P. Abbeel, and S. Levine, “MOKA: Open-World Robotic Manipulation through Mark-Based Visual Prompting,” in Robotics: Science and Systems
2024
Closest in time.
2024
Closest in time.
B. Yu, Y. Liu, L. Han, H. Kasaei, T. Li, and M. Cao, “VLN-Game: Vision-Language Equilibrium Search for Zero-Shot Semantic Navigation,” pp. 1–15, nov 2024
2024
Closest in time.
Technical Report
OpenAI, “Gpt-4o-mini: A compact and cost-effective multimodal model,” 2024 · 2024
Closest in time.