Fetching the paper…
Reading the bibliography…
In daily domestic settings, frequently used objects like cups often have unfixed positions and multiple instances within the same category, and their carriers frequently change as well.
F. Garcia , et al. , “Markov decision processes,” Markov Decision Processes in Artificial Intelligence , pp. 1–38, 2013
2013
Earlier work this paper cites.
2018
Earlier work this paper cites.
F. Xia , et al. , “Gibson env: Real-world perception for embodied agents,” in Proceedings of the IEEE conference on computer vision and pattern recognition , pp. 9068–9079, 2018
2018
Earlier work this paper cites.
2019
Earlier work this paper cites.
M. Savva , et al. , “Habitat: A platform for embodied ai research,” in Proceedings of the IEEE/CVF international conference on computer vision , pp. 9339–9347, 2019
2019
Earlier work this paper cites.
M. Chang , et al. , “Semantic visual navigation by watching youtube videos,” Advances in Neural Information Processing Systems , vol. 33, pp. 4283–4294, 2020
2020
Earlier work this paper cites.
D. S. Chaplot , et al. , “Object goal navigation using goal-oriented semantic exploration,” Advances in Neural Information Processing Systems , vol. 33, pp. 4247–4258, 2020
2020
Earlier work this paper cites.
2021
Earlier work this paper cites.
A. Radford , et al. , “Learning transferable visual models from natural language supervision,” in International conference on machine learning , pp. 8748–8763. PMLR, 2021
2021
Earlier work this paper cites.
H. Luo , et al. , “Stubborn: A strong baseline for indoor object navigation,” in 2022 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) , pp. 3287–3293. IEEE, 2022
2022
Earlier work this paper cites.
S. K. Ramakrishnan , et al. , “Poni: Potential functions for objectgoal navigation with interaction-free learning,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pp. 18 890–18 900, 2022
2022
Earlier work this paper cites.
Y. Deng , et al. , “S-mki: Incremental dense semantic occupancy reconstruction through multi-entropy kernel inference,” in 2022 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) , pp. 3824–3829. IEEE, 2022
2022
Cited alongside, same era.
2022
Cited alongside, same era.
N. Hughes , et al. , “Hydra: A real-time spatial perception system for 3D scene graph construction and optimization,” 2022
2022
Cited alongside, same era.
L. Qi , et al. , “High-quality entity segmentation,” arXiv preprint arXiv:2211.05776 , 2022
2022
Cited alongside, same era.
T. Pan , et al. , “Tokenize anything via prompting,” arXiv preprint arXiv:2312.09128 , 2023
2023
Later among the works it cites.
A. Rajvanshi , et al. , “Saynav: Grounding large language models for dynamic planning to navigation in new environments,” in Proceedings of the International Conference on Automated Planning and Scheduling , vol. 34, pp. 464–474, 2024
2024
Later among the works it cites.
J. Zhang , et al. , “Vision-language models for vision tasks: A survey,” IEEE Transactions on Pattern Analysis and Machine Intelligence , 2024
2024
Later among the works it cites.
Y. Chang , et al. , “A survey on evaluation of large language models,” ACM Transactions on Intelligent Systems and Technology , vol. 15, no. 3, pp. 1–45, 2024
2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Zhang , et al. , “3d-aware object goal navigation via simultaneous exploration and identification,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pp. 6672–6682, 2023
2023
Cited alongside, same era.
S. Peng , et al. , “Openscene: 3d scene understanding with open vocabularies,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pp. 815–824, 2023
2023
Cited alongside, same era.
2023
Cited alongside, same era.
K. Zhou , et al. , “Esc: Exploration with soft commonsense constraints for zero-shot object navigation,” in International Conference on Machine Learning , pp. 42 829–42 842. PMLR, 2023
2023
Cited alongside, same era.
V. S. Dorbala , et al. , “Can an embodied agent find your ”cat-shaped mug”? llm-based zero-shot object navigation,” IEEE Robotics and Automation Letters , 2023
2023
Cited alongside, same era.
C. Huang , et al. , “Visual language maps for robot navigation,” in 2023 IEEE International Conference on Robotics and Automation (ICRA) , pp. 10 608–10 615. IEEE, 2023
2023
Cited alongside, same era.
N. Yokoyama , et al. , “Vlfm: Vision-language frontier maps for zero-shot semantic navigation,” in 2024 IEEE International Conference on Robotics and Automation (ICRA) , pp. 42–48. IEEE, 2024
2024
Later among the works it cites.
Q. Gu , et al. , “Conceptgraphs: Open-vocabulary 3d scene graphs for perception and planning,” in 2024 IEEE International Conference on Robotics and Automation (ICRA) , pp. 5021–5028. IEEE, 2024
2024
Later among the works it cites.
A. Werby , et al. , “Hierarchical open-vocabulary 3d scene graphs for language-grounded robot navigation,” Robotics: Science and Systems , 2024
2024
Later among the works it cites.
Y. Kuang , et al. , “Openfmnav: Towards open-set zero-shot object navigation via vision-language foundation models,” in Findings of the Association for Computational Linguistics: NAACL 2024 , pp. 338–351, 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
M. Chang , et al. , “Goat: Go to any thing,” in Proceedings of Robotics: Science and Systems (RSS) , 2024
2024
Later among the works it cites.