Fetching the paper…
Reading the bibliography…
There is no limit to how much a robot might explore and learn, but all of that knowledge needs to be searchable and actionable.
P. H. Sneath and R. R. Sokal, Numerical Taxonomy: The Principles and Practice of Numerical Classification . W.H. Freeman, 1973
1973
Earlier work this paper cites.
H. Yang, L. Chaisorn, Y. Zhao, S.-Y. Neo, and T.-S. Chua, “Videoqa: question answering on news video,” in Proceedings of the eleventh ACM international conference on Multimedia , 2003, pp. 632–641
2003
Earlier work this paper cites.
2011
Earlier work this paper cites.
A. Hornung, K. M. Wurm, M. Bennewitz, C. Stachniss, and W. Burgard, “Octomap: An efficient probabilistic 3d mapping framework based on octrees,” Autonomous robots , vol. 34, pp. 189–206, 2013
2013
Earlier work this paper cites.
Y. Zhu, R. Mottaghi, E. Kolve, J. J. Lim, A. Gupta, L. Fei-Fei, and A. Farhadi, “Target-driven visual navigation in indoor scenes using deep reinforcement learning,” in 2017 ICRA . IEEE, 2017, pp. 3357–3364
2017
Earlier work this paper cites.
L. Zhang, L. Wei, P. Shen, W. Wei, G. Zhu, and J. Song, “Semantic slam based on object detection and improved octomap,” IEEE Access , vol. 6, pp. 75 545–75 559, 2018
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
A. Das, S. Datta, G. Gkioxari, S. Lee, D. Parikh, and D. Batra, “Embodied question answering,” in CVPR , 2018, pp. 1–10
2018
Earlier work this paper cites.
S. Shah, D. Dey, C. Lovett, and A. Kapoor, “Airsim: High-fidelity visual and physical simulation for autonomous vehicles,” in FSR , 2018
2018
Earlier work this paper cites.
K. Liu, Z. Fan, M. Liu, and S. Zhang, “Object-aware semantic mapping of indoor scenes using octomap,” in 2019 Chinese Control Conference (CCC) . IEEE, 2019, pp. 8671–8676
2019
Earlier work this paper cites.
L. Yu, X. Chen, G. Gkioxari, M. Bansal, T. L. Berg, and D. Batra, “Multi-target embodied question answering,” in ICCV , 2019, p. 6309
2019
Earlier work this paper cites.
Chaplot, R. R et al. , “Object goal navigation using goal-oriented semantic exploration,” NeurIPS , vol. 33, 2020
2020
Earlier work this paper cites.
Castro, Rada et al. , “Lifeqa: A real-life dataset for video question answering,” in LREC , 2020
2020
Earlier work this paper cites.
P. Lewis, D. Kiela et al. , “Retrieval-augmented generation for knowledge-intensive nlp tasks,” 2021
2021
Earlier work this paper cites.
S. Y. Min, D. S. Chaplot, P. Ravikumar, Y. Bisk, and R. Salakhutdinov, “Film: Following instructions in language with modular methods,” ICLR , 2021
2021
Earlier work this paper cites.
J. Xiao, X. Shang, A. Yao, and T.-S. Chua, “Next-qa: Next phase of question-answering to explaining temporal actions,” in ICCV , 2021
2021
Cited alongside, same era.
N. Hughes et al. , “Hydra: A real-time spatial perception system for 3d scene graph construction and optimization,” RSS , 2022
2022
Cited alongside, same era.
2022
Cited alongside, same era.
S. Y. Min, Yonatan et al. , “Don’t copy the teacher: Data and model challenges in embodied dialogue,” EMNLP , 2022
2022
Cited alongside, same era.
N. M. M. Shafiullah, C. Paxton, L. Pinto, S. Chintala, and A. Szlam, “Clip-fields: Weakly supervised semantic fields for robotic memory,” arXiv: Arxiv-2210.05663 , 2022
2022
C. Huang, O. Mees, A. Zeng, and W. Burgard, “Visual language maps for robot navigation,” in Proceedings of the ICRA , London, UK, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
S. Y. Min, Y.-H. H. Tsai, W. Ding, A. Farhadi, R. Salakhutdinov, Y. Bisk, and J. Zhang, “Self-supervised object goal navigation with in-situ finetuning,” in 2023 IROS . IEEE, 2023, pp. 7119–7126
2023
Later among the works it cites.
K. Rana et al. , “Sayplan: Grounding large language models using 3d scene graphs for scalable task planning,” in CoRL , 2023
2023
Later among the works it cites.
K. Zheng, A. Paul, and S. Tellex, “Asystem for generalized 3d multi-object search,” in 2023 ICRA . IEEE, 2023, pp. 1638–1644
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
S. K. Ramakrishnan, D. S. Chaplot, Z. Al-Halah, J. Malik, and K. Grauman, “Poni: Potential functions for objectgoal navigation with interaction-free learning,” in ICCV , 2022, pp. 18 890–18 900
2022
Cited alongside, same era.
Li, Fuchun et al. , “Embodied semantic scene graph generation,” in CoRL , A. Faust, D. Hsu, and G. Neumann, Eds. PMLR, 2022
2022
Cited alongside, same era.
L. Mezghan, S. Sukhbaatar, T. Lavril, O. Maksymets, D. Batra, P. Bojanowski, and K. Alahari, “Memory-augmented reinforcement learning for image-goal navigation,” in 2022 IROS . IEEE, 2022, pp. 3316–3323
2022
Cited alongside, same era.
J. Krantz, S. Lee, J. Malik, D. Batra, and D. S. Chaplot, “Instance-specific image goal navigation: Training embodied agents to find object instances,” CVPR , 2022
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2022
Cited alongside, same era.
A. Asai, S. Min, Z. Zhong, and D. Chen, “Acl 2023 tutorial: Retrieval-based language models and applications,” ACL 2023 , 2023
2023
Cited alongside, same era.
2023
Later among the works it cites.
2023
Later among the works it cites.
S. Tan, M. Ge, D. Guo, H. Liu, and F. Sun, “Knowledge-based embodied question answering,” IEEE Transactions on Pattern Analysis and Machine Intelligence , 2023
2023
Later among the works it cites.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
S. Yao, D. Yu, J. Zhao, I. Shafran, T. Griffiths, Y. Cao, and K. Narasimhan, “Tree of thoughts: Deliberate problem solving with large language models,” NeurIPS , vol. 36, 2024
2024
Closest in time.