Fetching the paper…
Reading the bibliography…
Recent works have shown that Large Language Models (LLMs) can facilitate the grounding of instructions for robotic task planning.
P. Velickovic, G. Cucurull, A. Casanova et al. , “Graph attention networks,” stat , vol. 1050, no. 20, pp. 10–48 550, 2017
2017
Earlier work this paper cites.
G. Sepulveda, J. C. Niebles, and A. Soto, “A deep learning based behavioral approach to indoor autonomous navigation,” in 2018 IEEE international conference on robotics and automation (ICRA) . IEEE, 2018, pp. 4646–4653
2018
Earlier work this paper cites.
I. Armeni, Z.-Y. He, J. Gwak et al. , “3d scene graph: A structure for unified semantics, 3d space, and camera,” in Proceedings of the IEEE/CVF international conference on computer vision , 2019, pp. 5664–5673
2019
Earlier work this paper cites.
2020
Earlier work this paper cites.
J. Wald, H. Dhamo, N. Navab et al. , “Learning 3d semantic scene graphs from 3d indoor reconstructions,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2020, pp. 3961–3970
2020
Earlier work this paper cites.
F. Kenghagho Kenfack, F. Ahmed Siddiky, F. Balint-Benczedi et al. , “Robotvqa — a scene-graph- and deep-learning-based visual question answering system for robot manipulation,” in 2020 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) , 2020, pp. 9667–9674
2020
Earlier work this paper cites.
S. Tellex, N. Gopalan, H. Kress-Gazit et al. , “Robots that use language,” Annual Review of Control, Robotics, and Autonomous Systems , vol. 3, pp. 25–55, 2020
2020
Earlier work this paper cites.
M. Shridhar, J. Thomason, D. Gordon et al. , “Alfred: A benchmark for interpreting grounded instructions for everyday tasks,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2020, pp. 10 740–10 749
2020
Earlier work this paper cites.
S.-C. Wu, J. Wald, K. Tateno et al. , “Scenegraphfusion: Incremental 3d scene graph prediction from rgb-d sequences,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2021, pp. 7515–7525
2021
Earlier work this paper cites.
Y. Zhu, J. Tremblay, S. Birchfield et al. , “Hierarchical planning for long-horizon manipulation with geometric and symbolic scene graphs,” in 2021 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2021, pp. 6541–6548
2021
Earlier work this paper cites.
A. Radford, J. W. Kim, C. Hallacy et al. , “Learning transferable visual models from natural language supervision,” in International conference on machine learning . PMLR, 2021, pp. 8748–8763
2021
Earlier work this paper cites.
M. Shridhar, L. Manuelli, and D. Fox, “Cliport: What and where pathways for robotic manipulation,” in Conference on Robot Learning . PMLR, 2022, pp. 894–906
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
S. Amiri, K. Chandan, and S. Zhang, “Reasoning with scene graphs for robot planning under partial observability,” IEEE Robotics and Automation Letters , vol. 7, no. 2, pp. 5560–5567, 2022
2022
Cited alongside, same era.
M. Han, Z. Zhang, Z. Jiao et al. , “Scene reconstruction with functional objects for robot autonomy,” International Journal of Computer Vision , vol. 130, no. 12, pp. 2940–2961, 2022
2022
Cited alongside, same era.
2022
Cited alongside, same era.
O. Mees, L. Hermann, E. Rosete-Beas et al. , “Calvin: A benchmark for language-conditioned policy learning for long-horizon robot manipulation tasks,” IEEE Robotics and Automation Letters , vol. 7, no. 3, pp. 7327–7334, 2022
2022
Cited alongside, same era.
O. Mees, J. Borja-Diaz, and W. Burgard, “Grounding language with visual affordances over unstructured data,” in 2023 IEEE International Conference on Robotics and Automation (ICRA) . IEEE, 2023, pp. 11 576–11 582
2023
Closest in time.
S.-C. Wu, K. Tateno, N. Navab et al. , “Incremental 3d semantic scene graph prediction from rgb sequences,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 5064–5074
2023
Closest in time.
R. Liu, X. Wang, W. Wang et al. , “Bird’s-eye-view scene graph for vision-language navigation,” in Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , October 2023, pp. 10 968–10 980
2023
Closest in time.
J. Bae, D. Shin, K. Ko et al. , “A Survey on 3D Scene Graphs: Definition, Generation and Application,” in Robot Intelligence Technology and Applications 7 , ser. Lecture Notes in Networks and Systems, J. Jo, H.-L. Choi, M. Helbig, H. Oh, J. Hwangbo, C.-H. Lee, and B. Stantic, Eds. Cham: Springer International Publishing, 2023, pp. 136–147
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Z. Ravichandran, L. Peng, N. Hughes et al. , “Hierarchical Representations and Explicit Memory: Learning Effective Navigation Policies on 3D Scene Graphs using Graph Neural Networks,” in 2022 International Conference on Robotics and Automation (ICRA) , May 2022, pp. 9272–9279
2022
Cited alongside, same era.
Z. Jiao, Y. Niu, Z. Zhang et al. , “Sequential Manipulation Planning on Scene Graph,” in 2022 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) , October 2022, pp. 8203–8210, iSSN: 2153-0866
2022
Cited alongside, same era.
E. Jang, A. Irpan, M. Khansari et al. , “Bc-z: Zero-shot task generalization with robotic imitation learning,” in Conference on Robot Learning . PMLR, 2022, pp. 991–1002
2022
Cited alongside, same era.
2022
Cited alongside, same era.
2023
Cited alongside, same era.
B. Zitkovich, T. Yu, S. Xu et al. , “Rt-2: Vision-language-action models transfer web knowledge to robotic control,” in Conference on Robot Learning . PMLR, 2023, pp. 2165–2183
2023
Cited alongside, same era.
A. Brohan, Y. Chebotar, C. Finn et al. , “Do as i can, not as i say: Grounding language in robotic affordances,” in Conference on Robot Learning . PMLR, 2023, pp. 287–318
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Closest in time.
2023
Closest in time.
V. S. Dorbala, J. F. Mullen Jr, and D. Manocha, “Can an embodied agent find your “cat-shaped mug”? llm-based zero-shot object navigation,” IEEE Robotics and Automation Letters , 2023
2023
Closest in time.
C. Lynch, A. Wahid, J. Tompson et al. , “Interactive language: Talking to robots in real time,” IEEE Robotics and Automation Letters , 2023
2023
Closest in time.
2023
Closest in time.
G. Chalvatzaki, A. Younes, D. Nandha et al. , “Learning to reason over scene graphs: a case study of finetuning gpt-2 into a robot language model for grounded task planning,” Frontiers in Robotics and AI , vol. 10, 2023
2023
Closest in time.
2023
Closest in time.
Y. Mu, Q. Zhang, M. Hu et al. , “Embodiedgpt: Vision-language pre-training via embodied chain of thought,” Advances in Neural Information Processing Systems , vol. 36, 2024
2024
Closest in time.