Fetching the paper…

Leveraging Large (Visual) Language Models for Robot 3D Scene Understanding · Around