Fetching the paper…

HumanVLA: Towards Vision-Language Directed Object Rearrangement by Physical Humanoid · Around