Fetching the paper…

HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction · Around