Fetching the paper…

Vision-Language Pre-training with Object Contrastive Learning for 3D Scene Understanding · Around