Fetching the paper…

Learning Visual Representation from Modality-Shared Contrastive Language-Image Pre-training · Around