Fetching the paper…

GridCLIP: One-Stage Object Detection by Grid-Level CLIP Representation Learning · Around