Fetching the paper…

Large-Scale Adversarial Training for Vision-and-Language Representation Learning · Around