Fetching the paper…

Towards Training Stronger Video Vision Transformers for EPIC-KITCHENS-100 Action Recognition · Around