Fetching the paper…

MOAT: Alternating Mobile Convolution and Attention Brings Strong Vision Models · Around