Fetching the paper…

Learnings from Scaling Visual Tokenizers for Reconstruction and Generation · Around