Fetching the paper…

Token Reduction Should Go Beyond Efficiency in Generative Models -- From Vision, Language to Multimodality · Around