Fetching the paper…

Pre-trained Text-to-Image Diffusion Models Are Versatile Representation Learners for Control · Around