Fetching the paper…

Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning · Around