Fetching the paper…

SayAnything: Audio-Driven Lip Synchronization with Conditional Video Diffusion · Around