Fetching the paper…

Multimodal-driven Talking Face Generation via a Unified Diffusion-based Generator · Around