Fetching the paper…

Diff-Foley: Synchronized Video-to-Audio Synthesis with Latent Diffusion Models · Around