Fetching the paper…

Hierarchical Cross-Modal Talking Face Generationwith Dynamic Pixel-Wise Loss · Around