Fetching the paper…

Audio2Gestures: Generating Diverse Gestures from Speech Audio with Conditional Variational Autoencoders · Around