Fetching the paper…

Accented Text-to-Speech Synthesis with a Conditional Variational Autoencoder · Around