Glow: Generative flow with invertible 1x1 convolutions
Diederik P. Kingma and Prafulla Dhariwal · 2018
Cited alongside, same era.
Deep voice 3: Scaling text-to-speech with convolutional sequence learning
Wei Ping, Kainan Peng, Andrew Gibiansky, Sercan Ömer Arik, Ajay Kannan, Sharan Narang, Jonathan Raiman, and John Miller · 2018
Cited alongside, same era.
Self-attention with relative position representations
Peter Shaw, Jakob Uszkoreit, and Ashish Vaswani · 2018
Cited alongside, same era.
Natural TTS synthesis by conditioning wavenet on mel spectrogram predictions
Jonathan Shen, Ruoming Pang, Ron J. Weiss, Mike Schuster, Navdeep Jaitly, Zongheng Yang, Zhifeng Chen, Yu Zhang, Yuxuan Wang, R. J. Skerry-Ryan, Rif A. Saurous, Yannis Agiomyrgiannakis, and Yonghui Wu · 2018
Cited alongside, same era.
Towards end-to-end prosody transfer for expressive speech synthesis with tacotron
R. J. Skerry-Ryan, Eric Battenberg, Ying Xiao, Yuxuan Wang, Daisy Stanton, Joel Shor, Ron J. Weiss, Rob Clark, and Rif A. Saurous · 2018
Cited alongside, same era.
Tacotron 2
Rafael Valle · 2018
Cited alongside, same era.
Style tokens: Unsupervised style modeling, control and transfer in end-to-end speech synthesis
Yuxuan Wang, Daisy Stanton, Yu Zhang, R. J. Skerry-Ryan, Eric Battenberg, Joel Shor, Ying Xiao, Ye Jia, Fei Ren, and Rif A. Saurous · 2018
Cited alongside, same era.
Neural spline flows
Conor Durkan, Artur Bekasov, Iain Murray, and George Papamakarios · 2019
Cited alongside, same era.
Emerging convolutions for generative normalizing flows
Emiel Hoogeboom, Rianne van den Berg, and Max Welling · 2019
Cited alongside, same era.
Flowavenet: A generative flow for raw audio
Sungwon Kim, Sang-Gil Lee, Jongyoon Song, Jaehyeon Kim, and Sungroh Yoon · 2019
Cited alongside, same era.
Neural speech synthesis with transformer network
Naihan Li, Shujie Liu, Yanqing Liu, Sheng Zhao, and Ming Liu · 2019
Cited alongside, same era.