Fetching the paper…
Reading the bibliography…
Text-to-speech models trained on large-scale datasets have demonstrated impressive in-context learning capabilities and naturalness.
G. Fairbanks, Voice and Articulation Drillbook , 1960
1960
Earlier work this paper cites.
D. Honorof, J. McCullough, and B. Somerville, “Comma gets a cure: A diagnostic passage for accent study,” Retrieved February , vol. 20, p. 2007, 2000
2000
Earlier work this paper cites.
2005
Cited alongside, same era.
Cited in the paper.
Cited in the paper.
Cited in the paper.
[Online]. Available: https://commoncrawl.org/
Cited in the paper.
Cited in the paper.
Cited in the paper.
Cited in the paper.
M. Le, A. Vyas, B. Shi, B. Karrer, L. Sari, R. Moritz, M. Williamson, V. M. Y. Adi, J. Mahadeokar, and W.-N. Hsu, “Voicebox: Text-Guided Multilingual Universal Speech Generation at Scale.”
Cited in the paper.
A. Team, A. Vyas, B. Shi, M. Le, A. Tjandra, Y.-C. Wu, B. Guo, J. Zhang, X. Zhang, R. Adkins, W. Ngan, J. Wang, I. Cruz, B. Akula, A. Akinyemi, B. Ellis, R. Moritz, Y. Yungster, A. Rakotoarison, L. Tan, C. Summers, C. Wood, J. Lane, M. Williamson, and W.-N. Hsu, “Audiobox: Unified Audio Generation with Natural Language Prompts.”
Cited in the paper.
Cited in the paper.
Cited in the paper.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…