Fetching the paper…
Reading the bibliography…
We describe our development of CSS10, a collection of single speaker speech datasets for ten languages.
“Voxforge,”
2006
Earlier work this paper cites.
“Mecab: Yet another part-of-speech and morphological analyzer,”
2006
Earlier work this paper cites.
J. Yamagishi, T. Nose, H. Zen, Z. H. Ling, T. Toda, K. Tokuda, S. King, and S. Renals, “Robust speaker-adaptive hmm-based text-to-speech synthesis,”
2009
Earlier work this paper cites.
“Pavoque corpus of expressive speech,”
2009
Earlier work this paper cites.
F. P. Ribeiro, D. Florencio, C. Zhang, and M. Seltzer, “Crowdmos: An approach for crowdsourcing mean opinion score studies,” in
2011
Earlier work this paper cites.
A. Stan, O. Watts, Y. Mamiya, M. Giurgiu, R. A. J. Clark, J. Yamagishi, and S. King, “Tundra: A multilingual corpus of found data for tts research created with light supervision,” 08 2013
2013
Earlier work this paper cites.
M. Yao, “python-romkan,”
2013
Earlier work this paper cites.
Sysko, “Tatoeba,” https://tatoeba.org, 03 2013
2013
Earlier work this paper cites.
A. Rousseau, P. Deléglise, and Y. Estève, “Ted-lium: an automatic speech recognition dedicated corpus,”
2014
Earlier work this paper cites.
V. Panayotov, G. Chen, D. Povey, and S. Khudanpur, “Librispeech: an asr corpus based on public domain audio books,”
2015
Earlier work this paper cites.
2016
Earlier work this paper cites.
T. Baumann, A. Kohn, and F. Hennig, “The spoken wikipedia corpus collection,”
2016
Earlier work this paper cites.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
J. Sotelo, S. Mehri, K. Kumar, J. F. Santos, K. Kastner, A. Courville, and Y. Bengio, “Char2wav: End-to-end speech synthesis,” 2017
2017
Cited alongside, same era.
2017
Cited alongside, same era.
R. Sonobe, S. Takamichi, and H. Saruwatari, “Jsut corpus: Free large-scale japanese speech corpus for end-to-end speech synthesis,”
2017
Later among the works it cites.
R. Ochshorn and M. Hawkins, “Gentle,”
2017
Later among the works it cites.
J. Sun, “Jieba,”
2017
Later among the works it cites.
“Blizzard challenge 2018,”
2018
Later among the works it cites.
P. Baljekar, “Speech synthesis from found data,” 2018
2018
Later among the works it cites.
“Librivox,”
2018
Later among the works it cites.
“Audacity,”
2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
2017
Cited alongside, same era.
K. Ito, “The lj speech dataset,”
2017
Cited alongside, same era.
Mozilla, “Common voice,”
2017
Cited alongside, same era.
K. Park, “The world english bible,”
2017
Cited alongside, same era.
P. Denisowski, “Cc-cedict,”
2018
Later among the works it cites.
K. Park and T. Mulc, “A (heavily documented) tensorflow implementation of tacotron: A fully end-to-end text-to-speech synthesis model,”
2018
Later among the works it cites.
K. Park, “A tensorflow implementation of dc-tts: yet another text-to-speech model,” https://github.com/Kyubyong/dc_tts, 2018
2018
Later among the works it cites.
I. Solak, “The m-ailabs speech dataset,”
2019
Closest in time.