Fetching the paper…
Reading the bibliography…
Thanks to improvements in machine learning techniques including deep learning, a free large-scale speech corpus that can be shared between academic institutions and commercial companies has an important role.
“ATR technical report,”
M. Abe, Y. Sagisaka, T. Umeda, and H. Kuwabara, · 1990
Earlier work this paper cites.
“Compilation of a multilingual parallel corpus,”
Y. Tanaka, · 2001
Earlier work this paper cites.
“English-japanese translation alignment data,” http://www2.nict.go.jp/astrec-att/member/mutiyama/align/index.html
M. Utiyama and M. Takahashi, · 2003
Earlier work this paper cites.
“Applying conditional random fields to Japanese morphological analysis,”
T. Kudo, K. Yamamoto, and Y. Matsumoto, · 2004
Earlier work this paper cites.
“XIMERA: a new TTS from ATR based on corpus-based technologies,”
H. Kawai, T. Toda, J. Ni, M. Tsuzaki, and K. Tokuda., · 2004
Earlier work this paper cites.
“Where does loanword prosody come from?: A case study of Japanese loanword accent,”
H. Kubozono, · 2006
Earlier work this paper cites.
“List of daily-use kanjis http://www.bunka.go.jp/kokugo_nihongo/sisaku/joho/joho/kijun/naikaku/kanji/index.html
Governments of Japan Agency for Cultural Affairs, · 2010
Earlier work this paper cites.
“SNOW E4: evaluation data set of japanese lexical simplification,” http://www.jnlp.org/SNOW/E4
2010
Earlier work this paper cites.
“Deep neural networks for acoustic modeling in speech recognition: The shared views of four research groups,”
G. Hinton, L. Deng, D. Yu, G. Dahl, A. r. Mohamed, N. Jaitly, A. Senior, V. Vanhoucke, P. Nguyen, T. Sainath, and B. Kingsbury, · 2012
Cited alongside, same era.
“Unsupervised learning for text-to-speech synthesis,”
O. Watts, · 2012
Cited alongside, same era.
“Automatic easy Japanese translation for information accessibility of foreigners,”
M. Moku, K. Yamamoto, and A. Makabi, · 2012
Cited alongside, same era.
“Narrow adaptive regularization of weights for grapheme-to-phoneme conversion,”
K. Kubo, S. Sakti, G. Neubig, T. Toda, and S. Nakamura, · 2014
Cited alongside, same era.
“Evaluation dataset and system for japanese lexical simplification,”
K. Tomoyuki and Y. Kazuhide, · 2015
Cited alongside, same era.
“Neologism dictionary based on the language resources on the web for Mecab,” 2015
“WORLD: a vocoder-based high-quality speech synthesis system for real-time applications,”
M. Morise, F. Yokomori, and K. Ozawa, · 2016
Later among the works it cites.
“Sampling-based speech parameter generation using moment-matching network,”
S. Takamichi, K. Tomoki, and H. Saruwatari, · 2017
Closest in time.
“Training algorithm to deceive anti-spoofing verification for DNN-based speech synthesis,”
Y. Saito, S. Takamichi, and H. Saruwatari, · 2017
Closest in time.
“Tacotron: Towards end-to-end speech synthesis,”
Y. Wang, RJ Skerry-Ryan, D. Stanton, Y. Wu, Ron J. Weiss, N. Jaitly, Z. Yang, Y. Xiao, Z. Chen, S. Bengio, Q. Le, Y. Agiomyrgiannakis, R. Clark, and R. A. Saurous, · 2017
Closest in time.
“Char2Wav: End-to-end speech synthesis,”
S. Jose, M. Soroush, K. Kundan, S. João F., K. Kyle, C. Aaron, and B. Yoshua, · 2017
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
T. Sato, · 2015
Cited alongside, same era.
“WaveNet: A generative model for raw audio,”
A. v. d. Oord, S. Dieleman, H. Zen, K. Simonyan, O. Vinyals, A. Graves, N. Kalchbrenner, A. W. Senior, and K. Kavukcuoglu, · 2016
Cited alongside, same era.
“First step towards end-toend parametric TTS synthesis: Generating spectral parameters with neural attention,”
W. Wang, S. Xu, and B. Xu, · 2016
Cited alongside, same era.
“JSUT: Japanese speech corpus of Saruwatari Lab, the University of Tokyo corpus,” https://sites.google.com/site/shinnosuketakamichi/publication/jsut
Cited in the paper.
“Voice-actress corpus,” http://voice-statistics.github.io/
y_benjo and MagnesiumRibbon,
Cited in the paper.
“Wikipedia,” https://ja.wikipedia.org/
Cited in the paper.
“COURTS IN JAPAN,” http://www.courts.go.jp/app/hanrei_jp/search1
Cited in the paper.
“Implementation of a word segmentation dictionary called mecab-ipadic-neologd and study on how to use it effectively for information retrieval (in Japanese),”
T. Sato, T. Hashimoro, and M. Okumura, · 2017
Closest in time.
“Voice conversion using sequence-to-sequence learning of context posterior probabilities,”
H. Miyoshi, Y. Saito, S. Takamichi, and H. Saruwatari, · 2017
Closest in time.