Fetching the paper…
Reading the bibliography…
Speech recognition systems for irregularly-spelled languages like English normally require hand-written pronunciations.
T. Hain, “Implicit pronunciation modelling in asr,” in
2002
Earlier work this paper cites.
M. Bisani and H. Ney, “Joint-sequence models for grapheme-to-phoneme conversion,”
2008
Earlier work this paper cites.
N. Goel, S. Thomas, M. Agarwal, P. Akyazi, L. Burget, K. Feng, A. Ghoshal, O. Glembek, M. Karafiát, D. Povey
2010
Earlier work this paper cites.
A. Laurent, S. Meignier, T. Merlin, P. Deléglise, and F. Spécinov-Trélazé, “Acoustics-based phonetic transcription method for proper nouns.” in
2010
Earlier work this paper cites.
D. Povey, A. Ghoshal, G. Boulianne, L. Burget, O. Glembek, N. Goel, M. Hannemann, P. Motlíček, Y. Qian, P. Schwarz
2011
Earlier work this paper cites.
R. Rasipuram
2012
Earlier work this paper cites.
D. Povey, M. Hannemann, G. Boulianne, L. Burget, A. Ghoshal, M. Janda, M. Karafiát, S. Kombrink, P. Motlíček, Y. Qian
2012
Earlier work this paper cites.
A. Rousseau, P. Deléglise, and Y. Estève, “Ted-lium: an automatic speech recognition dedicated corpus.” in
2012
Cited alongside, same era.
C.-y. Lee, Y. Zhang, and J. R. Glass, “Joint learning of phonetic units and word pronunciations for asr.” in
2013
Cited alongside, same era.
L. Lu, A. Ghoshal, and S. Renals, “Acoustic data-driven pronunciation lexicon for large vocabulary speech recognition,” in
2013
Cited alongside, same era.
I. McGraw, I. Badr, and J. R. Glass, “Learning lexicons from speech using a pronunciation mixture model,”
2013
Cited alongside, same era.
D. F. Harwath and J. R. Glass, “Speech recognition without a lexicon-bridging the gap between graphemic and phonetic systems.” in
2014
Cited alongside, same era.
C.-y. Lee, T. J. O’Donnell, and J. Glass, “Unsupervised lexicon discovery from acoustic input,”
2015
Later among the works it cites.
V. Panayotov, G. Chen, D. Povey, and S. Khudanpur, “Librispeech: an asr corpus based on public domain audio books,” in
2015
Later among the works it cites.
V. Peddinti, D. Povey, and S. Khudanpur, “A time delay neural network architecture for efficient modeling of long temporal contexts.” in
2015
Later among the works it cites.
G. Chen, D. Povey, and S. Khudanpur, “Acoustic data-driven pronunciation lexicon generation for logographic languages,” in
2016
Later among the works it cites.
S. Tsujioka, S. Sakti, K. Yoshino, G. Neubig, and S. Nakamura, “Unsupervised joint estimation of grapheme-to-phoneme conversion systems and acoustic model adaptation for non-native speech recognition,”
2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
M. J. Gales, K. M. Knill, and A. Ragni, “Unicode-based graphemic systems for limited resource languages,” in
2015
Cited alongside, same era.