Fetching the paper…
Reading the bibliography…
Connectionist Temporal Classification has recently attracted a lot of interest as it offers an elegant approach to building acoustic models (AMs) for speech recognition.
L. Rabiner and B. Juang, “An introduction to hidden markov models,”
1986
Earlier work this paper cites.
S. F. Chen and J. Goodman, “An empirical study of smoothing techniques for language modeling,” in
1996
Earlier work this paper cites.
S. Hochreiter and J. Schmidhuber, “Long short-term memory,”
1997
Earlier work this paper cites.
H. Soltau, F. Metze, C. Fugen, and A. Waibel, “A one-pass decoder based on polymorphic linguistic context assignment,” in
2001
Earlier work this paper cites.
A. Stolcke
2002
Earlier work this paper cites.
A. Graves, S. Fernández, F. Gomez, and J. Schmidhuber, “Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,” in
2006
Earlier work this paper cites.
D. Povey, A. Ghoshal, G. Boulianne, L. Burget, O. Glembek, N. Goel, M. Hannemann, P. Motlicek, Y. Qian, P. Schwarz
2011
Earlier work this paper cites.
2012
Earlier work this paper cites.
M. D. Zeiler, “Adadelta: an adaptive learning rate method,”
2012
Earlier work this paper cites.
2014
Earlier work this paper cites.
A. Graves and N. Jaitly, “Towards end-to-end speech recognition with recurrent neural networks.” in
2014
Cited alongside, same era.
2014
Cited alongside, same era.
D. Kingma and J. Ba, “Adam: A method for stochastic optimization,”
2014
Cited alongside, same era.
W. Chan, N. Jaitly, Q. V. Le, and O. Vinyals, “Listen, attend and spell,”
2015
Cited alongside, same era.
Y. Miao, M. Gowayyed, and F. Metze, “Eesen: End-to-end speech recognition using deep rnn models and wfst-based decoding,” in
2015
Cited alongside, same era.
2016
Later among the works it cites.
C. Mendis, J. Droppo, S. Maleki, M. Musuvathi, T. Mytkowicz, and G. Zweig, “Parallelizing wfst speech decoders,” in
2016
Later among the works it cites.
2016
Later among the works it cites.
B. Krause, L. Lu, I. Murray, and S. Renals, “Multiplicative lstm for sequence modelling,”
2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
H. Sak, A. Senior, K. Rao, O. Irsoy, A. Graves, F. Beaufays, and J. Schalkwyk, “Learning acoustic frame labeling for speech recognition with recurrent neural networks,” in
2015
Cited alongside, same era.
A. L. Maas, Z. Xie, D. Jurafsky, and A. Y. Ng, “Lexicon-free conversational speech recognition with networks.” in
2015
Cited alongside, same era.
D. Bahdanau, J. Chorowski, D. Serdyuk, P. Brakel, and Y. Bengio, “End-to-end attention-based large vocabulary speech recognition,” in
2016
Cited alongside, same era.
K. Hwang and W. Sung, “Character-level incremental speech recognition with recurrent neural networks,” in
2016
Cited alongside, same era.
G. Zweig, C. Yu, J. Droppo, and A. Stolcke, “Advances in all-neural speech recognition,”
2016
Cited alongside, same era.
2016
Later among the works it cites.
Y. Miao, M. Gowayyed, X. Na, T. Ko, F. Metze, and A. Waibel, “An empirical exploration of ctc acoustic models,” in
2016
Later among the works it cites.
2016
Later among the works it cites.
L. Lu, X. Zhang, and S. Renais, “On training the recurrent neural network encoder-decoder for large vocabulary end-to-end speech recognition,” in
2016
Later among the works it cites.
2017
Closest in time.
R. Sennrich, O. Firat, K. Cho, A. Birch, B. Haddow, J. Hitschler, M. Junczys-Dowmunt, S. L”aubli, A. V. Miceli Barone, J. Mokry, and M. Nadejde, “Nematus: a Toolkit for Neural Machine Translation,” in
2017
Closest in time.