Fetching the paper…
Reading the bibliography…
Attention-based sequence-to-sequence models for automatic speech recognition jointly train an acoustic model, language model, and alignment mechanism.
“Long Short-Term Memory,”
S. Hochreiter and J. Schmidhuber, · 1997
Earlier work this paper cites.
“Weighted finite state transducers in speech recognition,”
M. Mohri, F. Pereira, and M. Riley, · 2002
Earlier work this paper cites.
“Bidirectional LSTM Networks for Improved Phoneme Classification and Recognition,”
M. Schuster and K. K. Paliwal, · 2005
Earlier work this paper cites.
“Connectionist Temporal Classification: Labeling Unsegmented Seuqnece Data with Recurrent Neural Networks,”
A. Graves, S. Fernandez, F. Gomez, and J. Schmidhuber, · 2006
Earlier work this paper cites.
“Openfst: A general and efficient weighted finite-state transducer library,”
C. Allauzen, M. Riley, J. Schalkwyk, W. Skut, and M. Mohri, · 2007
Earlier work this paper cites.
“Recurrent neural network based language model,”
T.Mikolov, M. Karafiat, L. Burget, J. Cernocky, and S. Khudanbur, · 2010
Earlier work this paper cites.
“Bayesian language model interpolation for mobile speech input,”
C. Allauzen and M. Riley, · 2011
Earlier work this paper cites.
“The kaldi speech recognition toolkit,”
D. Povey, A. Ghoshal, G. Boulianne, L. Burget, O. Glembek, N. Goel, M. Hannemann, P. Motlicek, Y. Qian, J. Silovsky P. Schwarz, G. Stemmer, , and K. Vesely, · 2011
Earlier work this paper cites.
“Sequence transduction with recurrent neural networks,”
A. Graves, · 2012
Earlier work this paper cites.
“Japanese and Korean voice search,”
M. Schuster and K. Nakajima, · 2012
Cited alongside, same era.
W. Chan, N. Jaitly, Q. V. Le, and O. Vinyals, · 2015
Cited alongside, same era.
“On using monolingual corpora in neural machine translation,”
C. Gulcehre, O. Firat, K. Xu, K. Cho, L. Barrault, H. Lin, F. Bougares, H. Schwenk, and Y.Bengio, · 2015
Cited alongside, same era.
“R. sennrich and b. haddow and a. birch,”
Improving neural machine translation models with monolingual data, · 2015
Cited alongside, same era.
“Fast and Accurate Recurrent Neural Network Acoustic Models for Speech Recognition,”
H. Sak, A. Senior, K. Rao, and F. Beaufays, · 2015
Cited alongside, same era.
Y. Wu, M. Schuster, and et. al., · 2016
Later among the works it cites.
“Lower Frame Rate Neural Network Acoustic Models,”
G. Pundak and T. N. Sainath, · 2016
Later among the works it cites.
“A Comparison of Sequence-to-sequence Models for Speech Recognition,”
R. Prabhavalkar, K. Rao, B. Li, L. Johnson, and N. Jaitly, · 2017
Closest in time.
“Towards Better Decoding and Language Model Integration in Sequence to Sequence Models,”
J. K. Chorowski and N. Jaitly, · 2017
Closest in time.
“Advances in joint ctc-attention based end-to-end speech recognition with a deep cnn encoder and rnn-lm,”
T. Hori, S. Watanabe, Y. Zhang, and W. Chan, · 2017
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“TensorFlow: Large-Scale Machine Learning on Heterogeneous Distributed Systems,” Available online: http://download.tensorflow.org/paper/whitepaper2015.pdf, 2015
M. Abadi et al., · 2015
Cited alongside, same era.
“End-to-End Attention-based Large Vocabulary Speech Recognition,”
D. Bahdanau, J. Chorowski, D. Serdyuk, P. Brakel, and Y. Bengio, · 2016
Cited alongside, same era.
“Exploring the limits of language modeling,”
R. Jozefowicz, O. Vinyals, M. Schuster, N. Shazeer, and Y. Wu, · 2016
Cited alongside, same era.
“Modeling coverage for neural machine translation,”
Z. Tu, Z. Lu, Y. Liu, X. Liu, and H. Li, · 2016
Cited alongside, same era.
A. Sriram, H. Jun, S. Satheesh, and A. Coates, · 2017
Closest in time.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin, · 2017
Closest in time.
K. Rao, R. Prabhavalkar, and H. Sak, · 2017
Closest in time.