Fetching the paper…
Reading the bibliography…
Integrating an external language model into a sequence-to-sequence speech recognition system is non-trivial.
J. Li and M. Sun, “Scalable term selection for text categorization,” in
2007
Earlier work this paper cites.
R. C. Moore and W. Lewis, “Intelligent selection of language model training data,” in
2010
Earlier work this paper cites.
A. Rousseau, “Xenc: An open-source tool for data selection in natural language processing,”
2013
Earlier work this paper cites.
C. Gulcehre, O. Firat, K. Xu, K. Cho, L. Barrault, H. Lin, F. Bougares, H. Schwenk, and Y. Bengio, “On using monolingual corpora in neural machine translation,”
2015
Earlier work this paper cites.
G. Hinton, O. Vinyals, and J. Dean, “Distilling the knowledge in a neural network,”
2015
Earlier work this paper cites.
J. K. Chorowski, D. Bahdanau, D. Serdyuk, K. Cho, and Y. Bengio, “Attention-based models for speech recognition,” in
2015
Earlier work this paper cites.
D. Bahdanau, J. Chorowski, D. Serdyuk, P. Brakel, and Y. Bengio, “End-to-end attention-based large vocabulary speech recognition,”
2016
Earlier work this paper cites.
W. Chan, N. Jaitly, Q. Le, and O. Vinyals, “Listen, attend and spell: A neural network for large vocabulary conversational speech recognition,” in
2016
Earlier work this paper cites.
Y. Kim and A. M. Rush, “Sequence-level knowledge distillation,”
2016
Cited alongside, same era.
O. Press and L. Wolf, “Using the output embedding to improve language models,”
2016
Cited alongside, same era.
H. Bu, J. Du, X. Na, B. Wu, and H. Zheng, “AIShell-1: An open-source mandarin speech corpus and a speech recognition baseline,” in
2017
Cited alongside, same era.
2017
Cited alongside, same era.
J. Chorowski and N. Jaitly, “Towards better decoding and language model integration in sequence to sequence models,”
2017
Cited alongside, same era.
C.-C. Chiu, T. N. Sainath, Y. Wu, R. Prabhavalkar, P. Nguyen, Z. Chen, A. Kannan, R. J. Weiss, K. Rao, E. Gonina
2018
Later among the works it cites.
A. Kannan, Y. Wu, P. Nguyen, T. N. Sainath, Z. Chen, and R. Prabhavalkar, “An analysis of incorporating an external language model into a sequence-to-sequence model,” pp. 5824–5828, 2018
2018
Later among the works it cites.
A. Sriram, H. Jun, S. Satheesh, and A. Coates, “Cold fusion: Training seq2seq models together with language models.” pp. 387–391, 2018
2018
Later among the works it cites.
Y. Bai, J. Tao, J. Yi, Z. Wen, and C. Fan, “CLMAD: A chinese language model adaptation dataset,” in
2018
Later among the works it cites.
J. Andrés-Ferrer, N. Bodenstab, and P. Vozila, “Efficient language model adaptation with noise contrastive estimation and kullback-leibler regularization,”
2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
L. Dong, S. Xu, and B. Xu, “Speech-transformer: a no-recurrence sequence-to-sequence model for speech recognition,” in
2018
Cited alongside, same era.
Later among the works it cites.
2018
Later among the works it cites.
S. Changhao, C. Wen, G. Wang, D. Su, M. Luo, and D. Yu, “Component fusion: Learning replaceable language model component for end-to-end speech recognition system,” in
2019
Closest in time.