Fetching the paper…
Reading the bibliography…
While end-to-end ASR systems have proven competitive with the conventional hybrid approach, they are prone to accuracy degradation when it comes to noisy and low-resource conditions.
A. Graves, S. Fernández, F. Gomez, and J. Schmidhuber, “Connectionist temporal classification: Labelling unsegmented sequence data with recurrent neural networks,” in
2006
Earlier work this paper cites.
G. Hinton, l. Deng, D. Yu, G. Dahl
2012
Earlier work this paper cites.
A. Graves, “Sequence transduction with recurrent neural networks,” in
2012
Earlier work this paper cites.
V. Panayotov, G. Chen, D. Povey, and S. Khudanpur, “Librispeech: An ASR corpus based on public domain audio books,” in
2015
Earlier work this paper cites.
W. Chan, N. Jaitly, Q. V. Le, and O. Vinyals, “Listen, attend and spell: A neural network for large vocabulary conversational speech recognition,” in
2016
Earlier work this paper cites.
R. Sennrich, B. Haddow, and A. Birch, “Neural machine translation of rare words with subword units,” in
2016
Earlier work this paper cites.
S. Kim, T. Hori, and S. Watanabe, “Joint CTC-attention based end-to-end speech recognition using multi-task learning,” in
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit
2017
Earlier work this paper cites.
A. Zeyer, K. Irie, R. Schlüter, and H. Ney, “Improved training of end-to-end attention models for speech recognition,”
2018
Earlier work this paper cites.
V. Bataev, M. Korenevsky, I. Medennikov, and A. Zatvornitskiy, “Exploring end-to-end techniques for low-resource speech recognition,”
2018
Earlier work this paper cites.
J. Barker, S. Watanabe, E. Vincent, and J. Trmal, “The fifth ’CHiME’ speech separation and recognition challenge: Dataset, task and baselines,” in
2018
Earlier work this paper cites.
J. Du, T. Gao, L. Sun, F. Ma
2018
Earlier work this paper cites.
N. Kanda, R. Ikeshita, S. Horiguchi, Y. Fujita
2018
Cited alongside, same era.
I. Medennikov, I. Sorokin, A. Romanenko, D. Popov
2018
Cited alongside, same era.
S. Dalmia, S. Kim, and F. Metze, “Situation informed end-to-end ASR for noisy environments,” in
2018
Cited alongside, same era.
C. Boeddecker, J. Heitkaemper, J. Schmalenstroeer, L. Drude
2018
Cited alongside, same era.
S. Merity, N. S. Keskar, and R. Socher, “Regularizing and optimizing LSTM language models,” in
2018
Cited alongside, same era.
S. Watanabe, T. Hori, S. Karita, T. Hayashi
2018
Cited alongside, same era.
A. S. Subramanian, X. Wang, S. Watanabe, T. Taniguchi
2019
Later among the works it cites.
X. Chang, W. Zhang, Y. Qian, J. L. Roux, and S. Watanabe, “Mimo-speech: End-to-end multi-channel multi-speaker speech recognition,”
2019
Later among the works it cites.
Z. Tian, J. Yi, J. Tao, Y. Bai, and Z. Wen, “Self-attention transducers for end-to-end speech recognition,” in
2019
Later among the works it cites.
C.-F. Yeh, J. Mahadeokar, K. Kalgaonkar, Y. Wang
2019
Later among the works it cites.
M. Jain, K. Schubert, J. Mahadeokar, C.-F. Yeh
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
C. Lüscher, E. Beck, K. Irie, M. Kitza
2019
Cited alongside, same era.
G. Synnaeve, Q. Xu, J. Kahn, E. Grave
2019
Cited alongside, same era.
J. Li, V. Lavrukhin, B. Ginsburg, R. Leary
2019
Cited alongside, same era.
N. Yalta, S. Watanabe, T. Hori, K. Nakadai, and T. Ogata, “CNN-based multichannel end-to-end speech recognition for everyday home environments,” in
2019
Cited alongside, same era.
C. Zorila, C. Boeddeker, R. Doddipatla, and R. Haeb-Umbach, “An investigation into the effectiveness of enhancement in ASR training and test for CHiME-5 dinner party transcription,” in
2019
Later among the works it cites.
S. Karita, X. Wang, S. Watanabe, T. Yoshimura
2019
Later among the works it cites.
S. Schneider, A. Baevski, R. Collobert, and M. Auli, “Wav2vec: Unsupervised pre-training for speech recognition,” in
2019
Later among the works it cites.
D. S. Park, W. Chan, Y. Zhang, C.-C. Chiu
2019
Later among the works it cites.
I. Medennikov, M. Korenevsky, T. Prisyach, Y. Khokhlov
2020
Closest in time.