Fetching the paper…
Reading the bibliography…
Lingvo is a Tensorflow framework offering a complete solution for collaborative deep learning research, with a particular focus towards sequence-to-sequence models.
Japanese and Korean Voice Search
M. Schuster and K. Nakajima · 2012
Earlier work this paper cites.
Neural Machine Translation of Rare Words with Subword Units, 2015
R. Sennrich, B. Haddow, and A. Birch · 2015
Earlier work this paper cites.
Google’s Neural Machine Translation system: Bridging the gap between human and machine translation
Y. Wu, M. Schuster, Z. Chen, Q. V. Le, M. Norouzi, W. Macherey, M. Krikun, Y. Cao, Q. Gao, K. Macherey, J. Klingner, A. Shah, M. Johnson, X. Liu, L. Kaiser, S. Gouws, Y. Kato, T. Kudo, H. Kazawa, K. Stevens, G. Kurian, N. Patil, W. Wang, C. Young, J. Smith, J. Riesa, A. Rudnick, O. Vinyals, G. Corrado, M. Hughes, and J. Dean · 2016
Earlier work this paper cites.
Sequence-to-sequence models can directly translate foreign speech
Ron J. Weiss, Jan Chorowski, Navdeep Jaitly, Yonghui Wu, and Zhifeng Chen · 2017
Earlier work this paper cites.
Sequence-to-Sequence Models Can Directly Translate Foreign Speech
R. J. Weiss, J. Chorowski J, N. Jaitly, Y. Wu, and Z. Chen · 2017
Earlier work this paper cites.
The Best of Both Worlds: Combining Recent Advances in Neural Machine Translation
M. X. Chen, O. Firat, A.Bapna, M. Johnson, W. Macherey, G. Fosterand L. Jones, M. Schuster, N. Shazeer, N. Parmar, A. Vaswani, J. Uszkoreit, L. Kaiser, Z. Chen, Y. Wu, and M. Hughes · 2018
Earlier work this paper cites.
Monotonic Chunkwise Attention
C. C. Chiu and C. Raffel · 2018
Earlier work this paper cites.
State-of-the-art Speech Recognition With Sequence-to-Sequence Models
C. C. Chiu, T. N. Sainath, Y. Wu, R. Prabhavalkar, P. Nguyen, Z. Chen, A. Kannan, R. J. Weiss, K. Rao, E. Gonina, N. Jaitly, B. Li, J. Chorowski, and M. Bacchiani · 2018
Earlier work this paper cites.
Speech recognition for medical conversations
C. C. Chiu, A. Tripathi, K. Chou, C. Co, N. Jaitly, D. Jaunzeikare, A. Kannan, P. Nguyen, H. Sak, A. Sankar, J. Tansuwan, N. Wan, Y. Wu, and X. Zhang · 2018
Cited alongside, same era.
Hierarchical generative modeling for controllable speech synthesis
W. N. Hsu, Y. Zhang, R. J. Weiss, H. Zen, Y. Wu, Y. Wang, Y. Cao, Y. Jia, Z. Chen, J. Shen, et al · 2018
Cited alongside, same era.
Transfer Learning from Speaker Verification to Multispeaker Text-To-Speech Synthesis
Y. Jia, Y. Zhang, R. J. Weiss, Q. Wang, J. Shen, F. Ren, Z. Chen, P. Nguyen, R. Pang, I. Lopez-Moreno, and Y. Wu · 2018
Cited alongside, same era.
Leveraging weakly supervised data to improve end-to-end speech-to-text translation
Ye Jia, Melvin Johnson, Wolfgang Macherey, Ron J Weiss, Yuan Cao, Chung-Cheng Chiu, Naveen Ari, Stella Laurenzo, and Yonghui Wu · 2018
Cited alongside, same era.
An analysis of incorporating an external language model into a sequence-to-sequence model
Compression of End-to-End Models
R. Pang, T. N. Sainath, R. Prabhavalkar, S. Gupta, Y. Wu, S. Zhang, and C. C. Chiu · 2018
Later among the works it cites.
Minimum Word Error Rate Training for Attention-based Sequence-to-sequence Models
R. Prabhavalkar, T. N. Sainath, Y. Wu, P. Nguyen, Z. Chen, C. C. Chiu, and A. Kannan · 2018
Later among the works it cites.
Improving the Performance of Online Neural Transducer Models
T. N. Sainath, C. C. Chiu, R. Prabhavalkar, A. Kannan, Y. Wu, P. Nguyen, and Z. Chen Z · 2018
Later among the works it cites.
No Need for a Lexicon? Evaluating the Value of the Pronunciation Lexica in End-to-End Models
T. N. Sainath, P. Prabhavalkar, S. Kumar, S. Lee, A. Kannan, D. Rybach, V. Schogol, P. Nguyen, B. Li, Y. Wu, Z. Chen, and C. C. Chiu · 2018
Later among the works it cites.
Natural TTS Synthesis By Conditioning WaveNet on Mel Spectrogram Predictions
J. Shen, R. Pang, R. J. Weiss, M. Schuster, N. Jaitly, Z. Yang, Z. Chen, Y. Zhang, Y. Wang, R. J. Skerry-Ryan, R. A. Saurous, Y. Agiomyrgiannakis, and Y. Wu · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A. Kannan, Y. Wu, P. Nguyen, T. N. Sainath, Z. Chen, and R. Prabhavalkar · 2018
Cited alongside, same era.
Learning hard alignments with variational inference
D. Lawson, C. C. Chiu, G. Tucker, C. Raffel, K. Swersky, and N. Jaitly · 2018
Cited alongside, same era.
Multi-Dialect Speech Recognition With a Single Sequence-to-Sequence Model
B. Li, T. N. Sainath, K. Sim, M. Bacchiani, E. Weinstein, P. Nguyen, Z. Chen, Y. Wu, and K. Rao · 2018
Cited alongside, same era.
https://colab.research.google.com/github/tensorflow/lingvo/blob/master/codelabs/introduction.ipynb
Introduction to Lingvo
Cited in the paper.
End-to-End Multilingual Speech Recognition using Encoder-Decoder Models
S. Toshniwal, T. N. Sainath, R. J. Weiss, B. Li, P. Moreno, E. Weinstein, and K. Rao · 2018
Later among the works it cites.
Contextual Speech Recognition in End-to-End Neural Network Systems using Beam Search
I. Williams, A. Kannan, P. Aleksic, D. Rybach, and T. N. Sainath TN · 2018
Later among the works it cites.