Fetching the paper…
Reading the bibliography…
An end-to-end speech-to-text translation (ST) takes audio in a source language and outputs the text in a target language.
Explaining Sequence-Level Knowledge Distillation as Data-Augmentation for Neural Machine Translation
Gordon, M. A.; and Duh, K. 2019 · 1912
Earlier work this paper cites.
Distance measures for speech recognition, psychological and instrumental
Mermelstein, P. 1976 · 1976
Earlier work this paper cites.
Jointly Trained Transformers models for Spoken Language Translation
Vydana, H. K.; Karafi’at, M.; Zmolikova, K.; Burget, L.; and Cernocky, H. 2020 · 2004
Earlier work this paper cites.
Enhancing the TED-LIUM corpus with selected data for language modeling and more TED talks
Rousseau, A.; Deléglise, P.; and Esteve, Y. 2014 · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, D. P.; and Ba, J. 2015 · 2015
Earlier work this paper cites.
Listen and translate: A proof of concept for end-to-end speech-to-text translation
Bérard, A.; Pietquin, O.; Servan, C.; and Besacier, L. 2016 · 2016
Earlier work this paper cites.
An attentional model for speech translation without transcription
Duong, L.; Anastasopoulos, A.; Chiang, D.; Bird, S.; and Cohn, T. 2016 · 2016
Earlier work this paper cites.
Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks
Graves, A.; Fernández, S.; Gomez, F.; and Schmidhuber, J. 2006 · 2016
Earlier work this paper cites.
Sequence-Level Knowledge Distillation
Kim, Y.; and Rush, A. M. 2016 · 2016
Earlier work this paper cites.
Exploring neural transducers for end-to-end speech recognition
Battenberg, E.; Chen, J.; Child, R.; Coates, A.; Li, Y. G. Y.; Liu, H.; Satheesh, S.; Sriram, A.; and Zhu, Z. 2017 · 2017
Earlier work this paper cites.
Structured-Based Curriculum Learning for End-to-End English-Japanese Speech Translation
Kano, T.; Sakti, S.; and Nakamura, S. 2017 · 2017
Earlier work this paper cites.
Neural Lattice-to-Sequence Models for Uncertain Inputs
Sperber, M.; Neubig, G.; Niehues, J.; and Waibel, A. 2017 · 2017
Earlier work this paper cites.
Attention is All you Need
Vaswani, A.; Shazeer, N.; Parmar, N.; Uszkoreit, J.; Jones, L.; Gomez, A. N.; Kaiser, Ł.; and Polosukhin, I. 2017 · 2017
Earlier work this paper cites.
Sequence-to-Sequence Models Can Directly Translate Foreign Speech
Weiss, R. J.; Chorowski, J.; Jaitly, N.; Wu, Y.; and Chen, Z. 2017 · 2017
Earlier work this paper cites.
Tied Multitask Learning for Neural Speech Translation
Anastasopoulos, A.; and Chiang, D. 2018 · 2018
Earlier work this paper cites.
End-to-end automatic speech translation of audiobooks
Bérard, A.; Besacier, L.; Kocabiyikoglu, A. C.; and Pietquin, O. 2018 · 2018
Earlier work this paper cites.
Speech-transformer: a no-recurrence sequence-to-sequence model for speech recognition
Dong, L.; Xu, S.; and Xu, B. 2018 · 2018
Earlier work this paper cites.
The iwslt 2018 evaluation campaign
Jan, N.; Cattoni, R.; Sebastian, S.; Cettolo, M.; Turchi, M.; and Federico, M. 2018 · 2018
Earlier work this paper cites.
Augmenting Librispeech with French Translations: A Multimodal Corpus for Direct Speech Translation Evaluation
Kocabiyikoglu, A. C.; Besacier, L.; and Kraif, O. 2018 · 2018
Cited alongside, same era.
The USTC-NEL Speech Translation system at IWSLT 2018
Liu, D.; Liu, J.; Guo, W.; Xiong, S.; Ma, Z.; Song, R.; Wu, C.; and Liu, Q. 2018 · 2018
Cited alongside, same era.
Towards fluent translations from disfluent speech
Salesky, E.; Burger, S.; Niehues, J.; and Waibel, A. 2018 · 2018
Cited alongside, same era.
Semi-supervised disfluency detection
Wang, F.; Chen, W.; Yang, Z.; Dong, Q.; Xu, S.; and Xu, B. 2018 · 2018
Cited alongside, same era.
A comparative study on end-to-end speech to text translation
Bahar, P.; Bieschke, T.; and Ney, H. 2019 · 2019
Cited alongside, same era.
On using specaugment for end-to-end speech translation
Bahar, P.; Zeyer, A.; Schlüter, R.; and Ney, H. 2019 · 2019
End-to-End Speech Translation with Knowledge Distillation
Liu, Y.; Xiong, H.; Zhang, J.; He, Z.; Wu, H.; Wang, H.; and Zong, C. 2019 · 2019
Later among the works it cites.
Speech Model Pre-Training for End-to-End Spoken Language Understanding
Lugosch, L.; Ravanelli, M.; Ignoto, P.; Tomar, V. S.; and Bengio, Y. 2019 · 2019
Later among the works it cites.
SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition
Park, D. S.; Chan, W.; Zhang, Y.; Chiu, C.-C.; Zoph, B.; Cubuk, E. D.; and Le, Q. V. 2019 · 2019
Later among the works it cites.
Harnessing indirect training data for end-to-end automatic speech translation: Tricks of the trade
Pino, J.; Puzon, L.; Gu, J.; Ma, X.; McCarthy, A. D.; and Gopinath, D. 2019 · 2019
Later among the works it cites.
Exploring Phoneme-Level Speech Representations for End-to-End Speech Translation
Salesky, E.; Sperber, M.; and Black, A. W. 2019 · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Pre-training on high-resource speech recognition improves low-resource speech-to-text translation
Bansal, S.; Kamper, H.; Livescu, K.; Lopez, A.; and Goldwater, S. 2019 · 2019
Cited alongside, same era.
Neural Speech Translation using Lattice Transformations and Graph Networks
Beck, D.; Cohn, T.; and Haffari, G. 2019 · 2019
Cited alongside, same era.
Breaking the Data Barrier: Towards Robust Speech Translation via Adversarial Stability Training
Cheng, Q.; Fang, M.; Han, Y.; Huang, J.; and Duan, Y. 2019 · 2019
Cited alongside, same era.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Devlin, J.; Chang, M.-W.; Lee, K.; and Toutanova, K. 2019 · 2019
Cited alongside, same era.
Enhancing transformer for end-to-end speech-to-text translation
Di Gangi, M. A.; Negri, M.; Cattoni, R.; Roberto, D.; and Turchi, M. 2019 · 2019
Cited alongside, same era.
Adapting transformer to end-to-end spoken language translation
Di Gangi, M. A.; Negri, M.; and Turchi, M. 2019 · 2019
Cited alongside, same era.
Sperber, M.; Neubig, G.; Pham, N.-Q.; and Waibel, A. 2019b · 2019
Later among the works it cites.
Towards end-to-end speech-to-text translation with two-pass decoding
Sung, T.-W.; Liu, J.-Y.; Lee, H.-y.; and Lee, L.-s. 2019 · 2019
Later among the works it cites.
Lattice Transformer for Speech Translation
Zhang, P.; Ge, N.; Chen, B.; and Fan, K. 2019 · 2019
Later among the works it cites.
Understanding Knowledge Distillation in Non-autoregressive Machine Translation
Zhou, C.; Gu, J.; and Neubig, G. 2019 · 2019
Later among the works it cites.
Knowledge Distillation: A Survey
Gou, J.; Yu, B.; Maybank, S. J.; and Tao, D. 2020 · 2020
Closest in time.
ESPnet-ST: All-in-One Speech Translation Toolkit
Inaguma, H.; Kiyono, S.; Duh, K.; Karita, S.; Yalta, N.; Hayashi, T.; and Watanabe, S. 2020 · 2020
Closest in time.
Low-Latency Sequence-to-Sequence Speech Recognition and Translation by Partial Hypothesis Selection
Liu, D.; Spanakis, G.; and Niehues, J. 2020 · 2020
Closest in time.
Synchronous speech recognition and speech-to-text translation with interactive decoding
Liu, Y.; Zhang, J.; Xiong, H.; Zhou, L.; He, Z.; Wu, H.; Wang, H.; and Zong, C. 2020 · 2020
Closest in time.
Analyzing ASR pretraining for low-resource speech-to-text translation
Stoian, M. C.; Bansal, S.; and Goldwater, S. 2020 · 2020
Closest in time.
Bridging the gap between pre-training and fine-tuning for end-to-end speech translation
Wang, C.; Wu, Y.; Liu, S.; Yang, Z.; and Zhou, M. 2020a · 2020
Closest in time.
Curriculum Pre-training for End-to-End Speech Translation
Wang, C.; Wu, Y.; Liu, S.; Zhou, M.; and Yang, Z. 2020b · 2020
Closest in time.
Towards making the most of bert in neural machine translation
Yang, J.; Wang, M.; Zhou, H.; Zhao, C.; Zhang, W.; Yu, Y.; and Li, L. 2020 · 2020
Closest in time.