Fetching the paper…
Reading the bibliography…
Recently, end-to-end speech translation (ST) has gained significant attention as it avoids error propagation.
“Assessing the impact of speech recognition errors on machine translation quality,”
Nicholas Ruiz and Marcello Federico, · 2014
Earlier work this paper cites.
“Google’s multilingual neural machine translation system: Enabling zero-shot translation,”
Melvin Johnson, Mike Schuster, Quoc V Le, Maxim Krikun, Yonghui Wu, Zhifeng Chen, Nikhil Thorat, Fernanda Viégas, Martin Wattenberg, Greg Corrado, et al., · 2017
Earlier work this paper cites.
“Attention is all you need,”
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, L ukasz Kaiser, and Illia Polosukhin, · 2017
Earlier work this paper cites.
“Svcca: Singular vector canonical correlation analysis for deep learning dynamics and interpretability,”
Maithra Raghu, Justin Gilmer, Jason Yosinski, and Jascha Sohl-Dickstein, · 2017
Earlier work this paper cites.
“Low-resource speech-to-text translation,”
Sameer Bansal, Herman Kamper, Karen Livescu, Adam Lopez, and Sharon Goldwater, · 2018
Earlier work this paper cites.
“SentencePiece: A simple and language independent subword tokenizer and detokenizer for neural text processing,”
Taku Kudo and John Richardson, · 2018
Earlier work this paper cites.
“A call for clarity in reporting BLEU scores,”
Matt Post, · 2018
Cited alongside, same era.
“Pre-training on high-resource speech recognition improves low-resource speech-to-text translation,”
Sameer Bansal, Herman Kamper, Karen Livescu, Adam Lopez, and Sharon Goldwater, · 2019
Cited alongside, same era.
“Leveraging weakly supervised data to improve end-to-end speech-to-text translation,”
Ye Jia, Melvin Johnson, Wolfgang Macherey, Ron J Weiss, Yuan Cao, Chung-Cheng Chiu, Naveen Ari, Stella Laurenzo, and Yonghui Wu, · 2019
Cited alongside, same era.
“Improving zero-shot translation with language-independent constraints,”
Ngoc-Quan Pham, Jan Niehues, Thanh-Le Ha, and Alexander Waibel, · 2019
Cited alongside, same era.
“Very deep self-attention networks for end-to-end speech recognition,”
Ngoc-Quan Pham, Thai-Son Nguyen, Jan Niehues, Markus Müller, Sebastian Stüker, and Alexander Waibel, · 2019
Cited alongside, same era.
“Vizseq: A visual analysis toolkit for text generation tasks,”
Changhan Wang, Anirudh Jain, Danlu Chen, and Jiatao Gu, · 2019
Later among the works it cites.
“Speech translation and the end-to-end promise: Taking stock of where we are,”
Matthias Sperber and Matthias Paulik, · 2020
Later among the works it cites.
“Improving zero-shot translation by disentangling positional information,”
Danni Liu, Jan Niehues, James Cross, Francisco Guzmán, and Xian Li, · 2020
Later among the works it cites.
“Covost 2: A massively multilingual speech-to-text translation corpus,”
Changhan Wang, Anne Wu, and Juan Pino, · 2020
Later among the works it cites.
“A general multi-task learning framework to leverage text data for speech to text tasks,”
Yun Tang, Juan Pino, Changhan Wang, Xutai Ma, and Dmitriy Genzel, · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“The iwslt 2019 kit speech translation system,”
Ngoc-Quan Pham, Thai-Son Nguyen, Thanh-Le Ha, Juan Hussain, Felix Schneider, Jan Niehues, Sebastian Stüker, and Alexander Waibel, · 2019
Cited alongside, same era.