Fetching the paper…
Reading the bibliography…
End-to-End Speech Translation (E2E-ST) has received increasing attention due to the potential of its less error propagation, lower latency, and fewer parameters.
Speech translation: coupling of recognition and translation
H. Ney. 1999 · 1999
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, S. Roukos, T. Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Statistical significance tests for machine translation evaluation
Philipp Koehn. 2004 · 2004
Earlier work this paper cites.
Europarl: A parallel corpus for statistical machine translation
Philipp Koehn. 2005 · 2005
Earlier work this paper cites.
A kernel method for the two-sample-problem
Arthur Gretton, Karsten M. Borgwardt, Malte J. Rasch, Bernhard Schölkopf, and Alex Smola. 2006 · 2006
Earlier work this paper cites.
Neural unsupervised domain adaptation in nlp—a survey
Alan Ramponi and Barbara Plank. 2020 · 2006
Earlier work this paper cites.
On using monolingual corpora in neural machine translation
Caglar Gulcehre, Orhan Firat, Kelvin Xu, Kyunghyun Cho, Loic Barrault, Huei-Chi Lin, Fethi Bougares, Holger Schwenk, and Yoshua Bengio. 2015 · 2015
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba. 2015 · 2015
Earlier work this paper cites.
Exploiting source-side monolingual data in neural machine translation
Jiajun Zhang and Chengqing Zong. 2016 · 2016
Earlier work this paper cites.
An empirical comparison of domain adaptation methods for neural machine translation
Chenhui Chu, Raj Dabre, and Sadao Kurohashi. 2017 · 2017
Earlier work this paper cites.
Neural lattice-to-sequence models for uncertain inputs
Matthias Sperber, Graham Neubig, J. Niehues, and A. Waibel. 2017 · 2017
Earlier work this paper cites.
Instance weighting for neural machine translation domain adaptation
Rui Wang, Masao Utiyama, Lemao Liu, Kehai Chen, and Eiichiro Sumita. 2017 · 2017
Earlier work this paper cites.
Sequence-to-sequence models can directly translate foreign speech
Ron J. Weiss, J. Chorowski, Navdeep Jaitly, Yonghui Wu, and Z. Chen. 2017 · 2017
Earlier work this paper cites.
Tied multitask learning for neural speech translation
Antonios Anastasopoulos and David Chiang. 2018 · 2018
Earlier work this paper cites.
Unsupervised cross-modal alignment of speech and text embedding spaces
Yu-An Chung, Wei-Hung Weng, Schrasing Tong, and James R. Glass. 2018 · 2018
Earlier work this paper cites.
Search engine guided neural machine translation
Jiatao Gu, Yong Wang, Kyunghyun Cho, and Victor O. K. Li. 2018 · 2018
Earlier work this paper cites.
Augmenting librispeech with French translations: A multimodal corpus for direct speech translation evaluation
Ali Can Kocabiyikoglu, Laurent Besacier, and Olivier Kraif. 2018 · 2018
Earlier work this paper cites.
End-to-end speech translation with the transformer
Laura Cross Vila, Carlos Escolano, José A. R. Fonollosa, and Marta Ruiz Costa-jussà. 2018 · 2018
Earlier work this paper cites.
Compact personalized models for neural machine translation
Joern Wuebker, Patrick Simianer, and John DeNero. 2018 · 2018
Cited alongside, same era.
Joint training for neural machine translation models with monolingual data
Zhirui Zhang, Shujie Liu, Mu Li, M. Zhou, and Enhong Chen. 2018 · 2018
Cited alongside, same era.
Pre-training on high-resource speech recognition improves low-resource speech-to-text translation
Sameer Bansal, H. Kamper, Karen Livescu, Adam Lopez, and S. Goldwater. 2019 · 2019
Cited alongside, same era.
Simple, scalable adaptation for neural machine translation
Ankur Bapna, N. Arivazhagan, and Orhan Firat. 2019 · 2019
Cited alongside, same era.
Must-c: a multilingual speech translation corpus
Mattia Antonino Di Gangi, R. Cattoni, L. Bentivogli, Matteo Negri, and M. Turchi. 2019 · 2019
Cited alongside, same era.
Domain adaptation of neural machine translation by lexicon induction
Vatt: Transformers for multimodal self-supervised learning from raw video, audio and text
Hassan Akbari, Li Yuan, Rui Qian, Wei-Hong Chuang, Shih-Fu Chang, Yin Cui, and Boqing Gong. 2021 · 2021
Later among the works it cites.
Slam: A unified encoder for speech and language modeling via speech-text joint pre-training
Ankur Bapna, Yu-An Chung, Nan Wu, Anmol Gulati, Ye Jia, J. Clark, Melvin Johnson, Jason Riesa, Alexis Conneau, and Yu Zhang. 2021 · 2021
Later among the works it cites.
Cascade versus direct speech translation: Do the differences still make a difference?
Luisa Bentivogli, Mauro Cettolo, Marco Gaido, Alina Karakanta, Alberto Martinelli, Matteo Negri, and Marco Turchi. 2021 · 2021
Later among the works it cites.
mbart: Multidimensional monotone bart
Hugh A. Chipman, Edward I. George, Robert E. McCulloch, and Thomas S. Shively. 2021 · 2021
Later among the works it cites.
Consecutive decoding for speech-to-text translation
Qianqian Dong, Mingxuan Wang, Hao Zhou, Shuang Xu, Bo Xu, and Lei Li. 2021 · 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Junjie Hu, M. Xia, Graham Neubig, and Jaime G. Carbonell. 2019 · 2019
Cited alongside, same era.
End-to-end speech translation with knowledge distillation
Yuchen Liu, Hao Xiong, Zhongjun He, Jiajun Zhang, Hua Wu, Haifeng Wang, and Chengqing Zong. 2019 · 2019
Cited alongside, same era.
fairseq: A fast, extensible toolkit for sequence modeling
Myle Ott, Sergey Edunov, Alexei Baevski, Angela Fan, S. Gross, Nathan Ng, David Grangier, and Michael Auli. 2019 · 2019
Cited alongside, same era.
Specaugment: A simple data augmentation method for automatic speech recognition
Daniel S. Park, William Chan, Y. Zhang, C. Chiu, Barret Zoph, E. D. Cubuk, and Quoc V. Le. 2019 · 2019
Cited alongside, same era.
Attention-passing models for robust and data-efficient end-to-end speech translation
Matthias Sperber, Graham Neubig, J. Niehues, and A. Waibel. 2019 · 2019
Cited alongside, same era.
Lattice transformer for speech translation
Peidong Zhang, Boxing Chen, Niyu Ge, and Kai Fan. 2019 · 2019
Cited alongside, same era.
wav2vec 2.0: A framework for self-supervised learning of speech representations
Alexei Baevski, Henry Zhou, Abdel rahman Mohamed, and Michael Auli. 2020 · 2020
Cited alongside, same era.
Later among the works it cites.
Parameter-efficient transfer learning with diff pruning
Demi Guo, Alexander M. Rush, and Yoon Kim. 2021 · 2021
Later among the works it cites.
Learning shared semantic space for speech-to-text translation
Chi Han, Mingxuan Wang, Heng Ji, and Lei Li. 2021 · 2021
Later among the works it cites.
Efficient nearest neighbor language models
Junxian He, Graham Neubig, and Taylor Berg-Kirkpatrick. 2021 · 2021
Later among the works it cites.
Billion-scale similarity search with gpus
Jeff Johnson, Matthijs Douze, and Hervé Jégou. 2021 · 2021
Later among the works it cites.
Cascaded models with cyclic feedback for direct speech translation
Tsz Kin Lam, Shigehiko Schamoni, and Stefan Riezler. 2021 · 2021
Later among the works it cites.
Lightweight adapter tuning for multilingual speech translation
Hang Le, J. Pino, Changhan Wang, Jiatao Gu, D. Schwab, and L. Besacier. 2021 · 2021
Later among the works it cites.
Multilingual speech translation from efficient finetuning of pretrained models
Xian Li, Changhan Wang, Yun Tang, C. Tran, Yuqing Tang, J. Pino, Alexei Baevski, Alexis Conneau, and Michael Auli. 2021 · 2021
Later among the works it cites.
Lost in interpreting: Speech translation from source or interpreter?
Dominik Machácek, Matús Zilinec, and Ondrej Bojar. 2021 · 2021
Later among the works it cites.
Fast nearest neighbor machine translation
Yuxian Meng, Xiaoya Li, Xiayu Zheng, Fei Wu, Xiaofei Sun, Tianwei Zhang, and Jiwei Li. 2021 · 2021
Later among the works it cites.
Improving speech translation by understanding and learning from the auxiliary text translation task
Yun Tang, Juan Pino, Xian Li, Changhan Wang, and Dmitriy Genzel. 2021 · 2021
Later among the works it cites.
Regularizing end-to-end speech translation with triangular decomposition agreement
Yichao Du, Zhirui Zhang, Weizhi Wang, Boxing Chen, Jun Xie, and Tong Xu. 2022 · 2022
Closest in time.
Non-parametric online learning from human feedback for neural machine translation
Dongqi Wang, Haoran Wei, Zhirui Zhang, Shujian Huang, Jun Xie, Weihua Luo, and Jiajun Chen. 2022 · 2022
Closest in time.