Fetching the paper…
Reading the bibliography…
Having numerous potential applications and great impact, end-to-end speech translation (ST) has long been treated as an independent task, failing to fully draw strength from the rapid advances of its sibling - text machine translation (MT).
Speech production: Wernicke, broca and beyond
S Catrin Blank, Sophie K Scott, Kevin Murphy, Elizabeth Warburton, and Richard JS Wise. 2002 · 2002
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Neural correlates of lexical access during visual word recognition
Jeffrey R Binder, Kristen A McKiernan, Melanie E Parsons, Chris F Westbury, Edward T Possing, Jacqueline N Kaufman, and Lori Buchanan. 2003 · 2003
Earlier work this paper cites.
Integration of letters and speech sounds in the human brain
Nienke van Atteveldt, Elia Formisano, Rainer Goebel, and Leo Blomert. 2004 · 2004
Earlier work this paper cites.
Automatic sentence segmentation and punctuation prediction for spoken language translation
Evgeny Matusov, Arne Mauser, and Hermann Ney. 2006 · 2006
Earlier work this paper cites.
Converging language streams in the human temporal lobe
Galina Spitsyna, Jane E Warren, Sophie K Scott, Federico E Turkheimer, and Richard JS Wise. 2006 · 2006
Earlier work this paper cites.
A system for simultaneous translation of lectures and speeches
Christian Fügen. 2008 · 2008
Earlier work this paper cites.
Reading differences and brain: Cortical integration of speech and print in sentence processing varies with reader skill
Donald Shankweiler, W Einar Mencl, David Braze, Whitney Tabor, Kenneth R Pugh, and Robert K Fulbright. 2008 · 2008
Earlier work this paper cites.
Reconstructing false start errors in spontaneous speech text
Erin Fitzgerald, Keith B. Hall, and Frederick Jelinek. 2009 · 2009
Earlier work this paper cites.
Bridging the modality gap for speech-to-text translation
Yuchen Liu, Junnan Zhu, Jiajun Zhang, and Chengqing Zong. 2020 · 2010
Earlier work this paper cites.
Findings of the 2014 workshop on statistical machine translation
Ondřej Bojar, Christian Buck, Christian Federmann, Barry Haddow, Philipp Koehn, Johannes Leveling, Christof Monz, Pavel Pecina, Matt Post, Herve Saint-Amand, et al. 2014 · 2014
Earlier work this paper cites.
Listen and translate: A proof of concept for end-to-end speech-to-text translation
Alexandre Bérard, Olivier Pietquin, Christophe Servan, and Laurent Besacier. 2016 · 2016
Earlier work this paper cites.
Findings of the 2016 conference on machine translation
Ondřej Bojar, Rajen Chatterjee, Christian Federmann, Yvette Graham, Barry Haddow, Matthias Huck, Antonio Jimeno Yepes, Philipp Koehn, Varvara Logacheva, Christof Monz, et al. 2016 · 2016
Earlier work this paper cites.
An attentional model for speech translation without transcription
Long Duong, Antonios Anastasopoulos, David Chiang, Steven Bird, and Trevor Cohn. 2016 · 2016
Earlier work this paper cites.
Opensubtitles2016: Extracting large parallel corpora from movie and TV subtitles
Pierre Lison and Jörg Tiedemann. 2016 · 2016
Earlier work this paper cites.
Must-c: a multilingual speech translation corpus
Mattia A Di Gangi, Roldano Cattoni, Luisa Bentivogli, Matteo Negri, and Marco Turchi. 2019a · 2017
Earlier work this paper cites.
Structured-based curriculum learning for end-to-end english-japanese speech translation
Takatomo Kano, Sakriani Sakti, and Satoshi Nakamura. 2017 · 2017
Earlier work this paper cites.
Neural lattice-to-sequence models for uncertain inputs
Matthias Sperber, Graham Neubig, Jan Niehues, and Alex Waibel. 2017 · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
Sequence-to-sequence models can directly translate foreign speech
Ron J. Weiss, Jan Chorowski, Navdeep Jaitly, Yonghui Wu, and Zhifeng Chen. 2017 · 2017
Cited alongside, same era.
End-to-end automatic speech translation of audiobooks
Alexandre Bérard, Laurent Besacier, Ali Can Kocabiyikoglu, and Olivier Pietquin. 2018 · 2018
Cited alongside, same era.
Towards robust neural machine translation
Yong Cheng, Zhaopeng Tu, Fandong Meng, Junjie Zhai, and Yang Liu. 2018 · 2018
Cited alongside, same era.
The iwslt 2018 evaluation campaign
Niehues Jan, Roldano Cattoni, Stüker Sebastian, Mauro Cettolo, Marco Turchi, and Marcello Federico. 2018 · 2018
Cited alongside, same era.
Augmenting librispeech with french translations: A multimodal corpus for direct speech translation evaluation
Ali Can Kocabiyikoglu, Laurent Besacier, and Olivier Kraif. 2018 · 2018
Cited alongside, same era.
End-to-end speech translation with knowledge distillation
Yuchen Liu, Hao Xiong, Jiajun Zhang, Zhongjun He, Hua Wu, Haifeng Wang, and Chengqing Zong. 2019 · 2019
Later among the works it cites.
fairseq: A fast, extensible toolkit for sequence modeling
Myle Ott, Sergey Edunov, Alexei Baevski, Angela Fan, Sam Gross, Nathan Ng, David Grangier, and Michael Auli. 2019 · 2019
Later among the works it cites.
Harnessing indirect training data for end-to-end automatic speech translation: Tricks of the trade
Juan Pino, Liezl Puzon, Jiatao Gu, Xutai Ma, Arya D McCarthy, and Deepak Gopinath. 2019 · 2019
Later among the works it cites.
Fluent translations from disfluent speech in end-to-end speech translation
Elizabeth Salesky, Matthias Sperber, and Alexander H. Waibel. 2019 · 2019
Later among the works it cites.
Self-attentional models for lattice inputs
Matthias Sperber, Graham Neubig, Ngoc-Quan Pham, and Alex Waibel. 2019 · 2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Dan Liu, Junhua Liu, Wu Guo, Shifu Xiong, Zhiqiang Ma, Rui Song, Chongliang Wu, and Quan Liu. 2018 · 2018
Cited alongside, same era.
A neural interlingua for multilingual machine translation
Yichao Lu, Phillip Keung, Faisal Ladhak, Vikas Bhardwaj, Shaonan Zhang, and Jason Sun. 2018 · 2018
Cited alongside, same era.
Towards fluent translations from disfluent speech
Elizabeth Salesky, Susanne Burger, Jan Niehues, and Alex Waibel. 2018 · 2018
Cited alongside, same era.
End-to-end speech translation with the transformer
Laura Cross Vila, Carlos Escolano, José AR Fonollosa, and Marta R Costa-jussà. 2018 · 2018
Cited alongside, same era.
Multilingual seq2seq training with similarity loss for cross-lingual document classification
Katherine Yu, Haoran Li, and Barlas Oguz. 2018 · 2018
Cited alongside, same era.
A comparative study on end-to-end speech to text translation
Parnia Bahar, Tobias Bieschke, and Hermann Ney. 2019a · 2019
Cited alongside, same era.
On using specaugment for end-to-end speech translation
Parnia Bahar, Albert Zeyer, Ralf Schlüter, and Hermann Ney. 2019b · 2019
Cited alongside, same era.
Juan Raul Vazquez Carrillo, Alessandro Raganato, Jörg Tiedemann, Mathias Creutz, et al. 2019 · 2019
Later among the works it cites.
Improving multilingual sentence embedding using bi-directional dual encoder with additive margin softmax
Yinfei Yang, Gustavo Hernández Ábrego, Steve Yuan, Mandy Guo, Qinlan Shen, Daniel Cer, Yun-Hsuan Sung, Brian Strope, and Ray Kurzweil. 2019 · 2019
Later among the works it cites.
Lattice transformer for speech translation
Pei Zhang, Niyu Ge, Boxing Chen, and Kai Fan. 2019 · 2019
Later among the works it cites.
wav2vec 2.0: A framework for self-supervised learning of speech representations
Alexei Baevski, Yuhao Zhou, Abdelrahman Mohamed, and Michael Auli. 2020 · 2020
Later among the works it cites.
Instance-based model adaptation for direct speech translation
Mattia A Di Gangi, Viet-Nhat Nguyen, Matteo Negri, and Marco Turchi. 2020 · 2020
Later among the works it cites.
Espnet-st: All-in-one speech translation toolkit
Hirofumi Inaguma, Shun Kiyono, Kevin Duh, Shigeki Karita, Nelson Yalta, Tomoki Hayashi, and Shinji Watanabe. 2020 · 2020
Later among the works it cites.
Data efficient direct speech-to-text translation with modality agnostic meta-learning
Sathish Indurthi, Houjeung Han, Nikhil Kumar Lakumarapu, Beomseok Lee, Insoo Chung, Sangha Kim, and Chanwoo Kim. 2020 · 2020
Later among the works it cites.
Dual-decoder transformer for joint automatic speech recognition and multilingual speech translation
Hang Le, Juan Pino, Changhan Wang, Jiatao Gu, Didier Schwab, and Laurent Besacier. 2020 · 2020
Later among the works it cites.
Self-training for end-to-end speech translation
Juan Pino, Qiantong Xu, Xutai Ma, Mohammad Javad Dousti, and Yun Tang. 2020 · 2020
Later among the works it cites.
Analyzing asr pretraining for low-resource speech-to-text translation
Mihaela C Stoian, Sameer Bansal, and Sharon Goldwater. 2020 · 2020
Later among the works it cites.
Language-aware interlingua for multilingual neural machine translation
Changfeng Zhu, Heng Yu, Shanbo Cheng, and Weihua Luo. 2020 · 2020
Later among the works it cites.
A general multi-task learning framework to leverage text data for speech to text tasks
Yun Tang, Juan Pino, Changhan Wang, Xutai Ma, and Dmitriy Genzel. 2021 · 2021
Closest in time.
Jointly trained transformers models for spoken language translation
Hari Krishna Vydana, Martin Karafiát, Katerina Zmolikova, Lukáš Burget, and Honza Černockỳ. 2021 · 2021
Closest in time.