Fetching the paper…
Reading the bibliography…
Speech translation has traditionally been approached through cascaded models consisting of a speech recognizer trained on a corpus of transcribed speech, and a machine translation system trained on parallel texts.
The Fisher Corpus: a Resource for the Next Generations of Speech-to-Text
Christopher Cieri, David Miller, and Kevin Walker. 2004 · 2004
Earlier work this paper cites.
Improved Speech-to-Text Translation with the Fisher and Callhome Spanish–English Speech Translation Corpus
Matt Post, Gaurav Kumar, Adam Lopez, Damianos Karakos, Chris Callison-Burch, and Sanjeev Khudanpur. 2013 · 2013
Earlier work this paper cites.
Adam: A Method for Stochastic Optimization
Diederik P. Kingma and Jimmy L. Ba. 2014 · 2014
Earlier work this paper cites.
Neural Machine Translation by Jointly Learning to Align and Translate
Dzmitry Bahdanau, KyungHyun Cho, and Yoshua Bengio. 2015 · 2015
Earlier work this paper cites.
Attention-Based Models for Speech Recognition
Jan K. Chorowski, Dzmitry Bahdanau, Dmitriy Serdyuk, Kyunghyun Cho, and Yoshua Bengio. 2015 · 2015
Earlier work this paper cites.
Librispeech: an ASR corpus based on public domain audio books
Vassil Panayotov, Guoguo Chen, Daniel Povey, and Sanjeev Khudanpur. 2015 · 2015
Earlier work this paper cites.
Waleed Ammar, George Mulcaire, Miguel Ballesteros, Chris Dyer, and Noah A. Smith. 2016 · 2016
Earlier work this paper cites.
Listen, attend and spell: A neural network for large vocabulary conversational speech recognition
William Chan, Navdeep Jaitly, Quoc V. Le, and Oriol Vinyals. 2016 · 2016
Earlier work this paper cites.
An Attentional Model for Speech Translation Without Transcription
Long Duong, Antonios Anastasopoulos, David Chiang, Steven Bird, and Trevor Cohn. 2016 · 2016
Earlier work this paper cites.
A Theoretically Grounded Application of Dropout in Recurrent Neural Networks
Yarin Gal and Zoubin Ghahramani. 2016 · 2016
Earlier work this paper cites.
OpenSubtitles2016: Extracting Large Parallel Corpora from Movie and TV Subtitles
Pierre Lison and Jörg Tiedemann. 2016 · 2016
Cited alongside, same era.
Rethinking the Inception Architecture for Computer Vision
Christian Szegedy, Vincent Vanhoucke, Sergey Ioffe, Jon Shlens, and Zbigniew Wojna. 2016 · 2016
Cited alongside, same era.
Structured-based Curriculum Learning for End-to-end English-Japanese Speech Translation
Takatomo Kano, Sakriani Sakti, and Satoshi Nakamura. 2017 · 2017
Cited alongside, same era.
Exploiting Linguistic Resources for Neural Machine Translation Using Multi-task Learning
Jan Niehues and Eunah Cho. 2017 · 2017
Cited alongside, same era.
Toward Robust Neural Machine Translation for Noisy Input Sequences
Matthias Sperber, Jan Niehues, and Alex Waibel. 2017 · 2017
Cited alongside, same era.
Deliberation Networks: Sequence Generation
Yingce Xia, Fei Tian, Lijun Wu, Jianxin Lin, Tao Qin, Nenghai Yu, and Tie-Yan Liu. 2017 · 2017
Later among the works it cites.
Very Deep Convolutional Networks for End-to-End Speech Recognition
Yu Zhang, William Chan, and Navdeep Jaitly. 2017 · 2017
Later among the works it cites.
Tied Multitask Learning for Neural Speech Translation
Antonios Anastasopoulos and David Chiang. 2018 · 2018
Later among the works it cites.
Low-Resource Speech-to-Text Translation
Sameer Bansal, Herman Kamper, Karen Livescu, Adam Lopez, and Sharon Goldwater. 2018 · 2018
Later among the works it cites.
End-to-End Automatic Speech Translation of Audiobooks
Alexandre Bérard, Laurent Besacier, Ali Can Kocabiyikoglu, and Olivier Pietquin. 2018 · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Andros Tjandra, Sakriani Sakti, and Satoshi Nakamura. 2017 · 2017
Cited alongside, same era.
Multitask Learning with Low-Level Auxiliary Tasks for Encoder-Decoder Based Speech Recognition
Shubham Toshniwal, Hao Tang, Liang Lu, and Karen Livescu. 2017 · 2017
Cited alongside, same era.
Neural Machine Translation with Reconstruction
Zhaopeng Tu, Yang Liu, Lifeng Shang, Xiaohua Liu, and Hang Li. 2017 · 2017
Cited alongside, same era.
Sequence-to-Sequence Models Can Directly Transcribe Foreign Speech
Ron J. Weiss, Jan Chorowski, Navdeep Jaitly, Yonghui Wu, and Zhifeng Chen. 2017 · 2017
Cited alongside, same era.
Later among the works it cites.
Ali Can Kocabiyikoglu, Laurent Besacier, and Olivier Kraif. 2018 · 2018
Later among the works it cites.
XNMT: The eXtensible Neural Machine Translation Toolkit
Graham Neubig, Matthias Sperber, Xinyi Wang, Matthieu Felix, Austin Matthews, Sarguna Padmanabhan, Ye Qi, Devendra Singh Sachan, Philip Arthur, Pierre Godard, John Hewitt, Rachid Riad, and Liming Wang. 2018 · 2018
Later among the works it cites.
Improving Lexical Choice in Neural Machine Translation
Toan Q. Nguyen and David Chiang. 2018 · 2018
Later among the works it cites.
Pre-training on high-resource speech recognition improves low-resource speech-to-text translation
Sameer Bansal, Herman Kamper, Karen Livescu, Adam Lopez, and Sharon Goldwater. 2019 · 2019
Closest in time.