Fetching the paper…
Reading the bibliography…
We present a simple approach to improve direct speech-to-text translation (ST) when the source language is low-resource: we pre-train the model on a high-resource automatic speech recognition (ASR) task, and then fine-tune its parameters for ST.
A learning algorithm for continually running fully recurrent neural networks
Ronald J. Williams and David Zipser. 1989 · 1989
Earlier work this paper cites.
Switchboard-1 Release 2 (LDC97S62)
John Godfrey and Edward Holliman. 1993 · 1993
Earlier work this paper cites.
Is learning the n-th thing any easier than learning the first?
Sebastian Thrun. 1995 · 1995
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
BLEU: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Globalphone: a multilingual speech and text database developed at karlsruhe university
Tanja Schultz. 2002 · 2002
Earlier work this paper cites.
Towards speech translation of non written languages
Laurent Besacier, Bowen Zhou, and Yuqing Gao. 2006 · 2006
Earlier work this paper cites.
Moses: Open source toolkit for statistical machine translation
Philipp Koehn, Hieu Hoang, Alexandra Birch, Chris Callison-Burch, Marcello Federico, Nicola Bertoldi, Brooke Cowan, Wade Shen, Christine Moran, Richard Zens, et al. 2007 · 2007
Earlier work this paper cites.
Meteor: An automatic metric for mt evaluation with high levels of correlation with human judgments
Alon Lavie and Abhaya Agarwal. 2007 · 2007
Earlier work this paper cites.
Fisher Spanish Speech (LDC2010S01)
David Graff, Shudong Huang, Ingrid Cartagena, Kevin Walker, and Christopher Cieri. 2010 · 2010
Earlier work this paper cites.
Towards spoken term discovery at scale with zero resources
Aren Jansen, Kenneth Church, and Hynek Hermansky. 2010 · 2010
Earlier work this paper cites.
Crowdsourced translation for emergency response in Haiti: The global collaboration of local knowledge
Robert Munro. 2010 · 2010
Earlier work this paper cites.
Rectified linear units improve restricted Boltzmann machines
Vinod Nair and Geoffrey E. Hinton. 2010 · 2010
Earlier work this paper cites.
Rapid evaluation of speech representations for spoken term discovery
Michael A Carlin, Samuel Thomas, Aren Jansen, and Hynek Hermansky. 2011 · 2011
Earlier work this paper cites.
The Kaldi Speech Recognition Toolkit
Daniel Povey, Arnab Ghoshal, Gilles Boulianne, Lukas Burget, Ondrej Glembek, Nagendra Goel, Mirko Hannemann, Petr Motlicek, Yanmin Qian, Petr Schwarz, Jan Silovsky, Georg Stemmer, and Karel Vesely. 2011 · 2011
Earlier work this paper cites.
Multilingual mlp features for low-resource LVCSR systems
Samuel Thomas, Sriram Ganapathy, and Hynek Hermansky. 2012 · 2012
Earlier work this paper cites.
An investigation on initialization schemes for multilayer perceptron training using multilingual data and their effect on ASR performance
Ngoc Thang Vu, Wojtek Breiter, Florian Metze, and Tanja Schultz. 2012 · 2012
Cited alongside, same era.
Recent advances in deep learning for speech research at Microsoft
Li Deng, Jinyu Li, Jui-Ting Huang, Kaisheng Yao, Dong Yu, Frank Seide, Mike Seltzer, Geoff Zweig, Xiaodong He, Jason Williams, Yifan Gong, and Alex Acero. 2013 · 2013
Cited alongside, same era.
Improved speech-to-text translation with the Fisher and Callhome Spanish-English speech translation corpus
Matt Post, Gaurav Kumar, Adam Lopez, Damianos Karakos, Chris Callison-Burch, and Sanjeev Khudanpur. 2013 · 2013
Cited alongside, same era.
Fisher and CALLHOME Spanish–English Speech Translation
Matt Post, Gaurav Kumar, Adam Lopez, Damianos Karakos, Chris Callison-Burch, and Sanjeev Khudanpur. 2014 · 2014
Cited alongside, same era.
Dropout: A simple way to prevent neural networks from overfitting
Nitish Srivastava, Geoffrey Hinton, Alex Krizhevsky, Ilya Sutskever, and Ruslan Salakhutdinov. 2014 · 2014
A theoretically grounded application of dropout in recurrent neural networks
Yarin Gal. 2016 · 2016
Later among the works it cites.
Neural machine translation of rare words with subword units
Rico Sennrich, Barry Haddow, and Alexandra Birch. 2016 · 2016
Later among the works it cites.
Google’s neural machine translation system: Bridging the gap between human and machine translation
Yonghui Wu, Mike Schuster, Zhifeng Chen, Quoc V. Le, Mohammad Norouzi, Wolfgang Macherey, Maxim Krikun, Yuan Cao, Qin Gao, Klaus Macherey, Jeff Klingner, Apurva Shah, Melvin Johnson, Xiaobing Liu, Lukasz Kaiser, Stephan Gouws, Yoshikiyo Kato, Taku Kudo, Hideto Kazawa, Keith Stevens, George Kurian, Nishant Patil, Wei Wang, Cliff Young, Jason Smith, Jason Riesa, Alex Rudnick, Oriol Vinyals, Gregory S. Corrado, Macduff Hughes, and Jeffrey Dean. 2016 · 2016
Later among the works it cites.
Learning neural network representations using cross-lingual bottleneck features with word-pair information
Yougen Yuan, Cheung-Chi Leung, Lei Xie, Bin Ma, and Haizhou Li. 2016 · 2016
Later among the works it cites.
Transfer learning for low-resource neural machine translation
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Multilingual representations for low resource speech recognition and keyword search
Jia Cui, Brian Kingsbury, Bhuvana Ramabhadran, Abhinav Sethy, Kartik Audhkhasi, Xiaodong Cui, Ellen Kislal, Lidia Mangu, Markus Nussbaum-Thom, Michael Picheny, et al. 2015 · 2015
Cited alongside, same era.
On using monolingual corpora in neural machine translation
Caglar Gülçehre, Orhan Firat, Kelvin Xu, Kyunghyun Cho, Loıc Barrault, Huei-Chi Lin, Fethi Bougares, Holger Schwenk, and Yoshua Bengio. 2015 · 2015
Cited alongside, same era.
Delving deep into rectifiers: Surpassing human-level performance on Imagenet classification
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2015 · 2015
Cited alongside, same era.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Sergey Ioffe and Christian Szegedy. 2015 · 2015
Cited alongside, same era.
Unsupervised neural network based feature extraction using weak top-down constraints
Herman Kamper, Micha Elsner, Aren Jansen, and Sharon Goldwater. 2015 · 2015
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik Kingma and Jimmy Ba. 2015 · 2015
Cited alongside, same era.
Effective approaches to attention-based neural machine translation
Minh-Thang Luong, Hieu Pham, and Christopher D Manning. 2015 · 2015
Cited alongside, same era.
Barret Zoph, Deniz Yuret, Jonathan May, and Kevin Knight. 2016 · 2016
Later among the works it cites.
A case study on using speech-to-translation alignments for language documentation
Antonios Anastasopoulos and David Chiang. 2017 · 2017
Later among the works it cites.
Google’s multilingual neural machine translation system: Enabling zero-shot translation
Melvin Johnson, Mike Schuster, Quoc V. Le, Maxim Krikun, Yonghui Wu, Zhifeng Chen, Nikhil Thorat, Fernanda B. Viégas, Martin Wattenberg, Gregory S. Corrado, Macduff Hughes, and Jeffrey Dean. 2017 · 2017
Later among the works it cites.
A segmental framework for fully-unsupervised large-vocabulary speech recognition
Herman Kamper, Aren Jansen, and Sharon Goldwater. 2017 · 2017
Later among the works it cites.
Unsupervised pretraining for sequence to sequence learning
Prajit Ramachandran, Peter J Liu, and Quoc V Le. 2017 · 2017
Later among the works it cites.
Sequence-to-sequence models can directly transcribe foreign speech
Ron J Weiss, Jan Chorowski, Navdeep Jaitly, Yonghui Wu, and Zhifeng Chen. 2017 · 2017
Later among the works it cites.
Tied multitask learning for neural speech translation
Antonios Anastasopoulos and David Chiang. 2018 · 2018
Closest in time.
Low-resource speech-to-text translation
Sameer Bansal, Herman Kamper, Karen Livescu, Adam Lopez, and Sharon Goldwater. 2018 · 2018
Closest in time.
End-to-end automatic speech translation of audiobooks
Alexandre Bérard, Laurent Besacier, Ali Can Kocabiyikoglu, and Olivier Pietquin. 2018 · 2018
Closest in time.
A very low resource language speech corpus for computational language documentation experiments
Pierre Godard, Gilles Adda, Martine Adda-Decker, Juan Benjumea, Laurent Besacier, Jamison Cooper-Leavitt, Guy-No”el Kouarata, Lori Lamel, H’el‘ene Maynard, Markus M”uller, Annie Rialland, Sebastian St”uker, François Yvon, and Marcely Zanon Boito. 2018 · 2018
Closest in time.
Multilingual bottleneck features for subword modeling in zero-resource languages
Enno Hermann and Sharon Goldwater. 2018 · 2018
Closest in time.