Fetching the paper…
Reading the bibliography…
We report on adaptation of multilingual end-to-end speech recognition models trained on as many as 100 languages.
Long short-term memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Bidirectional recurrent neural networks
Mike Schuster and Kuldip K Paliwal. 1997 · 1997
Earlier work this paper cites.
The architecture of the Festival speech synthesis system
Paul Taylor, Alan W Black, and Richard Caley. 1998 · 1998
Earlier work this paper cites.
Experiments on cross-language acoustic Modeling
Tanja Schultz and Alex Waibel. 2001 · 2001
Earlier work this paper cites.
GlobalPhone: a multilingual speech and text database developed at Karlsruhe University
Tanja Schultz. 2002 · 2002
Earlier work this paper cites.
First steps in fast acoustic modeling for a new target language: application to Vietnamese
Viet Bac Le and Laurent Besacier. 2005 · 2005
Earlier work this paper cites.
Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks
Alex Graves, Santiago Fernandez, Faustino Gomez, and Jurgen Schmidhuber. 2006 · 2006
Earlier work this paper cites.
Cross-domain and cross-language portability of acoustic features estimated by multilayer perceptrons
A. Stolcke, F. Grezl, Mei-Yuh Hwang, Xin Lei, N. Morgan, and D. Vergyri. 2006 · 2006
Earlier work this paper cites.
Multilingual transliteration using feature based phonetic method
Su-Youn Yoon, Kyoung-Young Kim, and Richard Sproat. 2007 · 2007
Earlier work this paper cites.
On the use of a multilingual neural network front-end
Stefano Scanzio, Pietro Laface, Luciano Fissore, Roberto Gemello, and Franco Mana. 2008 · 2008
Earlier work this paper cites.
Cross-lingual portability of MLP-based tandem features - a case study for English and Hungarian
László Tóth, Joe Frankel, Gábor Gosztolya, and Simon King. 2008 · 2008
Earlier work this paper cites.
Visualizing data using t-SNE
Laurens Van Der Maaten and Geoffrey Hinton. 2008 · 2008
Earlier work this paper cites.
Cross-lingual portability of Chinese and English neural network features for French and German LVCSR
Christian Plahl, Ralf Schlüter, and Hermann Ney. 2011 · 2011
Earlier work this paper cites.
The Kaldi speech recognition toolkit
Daniel Povey, Arnab Ghoshal, Gilles Boulianne, Lukas Burget, Ondrej Glembek, Nagendra Goel, Mirko Hannemann, Petr Motlicek, Yanmin Qian, Petr Schwarz, and Others. 2011 · 2011
Earlier work this paper cites.
Multilingual MLP features for low-resource LVCSR systems
Samuel Thomas, Sriram Ganapathy, Hynek Hermansky, and Speech Processing. 2012 · 2012
Earlier work this paper cites.
The language-independent bottleneck features
Karel Vesely, Martin Karafiát, Frantisek Grezl, Marcel Janda, and Ekaterina Egorova. 2012 · 2012
Earlier work this paper cites.
Multilingual bottle-neck features and its application for under-resourced languages
Ngoc Thang Vu, Florian Metze, and Tanja Schultz. 2012 · 2012
Cited alongside, same era.
Speech recognition with deep recurrent neural networks
A Graves, A.-R. Mohamed, and G Hinton. 2013 · 2013
Cited alongside, same era.
Multilingual acoustic models using distributed deep neural networks
Georg Heigold, Vincent Vanhoucke, Alan Senior, Patrick Nguyen, Marc’Aurelio Ranzato, Matthieu Devin, and Jeffrey Dean. 2013 · 2013
Cited alongside, same era.
Cross-lingual phone mapping for large vocabulary speech recognition of under-resourced languages
Van Hai Do, Xiong Xiao, Eng Siong Chng, and Haizhou Li. 2014 · 2014
Cited alongside, same era.
Speech recognition and keyword spotting for low-resource languages: BABEL project research at CUED
Mark J F Gales, Kate M Knill, Anton Ragni, and Shakti P Rath. 2014 · 2014
Cited alongside, same era.
Phonemic and Graphemic Multilingual CTC Based Speech Recognition
Markus Müller, Sebastian Stüker, and Alex Waibel. 2017 · 2017
Later among the works it cites.
Automatic differentiation in pytorch
Adam Paszke, Sam Gross, Soumith Chintala, Gregory Chanan, Edward Yang, Zachary DeVito, Zeming Lin, Alban Desmaison, Luca Antiga, and Adam Lerer. 2017 · 2017
Later among the works it cites.
Multi-accent speech recognition with hierarchical grapheme based models
Kanishka Rao and Haşim Sak. 2017 · 2017
Later among the works it cites.
Building an ASR system for a low-research language through the adaptation of a high-resource language ASR system: preliminary results
Odette Scharenborg, Francesco Ciannella, Shruti Palaskar, Alan Black, Florian Metze, Lucas Ondel, and Mark Hasegawa-Johnson. 2017 · 2017
Later among the works it cites.
Multilingual speech recognition with a single end-to-end model
Shubham Toshniwal, Tara N. Sainath, Ron J. Weiss, Bo Li, Pedro Moreno, Eugene Weinstein, and Kanishka Rao. 2017 · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Using out-of-language data to improve an under-resourced speech recognizer
David Imseng, Petr Motlicek, Hervé Bourlard, and Philip N Garner. 2014 · 2014
Cited alongside, same era.
Very deep convolutional networks for large-scale image recognition
Karen Simonyan and Andrew Zisserman. 2014 · 2014
Cited alongside, same era.
William Chan, Navdeep Jaitly, Quoc V. Le, and Oriol Vinyals. 2015 · 2015
Cited alongside, same era.
Attention-based models for speech recognition
Jan K Chorowski, Dzmitry Bahdanau, Dmitriy Serdyuk, Kyunghyun Cho, and Yoshua Bengio. 2015 · 2015
Cited alongside, same era.
Domain-adversarial training of neural networks
Yaroslav Ganin, Evgeniya Ustinova, Hana Ajakan, Pascal Germain, Hugo Larochelle, François Laviolette, Mario Marchand, and Victor Lempitsky. 2016 · 2016
Cited alongside, same era.
Joint CTC-attention based end-to-end speech recognition using multi-task learning
Suyoun Kim, Takaaki Hori, and Shinji Watanabe. 2016 · 2016
Cited alongside, same era.
Very deep multilingual convolutional neural networks for LVCSR
Tom Sercu, Christian Puhrsch, Brian Kingsbury, and Yorktown Heights. 2016 · 2016
Cited alongside, same era.
Later among the works it cites.
State-of-the-art speech recognition with sequence-to-sequence models
Chung-Cheng Chiu, Tara N Sainath, Yonghui Wu, Rohit Prabhavalkar, Patrick Nguyen, Zhifeng Chen, Anjuli Kannan, Ron J Weiss, Kanishka Rao, Katya Gonina, et al. 2018 · 2018
Later among the works it cites.
Sequence-based multi-lingual low resource speech recognition
Siddharth Dalmia, Ramon Sanabria, Florian Metze, and Alan W Black. 2018 · 2018
Later among the works it cites.
Transfer learning of language-independent end-to-end ASR with language model fusion
Hirofumi Inaguma, Jaejin Cho, Murali Karthick Baskar, Tatsuya Kawahara, and Shinji Watanabe. 2018 · 2018
Later among the works it cites.
Analysis of multilingual sequence-to-sequence speech recognition systems
Martin Karafiát, Murali Karthick Baskar, Shinji Watanabe, Takaaki Hori, Matthew Wiesner, and Jan “Honza” Černocký. 2018 · 2018
Later among the works it cites.
Hierarchical multitask learning for CTC-based speech recognition
Kalpesh Krishna, Shubham Toshniwal, and Karen Livescu. 2018 · 2018
Later among the works it cites.
Integrating automatic transcription into the language documentation workflow: Experiments with Na data and the Persephone toolkit
Alexis Michaud, Oliver Adams, Trevor Anthony Cohn, Graham Neubig, and Séverine Guillaume. 2018 · 2018
Later among the works it cites.
Hierarchical multi task learning with CTC
Ramon Sanabria and Florian Metze. 2018 · 2018
Later among the works it cites.
Domain adversarial training for accented speech recognition
Sining Sun. 2018 · 2018
Later among the works it cites.
Adversarial learning of raw speech features for domain invariant speech recognition
Aditay Tripathi, Aanchan Mohan, Saket Anand, and Maneesh Singh. 2018 · 2018
Later among the works it cites.
Adversarial multilingual training for low-resource speech recognition
Jiangyan Yi, Jianhua Tao, Zhengqi Wen, and Ye Bai. 2018 · 2018
Later among the works it cites.
CMU Wilderness Multilingual Speech Dataset
Alan W Black. 2019 · 2019
Closest in time.