Fetching the paper…
Reading the bibliography…
Automatic Speech Recognition (ASR) systems are often optimized to work best for speakers with canonical speech patterns.
Motor Speech Disorders
Frederick L. Darley, Arnold E. Aronson, and Joe R. Brown. 1975 · 1975
Earlier work this paper cites.
Speech recognition with deep recurrent neural networks
Alex Graves, Abdel-rahman Mohamed, and Geoffrey Hinton. 2013 · 2013
Earlier work this paper cites.
Speaker adaptation of neural network acoustic models using i-vectors
George Saon, Hagen Soltau, David Nahamoo, and Michael Picheny. 2013 · 2013
Earlier work this paper cites.
Severity-based adaptation with limited data for asr to aid dysarthric speakers
Mumtaz Begum Mustafa, Siti Salwah Salim, Noraini Mohamed, Bassam Al-Qatab, and Chng Eng Siong. 2014 · 2014
Earlier work this paper cites.
Fast domain adaptation for neural machine translation
Markus Freitag and Yaser Al-Onaizan. 2016 · 2016
Earlier work this paper cites.
Learning hidden unit contributions for unsupervised acoustic model adaptation
Pawel Swietojanski, Jinyu Li, and Steve Renals. 2016 · 2016
Earlier work this paper cites.
Learning multiple visual domains with residual adapters
Sylvestre-Alvise Rebuffi, Hakan Bilen, and Andrea Vedaldi. 2017 · 2017
Earlier work this paper cites.
Efficient implementation of recurrent neural network transducer in TensorFlow
Tom Bagby, Kanishka Rao, and Khe Chai Sim. 2018 · 2018
Earlier work this paper cites.
Leveraging native language information for improved accented speech recognition
Shahram Ghorbani and John Hansen. 2018 · 2018
Earlier work this paper cites.
Whistle-blowing asrs: Evaluating the need for more inclusive automatic speech recognition systems
Meredith Moore, Hemanth Demakethepalli Venkateswara, and Sethuraman Panchanathan. 2018 · 2018
Earlier work this paper cites.
Simple, scalable adaptation for neural machine translation
Ankur Bapna and Orhan Firat. 2019 · 2019
Cited alongside, same era.
Parrotron: An end-to-end speech-to-speech conversion model and its applications to hearing-impaired speech and speech separation
F. Biadsy, R. J. Weiss, P. J. Moreno, D. Kanvesky, and Y. Jia. 2019 · 2019
Cited alongside, same era.
BERT: pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Improving ASR Systems for Children with Autism and Language Impairment Using Domain-Focused DNN Transfer Techniques
Robert Gale, Liu Chen, Jill Dolata, Jan van Santen, and Meysam Asgari. 2019 · 2019
Cited alongside, same era.
Streaming end-to-end speech recognition for mobile devices
Yanzhang He, Tara N Sainath, Rohit Prabhavalkar, Ian McGraw, Raziel Alvarez, Ding Zhao, David Rybach, Anjuli Kannan, Yonghui Wu, Ruoming Pang, et al. 2019 · 2019
Cited alongside, same era.
How to fine-tune bert for text classification?
Chi Sun, Xipeng Qiu, Yige Xu, and Xuanjing Huang. 2019 · 2019
Later among the works it cites.
Multi-Accent Adaptation Based on Gate Mechanism
Han Zhu, Li Wang, Pengyuan Zhang, and Yonghong Yan. 2019 · 2019
Later among the works it cites.
Common voice: A massively-multilingual speech corpus
Rosana Ardila, Megan Branson, Kelly Davis, Michael Henretty, Michael Kohler, Josh Meyer, Reuben Morais, Lindsay Saunders, Francis M. Tyers, and Gregor Weber. 2020 · 2020
Later among the works it cites.
Extending parrotron: An end-to-end, speech conversion andspeech recognition model for atypical speech
Rohan Doshi, Youzheng Chen, Jiang Liyang, Xia Zhang, Biadsy Fadi, Ramabhadran Bhuvana, Chu Fang, Andrew Rosenberg, and Pedro J. Moreno. 2020 · 2020
Later among the works it cites.
A streaming on-device end-to-end model surpassing server-side conventional model quality and latency
Tara N Sainath, Yanzhang He, Bo Li, Arun Narayanan, Ruoming Pang, Antoine Bruguier, Shuo-yiin Chang, Wei Li, Raziel Alvarez, Zhifeng Chen, et al. 2020 · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Parameter-efficient transfer learning for nlp
Neil Houlsby, Andrei Giurgiu, Stanislaw Jastrzebski, Bruna Morrone, Quentin De Laroussilhe, Andrea Gesmundo, Mona Attariyan, and Sylvain Gelly. 2019 · 2019
Cited alongside, same era.
Large-scale multilingual speech recognition with a streaming end-to-end model
Anjuli Kannan, Arindrima Datta, Tara Sainath, Eugene Weinstein, Bhuvana Ramabhadran, Yonghui Wu, Ankur Bapna, and Zhifeng Chen. 2019 · 2019
Cited alongside, same era.
Recognizing long-form speech using streaming end-to-end models
Arun Narayanan, Rohit Prabhavalkar, Chung-Cheng Chiu, David Rybach, Tara N. Sainath, and Trevor Strohman. 2019 · 2019
Cited alongside, same era.
SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition
Daniel S. Park, William Chan, Yu Zhang, Chung-Cheng Chiu, Barret Zoph, Ekin D. Cubuk, and Quoc V. Le. 2019 · 2019
Cited alongside, same era.
Personalizing ASR for Dysarthric and Accented Speech with Limited Data
Joel Shor, Dotan Emanuel, Oran Lang, Omry Tuval, Michael Brenner, Julie Cattiau, Fernando Vieira, Maeve McNally, Taylor Charbonneau, Melissa Nollstadt, Avinatan Hassidim, and Yossi Matias. 2019 · 2019
Cited alongside, same era.
Transformer transducer: A streamable speech recognition model with transformer encoders and rnn-t loss
Qian Zhang, Han Lu, Hasim Sak, Anshuman Tripathi, Erik McDermott, Stephen Koo, and Shankar Kumar. 2020 · 2020
Later among the works it cites.
Automatic speech recognition of disordered speech: Personalized models now outperforming human listeners on short phrases
Jordan E. Green, Robert L. MacDonald, Pan-Pan Jiang, Julie Cattiau, Rus Heywood, Richard Cave, Katie Seaver, Marilyn A. Ladewig, Jimmy Tobin, Michael P. Brenner, Philip C. Nelson, and Katrin Tomanek. 2021 · 2021
Closest in time.
Disordered speech data collection: Lessons learned at 1 million utterances from project euphonia
Robert L. MacDonald, Pan-Pan Jiang, Julie Cattiau, Rus Heywood, Richard Cave, Katie Seaver, Marilyn Ladewig, Jimmy Tobin, Michael P. Brenner, Philip Q. Nelson, Jordan R. Green, and Katrin Tomanek. 2021 · 2021
Closest in time.
Advancing rnn transducer technology for speech recognition
George Saon, Zoltan Tuske, Daniel Bolanos, and Brian Kingsbury. 2021 · 2021
Closest in time.