Fetching the paper…
Reading the bibliography…
Modern automatic speech recognition (ASR) systems have achieved superhuman Word Error Rate (WER) on many common corpora despite lacking adequate performance on speech in the wild.
J. C. Wells, Accents of English: Volume 3: Beyond the British Isles . Cambridge University Press, 1982
1982
Earlier work this paper cites.
P. I. Good, Permutation, Parametric, and Bootstrap Tests of Hypotheses (Springer Series in Statistics) . Berlin, Heidelberg: Springer-Verlag, 2004
2004
Earlier work this paper cites.
B. R. Chiswick and P. W. Miller, “Linguistic distance: A quantitative measure of the distance between english and other languages,” Journal of Multilingual and Multicultural Development , vol. 26, no. 1, pp. 1–11, 2005
2005
Earlier work this paper cites.
D. A. Pharies, A Brief History of the Spanish Language . University Of Chicago Press, 2007
2007
Earlier work this paper cites.
S. Goldwater, D. Jurafsky, and C. D. Manning, “Which words are hard to recognize? prosodic, lexical, and disfluency factors that increase speech recognition error rates,” Speech Communication , vol. 52, no. 3, pp. 181–200, 2010
2010
Cited alongside, same era.
A. Ralli, Greek in Contact With Romance Greek in Contact With Romance , 10 2020
2020
Cited alongside, same era.
B. van Rooy, English in Africa , ser. Cambridge Handbooks in Language and Linguistics. Cambridge University Press, 2020, p. 210–235
2020
Cited alongside, same era.
P. K. O’Neill, V. Lavrukhin, S. Majumdar, V. Noroozi, Y. Zhang, O. Kuchaiev, J. Balam, Y. Dovzhenko, K. Freyberg, M. D. Shulman, B. Ginsburg, S. Watanabe, and G. Kucsko, “SPGISpeech: 5,000 Hours of Transcribed Financial Audio for Fully Formatted End-to-End Speech Recognition,” in Proc. Interspeech 2021 , 2021, pp. 1434–1438
2021
Cited alongside, same era.
M. Del Rio, N. Delworth, R. Westerman, M. Huang, N. Bhandari, J. Palakapilly, Q. McNamara, J. Dong, P. Żelasko, and M. Jetté, “Earnings-21: A Practical Benchmark for ASR in the Wild,” in Proc. Interspeech 2021 , 2021, pp. 3465–3469
2021
Later among the works it cites.
C. Miller, E. Tzoukermann, J. Doyon, and E. Mallard, “Corpus creation and evaluation for speech-to-text and speech translation,” in Proceedings of Machine Translation Summit XVIII: Users and Providers Track , 2021, pp. 44–53
2021
Later among the works it cites.
L. Kosmala and L. Crible, “The dual status of filled pauses: Evidence from genre, proficiency and co-occurrence,” Language and Speech , May 2021. [Online]. Available: https://halshs.archives-ouvertes.fr/halshs-03225622
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…