Fetching the paper…
Reading the bibliography…
English is the most widely spoken language in the world, used daily by millions of people as a first or second language in many different contexts.
H. Giles, N. Coupland, and J. Coupland, Accommodation theory: Communication, context, and consequence . Cambridge University Press, 1991
1991
Earlier work this paper cites.
J. J. Godfrey, E. C. Holliman, and J. McDaniel, “Switchboard: Telephone speech corpus for research and development,” in ICASSP , 1992
1992
Earlier work this paper cites.
D. B. Paul and J. Baker, “The design for the wall street journal-based csr corpus,” in Speech and Natural Language: Proceedings of a Workshop Held at Harriman , 1992
1992
Earlier work this paper cites.
J. S. Garofolo, L. F. Lamel, W. M. Fisher, J. G. Fiscus, and D. S. Pallett, “Darpa timit acoustic-phonetic continous speech corpus cd-rom,” NASA STI , 1993
1993
Earlier work this paper cites.
L. Bauer, An Introduction to International Varieties of English . Edinburgh University Press, 2003
2003
Earlier work this paper cites.
I. McCowan, J. Carletta, W. Kraaij, S. Ashby, S. Bourban, M. Flynn, M. Guillemot, T. Hain, J. Kadlec, V. Karaiskos et al. , “The ami meeting corpus,” in International Conference on Methods and Techniques in Behavioral Research , 2005
2005
Earlier work this paper cites.
L. Campbell, “Ethnologue: Languages of the world,” 2008
2008
Earlier work this paper cites.
D. Povey, A. Ghoshal, G. Boulianne, L. Burget, O. Glembek, N. Goel, M. Hannemann, P. Motlicek, Y. Qian, P. Schwarz, J. Silovsky, G. Stemmer, and K. Vesely, “The kaldi speech recognition toolkit,” in ASRU , Dec. 2011
2011
Earlier work this paper cites.
K. Heafield, “KenLM: Faster and smaller language model queries,” in SMT , 2011
2011
Earlier work this paper cites.
N. Schilling, Sociolinguistic Fieldwork , ser. Key Topics in Sociolinguistics. Cambridge University Press, 2013
2013
Earlier work this paper cites.
V. Panayotov, G. Chen, D. Povey, and S. Khudanpur, “Librispeech: an asr corpus based on public domain audio books,” in ICASSP , 2015
2015
Earlier work this paper cites.
S. Weinberger, “The speech accent archive,” online, 2015. [Online]. Available: http://accent.gmu.edu
2015
Earlier work this paper cites.
P. Bell, M. J. Gales, T. Hain, J. Kilgour, P. Lanchantin, X. Liu, A. McParland, S. Renals, O. Saz, M. Wester et al. , “The mgb challenge: Evaluating multi-genre broadcast media recognition,” in ASRU , 2015
2015
Cited alongside, same era.
G. Van Herk, What is sociolinguistics . John Wiley & Sons, Inc, 2018
2018
Cited alongside, same era.
G. Zhao, S. Sonsaat, A. Silpachai, I. Lucic, E. Chukharev-Hudilainen, J. Levis, and R. Gutierrez-Osuna, “L2-arctic: A non-native english speech corpus.” in Interspeech , 2018
2018
Cited alongside, same era.
R. Ardila, M. Branson, K. Davis, M. Kohler, J. Meyer, M. Henretty, R. Morais, L. Saunders, F. Tyers, and G. Weber, “Common voice: A massively-multilingual speech corpus,” in EACL , 2019
2019
Cited alongside, same era.
A. Koenecke, A. Nam, E. Lake, J. Nudell, M. Quartey, Z. Mengesha, C. Toups, J. R. Rickford, D. Jurafsky, and S. Goel, “Racial disparities in automated speech recognition,” Proceedings of the National Academy of Sciences , vol. 117, 2020
Z.-H. Tan, N. Dehak et al. , “rvad: An unsupervised segment-based robust voice activity detection method,” Computer speech & language , 2020
2020
Later among the works it cites.
Z. Tüske, G. Saon, and B. Kingsbury, “On the limit of english conversational speech recognition,” in Interspeech , 2021
2021
Later among the works it cites.
W.-N. Hsu, A. Sriram, A. Baevski, T. Likhomanenko, Q. Xu, V. Pratap, J. Kahn, A. Lee, R. Collobert, G. Synnaeve, and M. Auli, “Robust wav2vec 2.0: Analyzing Domain Shift in Self-Supervised Pre-Training,” in Interspeech , 2021
2021
Later among the works it cites.
E. M. Bender, T. Gebru, A. McMillan-Major, and S. Shmitchell, “On the dangers of stochastic parrots: Can language models be too big,” in Conference on Fairness, Accountability, and Transparency (ACM) , 2021
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2020
Cited alongside, same era.
S. L. Blodgett, S. Barocas, H. Daumé III, and H. Wallach, “Language (technology) is power: A critical survey of “bias” in NLP,” in ACL , 2020
2020
Cited alongside, same era.
A. Baevski, Y. Zhou, A. Mohamed, and M. Auli, “wav2vec 2.0: A framework for self-supervised learning of speech representations,” NeurIPS , 2020
2020
Cited alongside, same era.
J. Kahn, M. Rivière, W. Zheng, E. Kharitonov, Q. Xu, P.-E. Mazaré, J. Karadayi, V. Liptchinsky, R. Collobert, C. Fuegen et al. , “Libri-light: A benchmark for asr with limited or no supervision,” in ICASSP , 2020
2020
Cited alongside, same era.
I. Demirsahin, O. Kjartansson, A. Gutkin, and C. Rivera, “Open-source multi-speaker corpora of the english accents in the british isles,” in LREC , 2020
2020
Cited alongside, same era.
J. L. Martin and K. Tang, “Understanding racial disparities in automatic speech recognition: The case of habitual “be”,” in Proc. Interspeech 2020 , 2020, pp. 626–630. [Online]. Available: http://dx.doi.org/10.21437/Interspeech.2020-2893
2020
Cited alongside, same era.
J. Meyer, L. Rauchenstein, J. D. Eisenberg, and N. Howell, “Artie bias corpus: An open dataset for detecting demographic bias in speech applications,” in Proceedings of the 12th language resources and evaluation conference , 2020
2020
Cited alongside, same era.
X. Shi, F. Yu, Y. Lu, Y. Liang, Q. Feng, D. Wang, Y. Qian, and L. Xie, “The accented english speech recognition challenge 2020: open datasets, tracks, baselines, results and methods,” in ICASSP , 2021
2021
Later among the works it cites.
M. Ravanelli, T. Parcollet, P. Plantinga, A. Rouhe, S. Cornell, L. Lugosch, C. Subakan, N. Dawalatabad, A. Heba, J. Zhong, J.-C. Chou, S.-L. Yeh, S.-W. Fu, C.-F. Liao, E. Rastorgueva, F. Grondin, W. Aris, H. Na, Y. Gao, R. D. Mori, and Y. Bengio, “Speechbrain: A general-purpose speech toolkit,” 2021
2021
Later among the works it cites.
N. Dawalatabad, M. Ravanelli, F. Grondin, J. Thienpondt, B. Desplanques, and H. Na, “ECAPA-TDNN Embeddings for Speaker Diarization,” in Interspeech , 2021
2021
Later among the works it cites.
N. Markl, “Language variation and algorithmic bias: Understanding algorithmic bias in british english automatic speech recognition,” in Conference on Fairness, Accountability, and Transparency , 2022
2022
Later among the works it cites.
G.-T. Lin, C.-J. Hsu, D.-R. Liu, H.-Y. Lee, and Y. Tsao, “Analyzing the robustness of unsupervised speech recognition,” in ICASSP , 2022
2022
Later among the works it cites.
N. Markl, “Mind the data gap(s): Investigating power in speech and language datasets,” Second Workshop on Language Technology for Equality, Diversity and Inclusion (ACL) , 2022
2022
Later among the works it cites.
A. Radford, J. W. Kim, T. Xu, G. Brockman, C. McLeavey, and I. Sutskever, “Robust speech recognition via large-scale weak supervision,” OpenAI, Tech. Rep., 2022
2022
Later among the works it cites.