Fetching the paper…
Reading the bibliography…
A cornerstone in AI research has been the creation and adoption of standardized training and test datasets to earmark the progress of state-of-the-art models.
XTREME: A Massively Multilingual Multi-task Benchmark for Evaluating Cross-lingual Generalization
Hu, J.; Ruder, S.; Siddhant, A.; Neubig, G.; Firat, O.; and Johnson, M. 2020 · 2003
Earlier work this paper cites.
Computing and visualizing dynamic time warping alignments in R: the dtw package
Giorgino, T. 2009 · 2009
Earlier work this paper cites.
The IIIT-H Indic speech databases
Prahallad, K.; Kumar, E. N.; Keri, V.; Rajendran, S.; and Black, A. W. 2012 · 2012
Earlier work this paper cites.
Librispeech: an asr corpus based on public domain audio books
Panayotov, V.; Chen, G.; Povey, D.; and Khudanpur, S. 2015 · 2015
Earlier work this paper cites.
Voxceleb: a large-scale speaker identification dataset
Nagrani, A.; Chung, J. S.; and Zisserman, A. 2017 · 2017
Earlier work this paper cites.
Voxceleb2: Deep speaker recognition
Chung, J. S.; Nagrani, A.; and Zisserman, A. 2018 · 2018
Earlier work this paper cites.
X-Vectors: Robust DNN Embeddings for Speaker Recognition
Snyder, D.; Garcia-Romero, D.; Sell, G.; Povey, D.; and Khudanpur, S. 2018 · 2018
Earlier work this paper cites.
Interspeech 2018 Low Resource Automatic Speech Recognition Challenge for Indian Languages
Srivastava, B. M. L.; Sitaram, S.; Kumar Mehta, R.; Doss Mohan, K.; Matani, P.; Satpal, S.; Bali, K.; Srikanth, R.; and Nayak, N. 2018 · 2018
Earlier work this paper cites.
GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding
Wang, A.; Singh, A.; Michael, J.; Hill, F.; Levy, O.; and Bowman, S. R. 2018 · 2018
Earlier work this paper cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Devlin, J.; Chang, M.; Lee, K.; and Toutanova, K. 2019 · 2019
Cited alongside, same era.
Common Voice: A Massively-Multilingual Speech Corpus
Ardila, R.; Branson, M.; Davis, K.; Henretty, M.; Kohler, M.; Meyer, J.; Morais, R.; Saunders, L.; Tyers, F. M.; and Weber, G. 2020 · 2020
Cited alongside, same era.
Open-source Multi-speaker Speech Corpora for Building Gujarati, Kannada, Malayalam, Marathi, Tamil and Telugu Speech Synthesis Systems
He, F.; Chu, S.-H. C.; Kjartansson, O.; Rivera, C.; Katanova, A.; Gutkin, A.; Demirsahin, I.; Johny, C.; Jansche, M.; Sarin, S.; and Pipatsrisawat, K. 2020 · 2020
Cited alongside, same era.
The State and Fate of Linguistic Diversity and Inclusion in the NLP World
Joshi, P.; Santy, S.; Budhiraja, A.; Bali, K.; and Choudhury, M. 2020 · 2020
Cited alongside, same era.
IndicNLPSuite: Monolingual Corpora, Evaluation Benchmarks and Pre-trained Multilingual Language Models for Indian Languages
Kakwani, D.; Kunchukuttan, A.; Golla, S.; N.C., G.; Bhattacharyya, A.; Khapra, M. M.; and Kumar, P. 2020 · 2020
CSTD-Telugu Corpus: Crowd-Sourced Approach for Large-Scale Speech data collection
Ganesh, M. S.; Vegesna, V. V. R.; Naroju, M. D.; Maity, S.; Yalla, P.; and Vuppala, A. K. 2021 · 2021
Later among the works it cites.
Exploring the use of Common Label Set to Improve Speech Recognition of Low Resource Indian Languages
Shetty, V. M.; and Umesh, S. 2021 · 2021
Later among the works it cites.
SUPERB: Speech Processing Universal PERformance Benchmark
Yang, S.; Chi, P.; Chuang, Y.; Lai, C. J.; Lakhotia, K.; Lin, Y. Y.; Liu, A. T.; Shi, J.; Chang, X.; Lin, G.; Huang, T.; Tseng, W.; Lee, K.; Liu, D.; Huang, Z.; Dong, S.; Li, S.; Watanabe, S.; Mohamed, A.; and Lee, H. 2021a · 2021
Later among the works it cites.
Subword Dictionary Learning and Segmentation Techniques for Automatic Speech Recognition in Tamil and Kannada
A, M.; Pilar, B.; and G, R. A. 2022 · 2022
Closest in time.
Gram Vaani ASR Challenge on spontaneous telephone speech recordings in regional variations of Hindi
Bhanushali, A.; Bridgman, G.; G, D.; Ghosh, P.; Kumar, P.; Kumar, S.; Raj Kolladath, A.; Ravi, N.; Seth, A.; Seth, A.; Singh, A.; Sukhadia, V.; S, U.; Udupa, S.; and Prasad, L. V. S. V. D. 2022 · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
How Useful Is Self-Supervised Pretraining for Visual Tasks?
Newell, A.; and Deng, J. 2020 · 2020
Cited alongside, same era.
Automatic Speech Recognition in Sanskrit: A New Speech Corpus and Modelling Insights
Adiga, D.; Kumar, R.; Krishna, A.; Jyothi, P.; Ramakrishnan, G.; and Goyal, P. 2021 · 2021
Cited alongside, same era.
XLS-R: Self-supervised cross-lingual speech representation learning at scale
Babu, A.; Wang, C.; Tjandra, A.; Lakhotia, K.; Xu, Q.; Goyal, N.; Singh, K.; von Platen, P.; Saraf, Y.; Pino, J.; et al. 2021 · 2021
Cited alongside, same era.
Multilingual and code-switching ASR challenges for low resource Indian languages
Diwan, A.; Vaideeswaran, R.; Shah, S.; Singh, A.; Raghavan, S.; Khare, S.; Unni, V.; Vyas, S.; Rajpuria, A.; Yarra, C.; Mittal, A.; Ghosh, P. K.; Jyothi, P.; Bali, K.; Seshadri, V.; Sitaram, S.; Bharadwaj, S.; Nanavati, J.; Nanavati, R.; Sankaranarayanan, K.; Seeram, T.; and Abraham, B. 2021 · 2021
Cited alongside, same era.
Crowdsourcing Speech Data for Low-Resource Languages from Low-Income Workers
Abraham, B.; Goel, D.; Siddarth, D.; Bali, K.; Chopra, M.; Choudhury, M.; Joshi, P.; Jyoti, P.; Sitaram, S.; and Seshadri, V. 2020a
Cited in the paper.
Crowdsourcing Speech Data for Low-Resource Languages from Low-Income Workers
Abraham, B.; Goel, D.; Siddarth, D.; Bali, K.; Chopra, M.; Choudhury, M.; Joshi, P.; Jyoti, P.; Sitaram, S.; and Seshadri, V. 2020b
Cited in the paper.
Crowd-Sourced Speech Corpora for Javanese, Sundanese, Sinhala, Nepali, and Bangladeshi Bengali
Kjartansson, O.; Sarin, S.; Pipatsrisawat, K.; Jansche, M.; and Ha, L. 2018a
Cited in the paper.
Closest in time.
Towards building asr systems for the next billion users
Javed, T.; Doddapaneni, S.; Raman, A.; Bhogale, K. S.; Ramesh, G.; Kunchukuttan, A.; Kumar, P.; and Khapra, M. M. 2022 · 2022
Closest in time.
No Language Left Behind: Scaling Human-Centered Machine Translation
NLLB Team; Costa-jussà, M. R.; Cross, J.; Çelebi, O.; Elbayad, M.; Heafield, K.; Heffernan, K.; Kalbassi, E.; Lam, J.; Licht, D.; Maillard, J.; Sun, A.; Wang, S.; Wenzek, G.; Youngblood, A.; Akula, B.; Barrault, L.; Gonzalez, G. M.; Hansanti, P.; Hoffman, J.; Jarrett, S.; Sadagopan, K. R.; Rowe, D.; Spruit, S.; Tran, C.; Andrews, P.; Ayan, N. F.; Bhosale, S.; Edunov, S.; Fan, A.; Gao, C.; Goswami, V.; Guzmán, F.; Koehn, P.; Mourachko, A.; Ropers, C.; Saleem, S.; Schwenk, H.; and Wang, J. 2022 · 2022
Closest in time.
Tsai, H.-S.; Chang, H.-J.; Huang, W.-C.; Huang, Z.; Lakhotia, K.; Yang, S.-w.; Dong, S.; Liu, A. T.; Lai, C.-I. J.; Shi, J.; et al. 2022 · 2022
Closest in time.