Fetching the paper…
Reading the bibliography…
Improving ASR systems is necessary to make new LLM-based use-cases accessible to people across the globe.
1909
Earlier work this paper cites.
C. Chandramouli and R. General, “Census of india 2011,” Provisional Population Totals. New Delhi: Government of India , pp. 409–413, 2011
2011
Earlier work this paper cites.
K. J. Piczak, “Esc: Dataset for environmental sound classification,” in Proceedings of the 23rd ACM international conference on Multimedia , 2015, pp. 1015–1018
2015
Earlier work this paper cites.
A. Baby, A. L. Thomas, N. Nishanthi, T. Consortium et al. , “Resources for indian languages,” in Proceedings of Text, Speech and Dialogue , 2016
2016
Earlier work this paper cites.
K. Sodimana, K. Pipatsrisawat, L. Ha, M. Jansche, O. Kjartansson, P. D. Silva, and S. Sarin, “A Step-by-Step Process for Building TTS Voices Using Open Source Data and Framework for Bangla, Javanese, Khmer, Nepali, Sinhala, and Sundanese,” in Proc. The 6th Intl. Workshop on Spoken Language Technologies for Under-Resourced Languages (SLTU) , Gurugram, India, Aug. 2018, pp. 66–70. [Online]. Available: http://dx.doi.org/10.21437/SLTU.2018-14
2018
Earlier work this paper cites.
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell et al. , “Language models are few-shot learners,” Advances in neural information processing systems , vol. 33, pp. 1877–1901, 2020
2020
Earlier work this paper cites.
B. Abraham, D. Goel, D. Siddarth, K. Bali, M. Chopra, M. Choudhury, P. Joshi, P. Jyothi, S. Sitaram, and V. Seshadri, “Crowdsourcing speech data for low-resource languages from low-income workers,” in Proceedings of The 12th Language Resources and Evaluation Conference, LREC 2020, Marseille, France, May 11-16, 2020 , N. Calzolari, F. Béchet, P. Blache, K. Choukri, C. Cieri, T. Declerck, S. Goggi, H. Isahara, B. Maegaard, J. Mariani, H. Mazo, A. Moreno, J. Odijk, and S. Piperidis, Eds. European Language Resources Association, 2020, pp. 2819–2826. [Online]. Available: https://aclanthology.org/2020.lrec-1.343/
2020
Earlier work this paper cites.
R. Ardila, M. Branson, K. Davis, M. Kohler, J. Meyer, M. Henretty, R. Morais, L. Saunders, F. M. Tyers, and G. Weber, “Common voice: A massively-multilingual speech corpus,” in Proceedings of The 12th Language Resources and Evaluation Conference, LREC 2020, Marseille, France, May 11-16, 2020 , N. Calzolari, F. Béchet, P. Blache, K. Choukri, C. Cieri, T. Declerck, S. Goggi, H. Isahara, B. Maegaard, J. Mariani, H. Mazo, A. Moreno, J. Odijk, and S. Piperidis, Eds. European Language Resources Association, 2020, pp. 4218–4222. [Online]. Available: https://aclanthology.org/2020.lrec-1.520/
2020
Cited alongside, same era.
F. He, S.-H. C. Chu, O. Kjartansson, C. Rivera, A. Katanova, A. Gutkin, I. Demirsahin, C. Johny, M. Jansche, S. Sarin, and K. Pipatsrisawat, “Open-source Multi-speaker Speech Corpora for Building Gujarati, Kannada, Malayalam, Marathi, Tamil and Telugu Speech Synthesis Systems,” in Proceedings of The 12th Language Resources and Evaluation Conference (LREC) . Marseille, France: European Language Resources Association (ELRA), May 2020, pp. 6494–6503. [Online]. Available: https://www.aclweb.org/anthology/2020.lrec-1.800
2020
Cited alongside, same era.
T. Javed, S. Doddapaneni, A. Raman, K. S. Bhogale, G. Ramesh, A. Kunchukuttan, P. Kumar, and M. M. Khapra, “Towards building ASR systems for the next billion users,” in Thirty-Sixth AAAI Conference on Artificial Intelligence, AAAI 2022, Thirty-Fourth Conference on Innovative Applications of Artificial Intelligence, IAAI 2022, The Twelveth Symposium on Educational Advances in Artificial Intelligence, EAAI 2022 Virtual Event, February 22 - March 1, 2022 . AAAI Press, 2022, pp. 10 813–10 821. [Online]. Available: https://ojs.aaai.org/index.php/AAAI/article/view/21327
2022
Later among the works it cites.
2022
Later among the works it cites.
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
N. Srivastava, R. Mukhopadhyay, K. R. Prajwal, and C. V. Jawahar, “Indicspeech: Text-to-speech corpus for indian languages,” in Proceedings of The 12th Language Resources and Evaluation Conference, LREC 2020, Marseille, France, May 11-16, 2020 , N. Calzolari, F. Béchet, P. Blache, K. Choukri, C. Cieri, T. Declerck, S. Goggi, H. Isahara, B. Maegaard, J. Mariani, H. Mazo, A. Moreno, J. Odijk, and S. Piperidis, Eds. European Language Resources Association, 2020, pp. 6417–6422. [Online]. Available: https://aclanthology.org/2020.lrec-1.789/
2020
Cited alongside, same era.
A. Gulati, J. Qin, C. Chiu, N. Parmar, Y. Zhang, J. Yu, W. Han, S. Wang, Z. Zhang, Y. Wu, and R. Pang, “Conformer: Convolution-augmented transformer for speech recognition,” in Interspeech 2020, 21st Annual Conference of the International Speech Communication Association, Virtual Event, Shanghai, China, 25-29 October 2020 , H. Meng, B. Xu, and T. F. Zheng, Eds. ISCA, 2020, pp. 5036–5040. [Online]. Available: https://doi.org/10.21437/Interspeech.2020-3015
2020
Cited alongside, same era.
D. Adiga, R. Kumar, A. Krishna, P. Jyothi, G. Ramakrishnan, and P. Goyal, “Automatic speech recognition in sanskrit: A new speech corpus and modelling insights,” in Findings of the Association for Computational Linguistics: ACL/IJCNLP 2021, Online Event, August 1-6, 2021 , ser. Findings of ACL, C. Zong, F. Xia, W. Li, and R. Navigli, Eds., vol. ACL/IJCNLP 2021. Association for Computational Linguistics, 2021, pp. 5039–5050. [Online]. Available: https://doi.org/10.18653/v1/2021.findings-acl.447
2021
Cited alongside, same era.
A. Diwan, R. Vaideeswaran, S. Shah, A. Singh, S. R. K. M., S. Khare, V. Unni, S. Vyas, A. Rajpuria, C. Yarra, A. R. Mittal, P. K. Ghosh, P. Jyothi, K. Bali, V. Seshadri, S. Sitaram, S. Bharadwaj, J. Nanavati, R. Nanavati, and K. Sankaranarayanan, “MUCS 2021: Multilingual and code-switching ASR challenges for low resource indian languages,” in Interspeech 2021, 22nd Annual Conference of the International Speech Communication Association, Brno, Czechia, 30 August - 3 September 2021 , H. Hermansky, H. Cernocký, L. Burget, L. Lamel, O. Scharenborg, and P. Motlícek, Eds. ISCA, 2021, pp. 2446–2450. [Online]. Available: https://doi.org/10.21437/Interspeech.2021-1339
2021
Cited alongside, same era.
A. Bhanushali, G. Bridgman, D. G, P. K. Ghosh, P. Kumar, S. Kumar, A. R. Kolladath, N. Ravi, A. Seth, A. Seth, A. Singh, V. N. Sukhadia, S. Umesh, S. Udupa, and L. V. S. V. D. Prasad, “Gram vaani ASR challenge on spontaneous telephone speech recordings in regional variations of hindi,” in Interspeech 2022, 23rd Annual Conference of the International Speech Communication Association, Incheon, Korea, 18-22 September 2022 , H. Ko and J. H. L. Hansen, Eds. ISCA, 2022, pp. 3548–3552
2022
Cited alongside, same era.
2022
Later among the works it cites.
2022
Later among the works it cites.
A. Conneau, M. Ma, S. Khanuja, Y. Zhang, V. Axelrod, S. Dalmia, J. Riesa, C. Rivera, and A. Bapna, “FLEURS: few-shot learning evaluation of universal representations of speech,” in IEEE Spoken Language Technology Workshop, SLT 2022, Doha, Qatar, January 9-12, 2023 . IEEE, 2022, pp. 798–805. [Online]. Available: https://doi.org/10.1109/SLT54892.2023.10023141
2023
Closest in time.