Fetching the paper…
Reading the bibliography…
We present a preprocessed, ready-to-use automatic speech recognition corpus, BembaSpeech, consisting over 24 hours of read speech in the Bemba language, a written but low-resourced language spoken by over 30% of the population in Zambia.
Hidden markov models for speech recognition
B. H. Juang and L. R. Rabiner. 1991 · 1991
Earlier work this paper cites.
Long Short-Term Memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Towards speech technology for south african languages: Automatic speech recognition in xhosa
F. de Wet and E. C. Botha. 1999 · 1999
Earlier work this paper cites.
African Languages: An Introduction
B Heine and D Nurse. 2000 · 2000
Earlier work this paper cites.
Facts About the World‘s Languages: An Encyclopedia of the Worlds‘s Major Languages, Past and Present
Debra Spitulnik and Mubanga E Kashoki. 2001 · 2001
Earlier work this paper cites.
An amharic speech corpus for large vocabulary continuous speech recognition
Solomon Teferra Abate, Wolfgang Menzel, and Bairu Tafila. 2005 · 2005
Earlier work this paper cites.
Towards automatic transcription of Somali language
Nimaan Abdillahi, Nocera Pascal, and Bonastre Jean-François. 2006 · 2006
Earlier work this paper cites.
Connectionist temporal classification: Labelling unsegmented sequence data with recurrent neural networks
Alex Graves, Santiago Fernández, Faustino Gomez, and Jürgen Schmidhuber. 2006 · 2006
Earlier work this paper cites.
Language and National Identity in Africa
Nancy C Kula and Lutz Marten. 2008 · 2008
Earlier work this paper cites.
The HTK Book (for HTK Version 3.4)
S. J. Young, G. Evermann, M. J. F. Gales, T. Hain, D. Kershaw, X. Liu, G. Moore, J. Odell, D. Ollason, D. Povey, V. Valtchev, and P. C. Woodland. 2009 · 2009
Earlier work this paper cites.
Rectified linear units improve Restricted Boltzmann machines
Vinod Nair and Geoffrey E. Hinton. 2010 · 2010
Earlier work this paper cites.
Collecting and evaluating speech recognition corpora for 11 South African languages
Jaco Badenhorst, Charl van Heerden, Marelie Davel, and Etienne Barnard. 2011 · 2011
Earlier work this paper cites.
KenLM : Faster and Smaller Language Model Queries
Kenneth Heafield. 2011 · 2011
Earlier work this paper cites.
The Kaldi speech recognition toolkit
Daniel Povey, Arnab Ghoshal, Gilles Boulianne, Lukas Burget, Ondrej Glembek, Nagendra Goel, Mirko Hannemann, Petr Motlıcek, Yanmin Qian, Petr Schwarz, Jan Silovsky, Georg Stemmer, and Karel Vesely. 2011 · 2011
Earlier work this paper cites.
Stochastic Gradient Descent Tricks
Léon Bottou. 2012 · 2012
Cited alongside, same era.
Developments of Swahili resources for an automatic speech recognition system
Hadrien Gelas, Laurent Besacier, and Francois Pellegrino. 2012 · 2012
Cited alongside, same era.
Hausa large vocabulary continuous speech recognition
Tim Schlippe, Edy Guevara Komgang Djomgang, Ngoc Thang Vu, Sebastian Ochs, and Tanja Schultz. 2012 · 2012
Cited alongside, same era.
Baseline Speech Recognition of South African Languages using Lwazi and AST
Daan Henselmans, Thomas Niesler, and David Van Leeuwen. 2013 · 2013
Cited alongside, same era.
Deep Speech: Scaling up end-to-end speech recognition
Awni Hannun, Carl Case, Jared Casper, Bryan Catanzaro, Greg Diamos, Erich Elsen, Ryan Prenger, Sanjeev Satheesh, Shubho Sengupta, Adam Coates, and Andrew Y. Ng. 2014 · 2014
Cited alongside, same era.
First automatic fongbe continuous speech recognition system: Development of acoustic models and language models
Frejus A.A. Laleye, Laurent Besacier, Eugene C. Ezin, and Cina Motamed. 2016 · 2016
Later among the works it cites.
Improving the Lwazi ASR baseline
Charl Van Heerden, Neil Kleynhans, and Marelie Davel. 2016 · 2016
Later among the works it cites.
Speech recognition for under-resourced languages: Data sharing in hidden Markov model systems
Febe De Wet, Neil Kleynhans, Dirk Van Compernolle, and Reza Sahraeian. 2017 · 2017
Later among the works it cites.
Theorectical Reflections on the Teaching of Literacy in Zambian Bantu Languages
Joseph M Mwansa. 2017 · 2017
Later among the works it cites.
An Introduction to Zambia’s Bemba Tribe
Mazuba Kapambwe. 2018 · 2018
Later among the works it cites.
WAV2LETTER++: THE FASTEST OPEN-SOURCE SPEECH RECOGNITION SYSTEM
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Vidali D Spitulnik and Mubanga E Kashoki. 2014 · 2014
Cited alongside, same era.
Bemba Phonology
Vidali D Spitulnik and Mubanga E. Kashoki. 2014 · 2014
Cited alongside, same era.
Using different acoustic, lexical and language modeling units for ASR of an under-resourced language - Amharic
Martha Yifiru Tachbelie and Laurent Besacier. 2014 · 2014
Cited alongside, same era.
Librispeech: An ASR corpus based on public domain audio books
Vassil Panayotov, Guoguo Chen, Daniel Povey, and Sanjeev Khudanpur. 2015 · 2015
Cited alongside, same era.
Deep speech 2: End-to-end speech recognition in English and Mandarin
Dario Amodei, Sundaram Ananthanarayanan, Rishita Anubhai, Jingliang Bai, Eric Battenberg, Carl Case, Jared Casper, Bryan Catanzaro, Qiang Cheng, Guoliang Chen, Jie Chen, Jingdong Chen, Zhijie Chen, Mike Chrzanowski, Adam Coates, Greg Diamos, Ke Ding, Niandong Du, Erich Elsen, Jesse Engel, Weiwei Fang, Linxi Fan, Christopher Fougner, Liang Gao, Caixia Gong, Aw Ni Hannun, Tony Han, Lappi Vaino Johannes, Bing Jiang, Cai Ju, Billy Jun, Patrick Legresley, Libby Lin, Junjie Liu, Yang Liu, Weigao Li, Xiangang Li, Dongpeng Ma, Sharan Narang, Andrew Ng, Sherjil Ozair, Yiping Peng, Ryan Prenger, Sheng Qian, Zongfeng Quan, Jonathan Raiman, Vinay Rao, Sanjeev Satheesh, David Seetapun, Shubho Sengupta, Kavya Srinet, Anuroop Sriram, Haiyuan Tang, Liliang Tang, Chong Wang, Jidong Wang, Kaifu Wang, Yi Wang, Zhijian Wang, Zhiqian Wang, Shuang Wu, Likai Wei, Bo Xiao, Wen Xie, Yan Xie, Dani Yogatama, Bin Yuan, Jun Zhan, and Zhenyao Zhu. 2016 · 2016
Cited alongside, same era.
Parallel Speech Collection for Under-resourced Language Studies Using the Lig-Aikuma Mobile Device App
David Blachon, Elodie Gauthier, Laurent Besacier, Guy Noël Kouarata, Martine Adda-Decker, and Annie Rialland. 2016 · 2016
Cited alongside, same era.
Collecting resources in sub-Saharan African languages for automatic speech recognition: A case study of Wolof
Elodie Gauthier, Laurent Besacier, Sylvie Voisin, Michael Melese, and Uriel Pascal Elingui. 2016b · 2016
Cited alongside, same era.
Vineel Pratap, Awni Hannun, Qiantong Xu, Jeff Cai, Jacob Kahn, Gabriel Synnaeve, Vitaliy Liptchinsky, and Ronan Collobert. 2018 · 2018
Later among the works it cites.
JW300: A wide-coverage parallel corpus for low-resource languages
Željko Agic and Ivan Vulic. 2020 · 2019
Later among the works it cites.
Benchmarking Neural Machine Translation for Southern African Languages
Laura Martinus and Jade Z. Abbott. 2019 · 2019
Later among the works it cites.
The Pytorch-kaldi Speech Recognition Toolkit
Mirco Ravanelli, Titouan Parcollet, and Yoshua Bengio. 2019 · 2019
Later among the works it cites.
Large vocabulary read speech corpora for four ethiopian languages: Amharic, Tigrigna, Oromo and Wolaytta
Solomon Teferra Abate, Martha Yifiru Tachbelie, Michael Melese, Hafte Abera, Tewodros Abebe, Wondwossen Mulugeta, Yaregal Assabie, Million Meshesha, Solomon Atinafu, and Binyam Ephrem. 2020 · 2020
Later among the works it cites.
Improving the Language Model for Low-Resource {ASR} with Online Text Corpora
Nils Hjortnaes, Timofey Arkhangelskiy, Niko Partanen, Michael Rießler, and Francis Tyers. 2020 · 2020
Later among the works it cites.
Multi-Task and Transfer Learning in Low-Resource Speech Recognition
Josh Meyer. 2020 · 2020
Later among the works it cites.
Transfer Learning for Less-Resourced {S}emitic Languages Speech Recognition: the Case of {A}mharic
Yonas Woldemariam. 2020 · 2020
Later among the works it cites.