Fetching the paper…
Reading the bibliography…
Self-supervised models for speech processing emerged recently as popular foundation blocks in speech processing pipelines.
M. Adda-Decker and L. Lamel, “Do speech recognizers prefer female speakers?” in Interspeech , 2005
2005
Earlier work this paper cites.
A. Graves, S. Fernández et al. , “Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,” in ICML , 2006
2006
Earlier work this paper cites.
Y. Estève, T. Bazillon et al. , “The EPAC Corpus: Manual and Automatic Annotations of Conversational Speech in French Broadcast News,” in LREC , 2010
2010
Earlier work this paper cites.
F. Torreira, M. Adda-Decker, and M. Ernestus, “The Nijmegen Corpus of Casual French,” Speech Communication , 2010
2010
Earlier work this paper cites.
I. Eshkol-Taravella, O. Baude et al. , “Un grand corpus oral ”disponible” : le corpus d’Orléans 1968-2012,” Ressources Linguistiques Libres - Traitement Automatique des Langues , 2011
2011
Earlier work this paper cites.
D. Povey, A. Ghoshal et al. , “The Kaldi speech recognition toolkit,” in IEEE Workshop on automatic speech recognition and understanding , 2011
2011
Earlier work this paper cites.
S. Branca-Rosoff, S. Fleury et al. , “Discours sur la ville. Présentation du Corpus de Français parlé Parisien des années 2000 (CFPP2000),” 2012
2012
Earlier work this paper cites.
T. Bänziger, M. Mortillaro, and K. Scherer, “Introducing the Geneva Multimodal Expression Corpus for Experimental Research on Emotion Perception,” Emotion (Washington, D.C.) , 2012
2012
Earlier work this paper cites.
F. Lefèvre, D. Mostefa et al. , “Robustesse et portabilités multilingue et multi-domaines des systèmes de compréhension de la parole : le projet PortMedia,” in JEP-TALN-RECITAL , 2012
2012
Earlier work this paper cites.
V. Peddinti, D. Povey, and S. Khudanpur, “A time delay neural network architecture for efficient modeling of long temporal contexts,” in Interspeech , 2015
2015
Earlier work this paper cites.
D. P. Kingma and J. Ba, “Adam: A method for stochastic optimization,” in ICLR 2015, Conference Track Proceedings , Y. Bengio and Y. LeCun, Eds., 2015
2015
Earlier work this paper cites.
D. Povey, V. Peddinti et al. , “Purely sequence-trained neural networks for ASR based on lattice-free MMI.” in Interspeech , 2016
2016
Earlier work this paper cites.
R. Tatman, “Gender and dialect bias in YouTube’s automatic captions,” in ACL Workshop on Ethics in NLProc , 2017
2017
Earlier work this paper cites.
R. Tatman and C. Kasten, “Effects of talker dialect, gender & race on accuracy of Bing Speech and YouTube automatic captions,” in Interspeech , 2017
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer et al. , “Attention is all you need,” Advances in neural information processing systems , vol. 30, 2017
2017
Earlier work this paper cites.
P. Gournay, O. Lahaie, and R. Lefebvre, “A canadian french emotional speech dataset,” in ACM Multimedia Systems Conference , 2018
2018
Earlier work this paper cites.
D. Povey, G. Cheng et al. , “Semi-orthogonal low-rank matrix factorization for deep neural networks.” in Interspeech , 2018
2018
Cited alongside, same era.
M. Post, “A call for clarity in reporting BLEU scores,” in Conference on Machine Translation: Research Papers , 2018
2018
Cited alongside, same era.
2019
Cited alongside, same era.
2019
Cited alongside, same era.
M. Garnerin, S. Rossato, and L. Besacier, “Gender representation in French broadcast corpora and its impact on ASR performance,” in International Workshop on AI for Smart TV Content Production, Access and Delivery , 2019
W.-N. Hsu, B. Bolte et al. , “Hubert: Self-supervised speech representation learning by masked prediction of hidden units,” IEEE/ACM Transactions on Audio, Speech, and Language Processing , 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
S. Evain, H. Nguyen et al. , “ LeBenchmark
2021
Later among the works it cites.
——, “Task agnostic and task specific self-supervised learning from speech with LeBenchmark
2021
Later among the works it cites.
S. wen Yang, P.-H. Chi et al. , “SUPERB: Speech Processing Universal PERformance Benchmark,” in Interspeech , 2021
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
M. Ott, S. Edunov et al. , “fairseq: A fast, extensible toolkit for sequence modeling,” in NAACL (Demonstrations) , 2019
2019
Cited alongside, same era.
A. Baevski, Y. Zhou et al. , “wav2vec 2.0: A framework for self-supervised learning of speech representations,” Advances in Neural Information Processing Systems , 2020
2020
Cited alongside, same era.
2020
Cited alongside, same era.
K. Kawakami, L. Wang et al. , “Learning robust and multilingual speech representations,” in EMNLP , 2020
2020
Cited alongside, same era.
V. Pratap, Q. Xu et al. , “MLS: A large-scale multilingual dataset for speech research,” in Interspeech , 2020
2020
Cited alongside, same era.
C. Le Moine and N. Obin, “Att-HACK: An expressive speech database with social attitudes,” in Speech Prosody , 2020
2020
Cited alongside, same era.
ATILF, “TCOF : Traitement de corpus oraux en français,” 2020, https://hdl.handle.net/11403/tcof/v2.1, ORTOLANG (Open Resources and TOols for LANGuage) –www.ortolang.fr
2020
Cited alongside, same era.
E. Salesky, M. Wiesner et al. , “The Multilingual TEDx Corpus for Speech Recognition and Translation,” in Interspeech , 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
——, “Investigating the impact of gender representation in ASR training data: a case study on librispeech,” in ACL Workshop on Gender Bias in Natural Language Processing , 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.
2022
Closest in time.
B. Savoldi, M. Gaido et al. , “Under the morphosyntactic lens: A multifaceted evaluation of gender bias in speech translation,” in ACL , 2022
2022
Closest in time.
M. R. Costa-jussà, C. Basta et al. , “Evaluating gender bias in speech translation,” in LREC , 2022
2022
Closest in time.