Fetching the paper…
Reading the bibliography…
Emotion recognition datasets are relatively small, making the use of the more sophisticated deep learning approaches challenging.
B. Schuller, G. Rigoll, and M. Lang, Hidden Markov Model-based Speech Emotion Recognition , 2003, vol. 2
2003
Earlier work this paper cites.
F. Burkhardt, A. Paeschke, M. Rolfes, W. F. Sendlmeier, and B. Weiss, “A database of german emotional speech,” in Ninth European Conference on Speech Communication and Technology , 2005
2005
Earlier work this paper cites.
M. Borchert and A. Dusterhoft, “Emotions in speech-experiments with prosody and quality features in speech for use in categorical and dimensional emotion recognition environments,” in International Conference on Natural Language Processing and Knowledge Engineering . IEEE, 2005
2005
Earlier work this paper cites.
A. Graves, S. Fernández, F. Gomez, and J. Schmidhuber, “Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,” in ICML , 2006
2006
Earlier work this paper cites.
N. Sato and Y. Obuchi, “Emotion recognition using mel-frequency cepstral coefficients,” Information and Media Technologies , vol. 2, no. 3, 2007
2007
Earlier work this paper cites.
C. Busso, M. Bulut, C.-C. Lee, A. Kazemzadeh, E. Mower, S. Kim, J. N. Chang, S. Lee, and S. S. Narayanan, “Iemocap: Interactive emotional dyadic motion capture database,” Language resources and evaluation , vol. 42, no. 4, 2008
2008
Earlier work this paper cites.
S. Brave and C. Nass, “Emotion in human–computer interaction,” in Human-computer interaction fundamentals . CRC Press Boca Raton, FL, USA, 2009, vol. 20094635
2009
Earlier work this paper cites.
S. Haq and P. Jackson, “Machine Audition: Principles, Algorithms and Systems,” W. Wang, Ed. IGI Global, 2010
2010
Earlier work this paper cites.
F. Eyben, M. Wöllmer, and B. Schuller, “Opensmile: the munich versatile and fast open-source audio feature extractor,” in Proceedings of the 18th ACM international conference on Multimedia , 2010
2010
Earlier work this paper cites.
P. Shen, Z. Changjun, and X. Chen, “Automatic speech emotion recognition using support vector machine,” in Proceedings of 2011 International Conference on Electronic & Mechanical Engineering and Information Technology , vol. 2. IEEE, 2011
2011
Earlier work this paper cites.
C.-C. Lee, E. Mower, C. Busso, S. Lee, and S. Narayanan, “Emotion recognition using a hierarchical binary decision tree approach,” Speech Communication , vol. 53, no. 9-10, 2011
2011
Earlier work this paper cites.
K. S. Rao, S. G. Koolagudi, and R. R. Vempada, “Emotion recognition from speech using global and local prosodic features,” International journal of speech technology , vol. 16, no. 2, 2013
2013
Earlier work this paper cites.
V. Panayotov, G. Chen, D. Povey, and S. Khudanpur, “Librispeech: An asr corpus based on public domain audio books,” in 2015 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , 2015
2015
Earlier work this paper cites.
F. Eyben, K. R. Scherer, B. W. Schuller, J. Sundberg, E. André, C. Busso, L. Y. Devillers, J. Epps, P. Laukka, S. S. Narayanan et al. , “The geneva minimalistic acoustic parameter set (gemaps) for voice research and affective computing,” IEEE transactions on affective computing , vol. 7, no. 2, 2015
2015
Cited alongside, same era.
P. Matějka, O. Glembek, O. Novotnỳ, O. Plchot, F. Grézl, L. Burget, and J. H. Cernockỳ, “Analysis of dnn approaches to speaker identification,” in ICASSP . IEEE, 2016
2016
Cited alongside, same era.
2016
Cited alongside, same era.
S. Mirsamadi, E. Barsoum, and C. Zhang, “Automatic speech emotion recognition using recurrent neural networks with local attention,” in 2017 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2017
2017
2019
Later among the works it cites.
P. Riera, L. Ferrer, A. Gravano, and L. Gauder, “No sample left behind: Towards a comprehensive evaluation of speech emotion recognition system,” in Proc. Workshop on Speech, Music and Mind 2019 , 2019
2019
Later among the works it cites.
2020
Later among the works it cites.
A. Baevski, S. Schneider, and M. Auli, “vq-wav2vec: Self-supervised learning of discrete speech representations,” in International Conference on Learning Representations , 2020
2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
A. Satt, S. Rozenberg, and R. Hoory, “Efficient emotion recognition from speech using deep learning on spectrograms.” in Interspeech , 2017
2017
Cited alongside, same era.
2017
Cited alongside, same era.
H. M. Fayek, M. Lech, and L. Cavedon, “Evaluating deep learning architectures for Speech Emotion Recognition,” Neural Networks , vol. 92, 2017
2017
Cited alongside, same era.
S. R. Livingstone and F. A. Russo, “The Ryerson Audio-Visual Database of Emotional Speech and Song (RAVDESS): A dynamic, multimodal set of facial and vocal expressions in North American English,” PLOS ONE , vol. 13, no. 5, 2018
2018
Cited alongside, same era.
M. Sarma, P. Ghahremani, D. Povey, N. K. Goel, K. K. Sarma, and N. Dehak, “Emotion identification from raw speech signals using dnns.” in Interspeech , 2018
2018
Cited alongside, same era.
2018
Cited alongside, same era.
M. Peters, M. Neumann, M. Iyyer, M. Gardner, C. Clark, K. Lee, and L. Zettlemoyer, “Deep contextualized word representations,” in NAACL , 2018
2018
Cited alongside, same era.
R. Lotfian and C. Busso, “Building naturalistic emotionally balanced speech corpus by retrieving emotional speech from existing podcast recordings,” IEEE Transactions on Affective Computing , vol. 10, no. 4, 2019
2019
Cited alongside, same era.
2020
Later among the works it cites.
2020
Later among the works it cites.
2020
Later among the works it cites.
2020
Later among the works it cites.
D. Issa, M. Fatih Demirci, and A. Yazici, “Speech emotion recognition with deep convolutional neural networks,” Biomedical Signal Processing and Control , vol. 59, 2020
2020
Later among the works it cites.
2021
Closest in time.
2021
Closest in time.