Fetching the paper…
Reading the bibliography…
Despite recent advancements in speech emotion recognition (SER) models, state-of-the-art deep learning (DL) approaches face the challenge of the limited availability of annotated data.
S. Yildirim, M. Bulut, C. M. Lee, A. Kazemzadeh, Z. Deng, S. Lee, S. Narayanan, and C. Busso, “An acoustic study of emotions expressed in speech,” in Eighth International Conference on Spoken Language Processing , 2004
2004
Earlier work this paper cites.
C. Busso, M. Bulut, C.-C. Lee, A. Kazemzadeh, E. Mower, S. Kim, J. N. Chang, S. Lee, and S. S. Narayanan, “Iemocap: Interactive emotional dyadic motion capture database,” Language resources and evaluation , vol. 42, no. 4, p. 335, 2008
2008
Earlier work this paper cites.
P. J. Fraccaro, B. C. Jones, J. Vukovic, F. G. Smith, C. D. Watkins, D. R. Feinberg, A. C. Little, and L. M. Debruine, “Experimental evidence that women speak in a higher voice pitch to men they find attractive,” Journal of Evolutionary Psychology , vol. 9, no. 1, pp. 57–67, 2011
2011
Earlier work this paper cites.
J. Pustejovsky and A. Stubbs, Natural Language Annotation for Machine Learning: A guide to corpus-building for applications . ” O’Reilly Media, Inc.”, 2012
2012
Earlier work this paper cites.
2013
Earlier work this paper cites.
V. Panayotov, G. Chen, D. Povey, and S. Khudanpur, “Librispeech: an asr corpus based on public domain audio books,” in 2015 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2015, pp. 5206–5210
2015
Earlier work this paper cites.
G. Trigeorgis, F. Ringeval, R. Brueckner, E. Marchi, M. A. Nicolaou, B. Schuller, and S. Zafeiriou, “Adieu features? end-to-end speech emotion recognition using a deep convolutional recurrent network,” in 2016 IEEE international conference on acoustics, speech and signal processing (ICASSP) . IEEE, 2016, pp. 5200–5204
2016
Earlier work this paper cites.
R. Lotfian and C. Busso, “Retrieving categorical emotions using a probabilistic framework to define preference learning samples,” in Interspeech 2016 , 2016, pp. 490–494
2016
Earlier work this paper cites.
Y. Kim and E. M. Provost, “Emotion spotting: Discovering regions of evidence in audio-visual emotion expressions,” in Proceedings of the 18th ACM International Conference on Multimodal Interaction . ACM, 2016, pp. 92–99
2016
Earlier work this paper cites.
A. Burmania, S. Parthasarathy, and C. Busso, “Increasing the reliability of crowdsourcing evaluations using online quality assessment,” IEEE Transactions on Affective Computing , vol. 7, no. 4, pp. 374–388, 2016
2016
Earlier work this paper cites.
C. Cioffi-Revilla and C. Cioffi-Revilla, “Computation and social science,” Introduction to computational social science: Principles and applications , pp. 35–102, 2017
2017
Earlier work this paper cites.
C. Busso, S. Parthasarathy, A. Burmania, M. AbdelWahab, N. Sadoughi, and E. M. Provost, “Msp-improv: An acted corpus of dyadic interactions to study emotion perception,” IEEE Transactions on Affective Computing , vol. 8, no. 1, pp. 67–80, 2017
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
A. van den Oord, O. Vinyals, and K. Kavukcuoglu, “Neural discrete representation learning,” 2018
2018
Earlier work this paper cites.
A. Qayyum, S. Latif, and J. Qadir, “Quran reciter identification: A deep learning approach,” in 2018 7th International Conference on Computer and Communication Engineering (ICCCE) . IEEE, 2018, pp. 492–497
2018
Earlier work this paper cites.
S. Latif, R. Rana, J. Qadir, and J. Epps, “Variational autoencoders for learning latent representations of speech emotion: A preliminary study,” Proc. Interspeech 2018 , pp. 3107–3111, 2018
2018
Earlier work this paper cites.
S. Sahu, R. Gupta, and C. Espy-Wilson, “On enhancing speech emotion recognition using generative adversarial networks,” Proc. Interspeech 2018 , pp. 3693–3697, 2018
2018
Earlier work this paper cites.
S. Latif, A. Qayyum, M. Usama, J. Qadir, A. Zwitter, and M. Shahzad, “Caveat emptor: the risks of using big data for human development,” Ieee technology and society magazine , vol. 38, no. 3, pp. 82–90, 2019
2019
Earlier work this paper cites.
X. Liao and Z. Zhao, “Unsupervised approaches for textual semantic annotation, a survey,” ACM Computing Surveys (CSUR) , vol. 52, no. 4, pp. 1–45, 2019
2019
Earlier work this paper cites.
S. Ding and R. Gutierrez-Osuna, “Group latent embedding for vector quantized variational autoencoder in non-parallel voice conversion.” in INTERSPEECH , 2019, pp. 724–728
2019
Earlier work this paper cites.
2019
Cited alongside, same era.
S. Latif, R. Rana, S. Khalifa, R. Jurdak, and J. Epps, “Direct Modelling of Speech Emotion from Raw Speech,” in Proc. Interspeech 2019 , 2019, pp. 3920–3924
2019
Cited alongside, same era.
S. Poria, D. Hazarika, N. Majumder, G. Naik, E. Cambria, and R. Mihalcea, “MELD: A multimodal multi-party dataset for emotion recognition in conversations,” in Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics . Florence, Italy: Association for Computational Linguistics, Jul. 2019, pp. 527–536
2019
Cited alongside, same era.
D. Dai, Z. Wu, R. Li, X. Wu, J. Jia, and H. Meng, “Learning discriminative features from spectrograms using center loss for speech emotion recognition,” in ICASSP 2019-2019 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2019, pp. 7405–7409
S. Latif, R. Rana, S. Khalifa, R. Jurdak, and B. W. Schuller, “Multitask learning from augmented auxiliary data for improving speech emotion recognition,” IEEE Transactions on Affective Computing , 2022
2022
Later among the works it cites.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
J. Gideon, M. G. McInnis, and E. M. Provost, “Improving cross-corpus speech emotion recognition with adversarial discriminative domain generalization (addog),” IEEE Transactions on Affective Computing , vol. 12, no. 4, pp. 1055–1068, 2019
2019
Cited alongside, same era.
F. Bao, M. Neumann, and N. T. Vu, “Cyclegan-based emotion style transfer as data augmentation for speech emotion recognition.” in INTERSPEECH , 2019, pp. 2828–2832
2019
Cited alongside, same era.
M. Neumann and N. T. Vu, “Improving speech emotion recognition with unsupervised representation learning on unlabeled speech,” in ICASSP 2019-2019 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) . IEEE, 2019, pp. 7390–7394
2019
Cited alongside, same era.
N. Majumder, S. Poria, D. Hazarika, R. Mihalcea, A. Gelbukh, and E. Cambria, “Dialoguernn: An attentive rnn for emotion detection in conversations,” in Proceedings of the AAAI conference on artificial intelligence , vol. 33, no. 01, 2019, pp. 6818–6825
2019
Cited alongside, same era.
T. Brown, B. Mann, N. Ryder, M. Subbiah, J. D. Kaplan, P. Dhariwal, A. Neelakantan, P. Shyam, G. Sastry, A. Askell et al. , “Language models are few-shot learners,” Advances in neural information processing systems , vol. 33, pp. 1877–1901, 2020
2020
Cited alongside, same era.
2020
Cited alongside, same era.
S. Latif, R. Rana, S. Khalifa, R. Jurdak, J. Epps, and B. W. Schuller, “Multi-task semi-supervised adversarial autoencoding for speech emotion recognition,” IEEE Transactions on Affective Computing , 2020
2020
Cited alongside, same era.
2021
Cited alongside, same era.
2023
Closest in time.
E. Hoes, S. Altay, and J. Bermeo, “Using chatgpt to fight misinformation: Chatgpt nails 72% of 12,000 verified claims,” 2023
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
I. Malik, S. Latif, R. Jurdak, and B. W. Schuller, “A preliminary study on augmenting speech emotion recognition using a diffusion model,” Proceedings of Interspeech, Dublin, Ireland, August, 2023 , 2023
2023
Closest in time.