Fetching the paper…
Reading the bibliography…
Automatic emotion recognition plays a key role in computer-human interaction as it has the potential to enrich the next-generation artificial intelligence with emotional intelligence.
R. Cowie, E. Douglas-Cowie, N. Tsapatsoulis, G. Votsis, S. Kollias, W. Fellenz, and J. G. Taylor, “Emotion recognition in human-computer interaction,”
2001
Earlier work this paper cites.
O.-W. Kwon, K. Chan, J. Hao, and T.-W. Lee, “Emotion recognition by speech signals,” in
2003
Earlier work this paper cites.
N. Sebe, I. Cohen, and T. S. Huang, “Multimodal emotion recognition,” in
2005
Earlier work this paper cites.
D. Ververidis and C. Kotropoulos, “Emotional speech recognition: Resources, features, and methods,”
2006
Earlier work this paper cites.
H. Richard, R. Tom, R. Yvonne, and S. Abigail,
2008
Earlier work this paper cites.
R. Beale and C. Peter, “The role of affect and emotion in HCI,” in
2008
Earlier work this paper cites.
C. Busso, M. Bulut, C.-C. Lee, A. Kazemzadeh, E. Mower, S. Kim, J. N. Chang, S. Lee, and S. S. Narayanan, “IEMOCAP: Interactive emotional dyadic motion capture database,”
2008
Earlier work this paper cites.
S. Haq and P. J. Jackson, “Multimodal emotion recognition,” in
2011
Earlier work this paper cites.
2013
Earlier work this paper cites.
J. Deng, Z. Zhang, E. Marchi, and B. Schuller, “Sparse autoencoder-based feature transfer learning for speech emotion recognition,” in
2013
Earlier work this paper cites.
K. Han, D. Yu, and I. Tashev, “Speech emotion recognition using deep neural network and extreme learning machine,” in
2014
Earlier work this paper cites.
J. Pennington, R. Socher, and C. D. Manning, “Glove: Global vectors for word representation,” in
2014
Earlier work this paper cites.
P. Song, Y. Jin, L. Zhao, and M. Xin, “Speech emotion recognition using transfer learning,”
2014
Earlier work this paper cites.
G. Keren and B. Schuller, “Convolutional RNN: an enhanced model for extracting features from sequential data,” in
2016
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in
2016
Earlier work this paper cites.
C.-W. Huang and S. Narayanan, “Attention assisted discovery of sub-utterance structure in speech emotion recognition,” in
2016
Earlier work this paper cites.
S. Ghosh, E. Laksana, L.-P. Morency, and S. Scherer, “Representation learning for speech emotion recognition,” in
2016
Earlier work this paper cites.
H. Ranganathan, S. Chakraborty, and S. Panchanathan, “Multimodal emotion recognition using deep learning architectures,” in
2016
Earlier work this paper cites.
A. Satt, S. Rozenberg, and R. Hoory, “Efficient emotion recognition from speech using deep learning on spectrograms,” in
2017
Earlier work this paper cites.
M. Neumann and N. T. Vu, “Attentive convolutional neural network based speech emotion recognition: A study on the impact of input features, signal length, and acted speech,” in
2017
Earlier work this paper cites.
S. Mirsamadi, E. Barsoum, and C. Zhang, “Automatic speech emotion recognition using recurrent neural networks with local attention,” in
2017
Cited alongside, same era.
J. Gideon, S. Khorram, Z. Aldeneh, D. Dimitriadis, and E. M. Provost, “Progressive neural networks for transfer learning in emotion recognition,” in
2017
Cited alongside, same era.
P. Tzirakis, G. Trigeorgis, M. A. Nicolaou, B. W. Schuller, and S. Zafeiriou, “End-to-end multimodal emotion recognition using deep neural networks,”
2017
Cited alongside, same era.
2017
Cited alongside, same era.
B. W. Schuller, “Speech emotion recognition: Two decades in a nutshell, benchmarks, and ongoing trends,”
2018
D. S. Park, W. Chan, Y. Zhang, C.-C. Chiu, B. Zoph, E. D. Cubuk, and Q. V. Le, “SpecAugment: A simple data augmentation method for automatic speech recognition,” in
2019
Later among the works it cites.
A. Chatziagapi, G. Paraskevopoulos, D. Sgouropoulos, G. Pantazopoulos, M. Nikandrou, T. Giannakopoulos, A. Katsamanis, A. Potamianos, and S. Narayanan, “Data augmentation using GANs for speech emotion recognition,” in
2019
Later among the works it cites.
J. Gideon, M. McInnis, and E. M. Provost, “Improving cross-corpus speech emotion recognition with adversarial discriminative domain generalization (ADDoG),”
2019
Later among the works it cites.
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Z. Yang and J. Hirschberg, “Predicting arousal and valence from waveforms and spectrograms using deep neural networks,” in
2018
Cited alongside, same era.
X. Ma, Z. Wu, J. Jia, M. Xu, H. Meng, and L. Cai, “Emotion recognition from variable-length speech segments using deep learning on spectrograms,” in
2018
Cited alongside, same era.
P. Yenigalla, A. Kumar, S. Tripathi, C. Singh, S. Kar, and J. Vepa, “Speech emotion recognition using spectrogram & phoneme embedding,” in
2018
Cited alongside, same era.
M. Sarma, P. Ghahremani, D. Povey, N. K. Goel, K. K. Sarma, and N. Dehak, “Emotion identification from raw speech signals using DNNs,” in
2018
Cited alongside, same era.
2018
Cited alongside, same era.
G. Ramet, P. N. Garner, M. Baeriswyl, and A. Lazaridis, “Context-aware attention mechanism for speech emotion recognition,” in
2018
Cited alongside, same era.
2018
Cited alongside, same era.
2019
Later among the works it cites.
2019
Later among the works it cites.
M. Neumann and N. T. Vu, “Improving speech emotion recognition with unsupervised representation learning on unlabeled speech,” in
2019
Later among the works it cites.
L. Tarantino, P. N. Garner, and A. Lazaridis, “Self-attention for speech emotion recognition,” in
2019
Later among the works it cites.
R. Pappagari, T. Wang, J. Villalba, N. Chen, and N. Dehak, “X-vectors meet emotions: A study on dependencies between emotion and speaker recognition,” in
2020
Later among the works it cites.
K. Feng and T. Chaspari, “A review of generalizable transfer learning in automatic emotion recognition,”
2020
Later among the works it cites.
G. Boateng and T. Kowatsch, “Speech emotion recognition among elderly individuals using multimodal fusion and transfer learning,” in
2020
Later among the works it cites.
L. Pepino, P. Riera, L. Ferrer, and A. Gravano, “Fusion approaches for emotion recognition from speech using acoustic and text-based features,” in
2020
Later among the works it cites.
S. O. Sadjadi, C. Greenberg, E. Singer, D. Reynolds, L. Mason, and J. Hernandez-Cordero, “The 2019 NIST audio-visual speaker recognition evaluation,” in
2020
Later among the works it cites.
A. Nagrani, J. S. Chung, W. Xie, and A. Zisserman, “Voxceleb: Large-scale speaker verification in the wild,”
2020
Later among the works it cites.
W. Wu, C. Zhang, and P. C. Woodland, “Emotion recognition by fusing time synchronous and time asynchronous representations,” in
2021
Later among the works it cites.
S. Padi, S. O. Sadjadi, R. D. Sriram, and D. Manocha, “Improved speech emotion recognition using transfer learning and spectrogram augmentation,” in
2021
Later among the works it cites.
2021
Later among the works it cites.
S.-w. Yang, P.-H. Chi, Y.-S. Chuang, C.-I. J. Lai, K. Lakhotia, Y. Y. Lin, A. T. Liu, J. Shi, X. Chang, G.-T. Lin
2021
Later among the works it cites.
S. Chen, C. Wang, Z. Chen, Y. Wu, S. Liu, Z. Chen, J. Li, N. Kanda, T. Yoshioka, X. Xiao
2021
Later among the works it cites.
U. Evci, V. Dumoulin, H. Larochelle, and M. C. Mozer, “Head2Toe: Utilizing intermediate representations for better transfer learning,” 2022
2022
Closest in time.