Fetching the paper…
Reading the bibliography…
Speech Emotion Recognition (SER) application is frequently associated with privacy concerns as it often acquires and transmits speech data at the client-side to remote cloud platforms for further processing.
W. Li, Y. Zhang, and Y. Fu, “Speech emotion recognition in e-learning system based on affective computing,” in Third International Conference on Natural Computation (ICNC 2007) , vol. 5. IEEE, 2007, pp. 809–813
2007
Earlier work this paper cites.
C. Busso, M. Bulut, C.-C. Lee, A. Kazemzadeh, E. Mower, S. Kim, J. N. Chang, S. Lee, and S. S. Narayanan, “IEMOCAP: Interactive emotional dyadic motion capture database,” Language resources and evaluation , vol. 42, no. 4, pp. 335–359, 2008
2008
Earlier work this paper cites.
F. Eyben, M. Wöllmer, and B. Schuller, “Opensmile: the munich versatile and fast open-source audio feature extractor,” in Proceedings of the 18th ACM international conference on Multimedia , 2010, pp. 1459–1462
2010
Earlier work this paper cites.
S. G. Koolagudi and K. S. Rao, “Emotion recognition from speech: a review,” International journal of speech technology , vol. 15, no. 2, pp. 99–117, 2012
2012
Earlier work this paper cites.
S. Ramakrishnan and I. M. El Emary, “Speech emotion recognition approaches in human computer interaction,” Telecommunication Systems , vol. 52, no. 3, pp. 1467–1478, 2013
2013
Earlier work this paper cites.
C. Busso, S. Parthasarathy, A. Burmania, M. AbdelWahab, N. Sadoughi, and E. M. Provost, “Msp-improv: An acted corpus of dyadic interactions to study emotion perception,” IEEE Transactions on Affective Computing , vol. 8, no. 1, pp. 67–80, 2016
2016
Earlier work this paper cites.
D. Bone, C.-C. Lee, T. Chaspari, J. Gibson, and S. Narayanan, “Signal processing and machine learning for mental health research and clinical applications,” IEEE Signal Processing Magazine , vol. 34, no. 5, pp. 189–196, September 2017
2017
Earlier work this paper cites.
B. McMahan, E. Moore, D. Ramage, S. Hampson, and B. A. y Arcas, “Communication-efficient learning of deep networks from decentralized data,” in Artificial intelligence and statistics . PMLR, 2017, pp. 1273–1282
2017
Earlier work this paper cites.
A. Satt, S. Rozenberg, and R. Hoory, “Efficient emotion recognition from speech using deep learning on spectrograms.” in Interspeech , 2017, pp. 1089–1093
2017
Earlier work this paper cites.
Y. Zhang, J. Du, Z. Wang, J. Zhang, and Y. Tu, “Attention based fully convolutional network for speech emotion recognition,” in 2018 Asia-Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA ASC) . IEEE, 2018, pp. 1771–1775
2018
Earlier work this paper cites.
G. Ramet, P. N. Garner, M. Baeriswyl, and A. Lazaridis, “Context-aware attention mechanism for speech emotion recognition,” in 2018 IEEE Spoken Language Technology Workshop (SLT) . IEEE, 2018, pp. 126–131
2018
Cited alongside, same era.
Y.-A. Chung, W.-N. Hsu, H. Tang, and J. Glass, “An unsupervised autoregressive model for speech representation learning,” in Interspeech , 2019
2019
Cited alongside, same era.
M.-C. Lee, S.-Y. Chiang, S.-C. Yeh, and T.-F. Wen, “Study on emotion recognition and companion chatbot using deep neural network,” Multimedia Tools and Applications , vol. 79, no. 27, pp. 19 629–19 657, 2020
2020
Cited alongside, same era.
2020
Cited alongside, same era.
2021
Later among the works it cites.
P. Li, D. Li, W. Li, S. Gong, Y. Fu, and T. M. Hospedales, “A simple feature augmentation for domain generalization,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2021, pp. 8886–8895
2021
Later among the works it cites.
B. Xiong, H. Fan, K. Grauman, and C. Feichtenhofer, “Multiview pseudo-labeling for semi-supervised learning from video,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2021, pp. 7209–7219
2021
Later among the works it cites.
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
L. Zhu and S. Han, “Deep leakage from gradients,” in Federated learning . Springer, 2020, pp. 17–31
2020
Cited alongside, same era.
K. Sohn, D. Berthelot, N. Carlini, Z. Zhang, H. Zhang, C. A. Raffel, E. D. Cubuk, A. Kurakin, and C.-L. Li, “Fixmatch: Simplifying semi-supervised learning with consistency and confidence,” Advances in Neural Information Processing Systems , vol. 33, pp. 596–608, 2020
2020
Cited alongside, same era.
S. P. Karimireddy, S. Kale, M. Mohri, S. Reddi, S. Stich, and A. T. Suresh, “Scaffold: Stochastic controlled averaging for federated learning,” in International Conference on Machine Learning . PMLR, 2020, pp. 5132–5143
2020
Cited alongside, same era.
S. Ling and Y. Liu, “Decoar 2.0: Deep contextualized acoustic representations with vector quantization,” 2020
2020
Cited alongside, same era.
2020
Cited alongside, same era.
A. T. Liu, S.-w. Yang, P.-H. Chi, P.-c. Hsu, and H.-y. Lee, “Mockingjay: Unsupervised speech representation learning with deep bidirectional transformer encoders,” ICASSP 2020 - 2020 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) , May 2020
2020
Cited alongside, same era.
S. wen Yang, P.-H. Chi, Y.-S. Chuang, C.-I. J. Lai, K. Lakhotia, Y. Y. Lin, A. T. Liu, J. Shi, X. Chang, G.-T. Lin, T.-H. Huang, W.-C. Tseng, K. tik Lee, D.-R. Liu, Z. Huang, S. Dong, S.-W. Li, S. Watanabe, A. Mohamed, and H. yi Lee, “SUPERB: Speech Processing Universal PERformance Benchmark,” in Proc. Interspeech 2021 , 2021, pp. 1194–1198
2021
Later among the works it cites.
A. T. Liu, S.-W. Li, and H.-y. Lee, “Tera: Self-supervised learning of transformer encoder representation for speech,” IEEE/ACM Transactions on Audio, Speech, and Language Processing , vol. 29, pp. 2351–2366, 2021
2021
Later among the works it cites.
2021
Later among the works it cites.
2021
Later among the works it cites.