Fetching the paper…
Reading the bibliography…
In this work, we conduct an extensive comparison of various approaches to speech based emotion recognition systems.
Stevens, S. S., Volkmann, J., & Newman, E. B. (1937). A Scale for the Measurement of the Psychological Magnitude Pitch
1937
Earlier work this paper cites.
Wiener, Norbert (1949). Extrapolation, Interpolation, and Smoothing of Stationary Time Series. New York: Wiley. ISBN 978-0-262-73005-1
1949
Earlier work this paper cites.
L.R. Rabiner and B.H. Juang (1986), "An Introduction to Hidden Markov Models," IEEE ASSP Magazine, January 1986
1986
Earlier work this paper cites.
Chen, C. H., Signal processing handbook, Dekker, New York, 1988
1988
Earlier work this paper cites.
Hochreiter, Sepp and Schmidhuber, Jürgen (1997), Long Short-Term Memory, Neural Computation, MIT Press,Cambridge, MA, USA. DOI = 10.1162/neco.1997.9.8.1735
1997
Earlier work this paper cites.
B. Schuller, G. Rigoll and M. Lang, "Hidden Markov model-based speech emotion recognition," 2003 International Conference on Multimedia and Expo. ICME ’03. Proceedings (Cat. No.03TH8698), Baltimore, MD, USA, 2003, pp. I-401. doi: 10.1109/ICME.2003.1220939
2003
Earlier work this paper cites.
S. Haq and P.J.B. Jackson. "Speaker-Dependent Audio-Visual Emotion Recognition", In Proc. Int’l Conf. on Auditory-Visual Speech Processing, pages 53-58, 2009
2009
Earlier work this paper cites.
Ayadi, M. E., Kamel, M. S., & Karray, F. (2011). Survey on speech emotion recognition: Features, classification schemes, and databases
2010
Earlier work this paper cites.
Kate Dupuis and M. Kathleen Pichora-Fuller (2010),Toronto emotional speech set (TESS), University of Toronto, Psychology Department
2010
Cited alongside, same era.
Ji, S., Xu, W., Yang, M., & Yu, K. (2010). 3D Convolutional Neural Networks for Human Action Recognition. IEEE Transactions on Pattern Analysis and Machine Intelligence, 35, 221-231
2010
Cited alongside, same era.
Bottou, Léon; Bousquet, Olivier (2012). "The Tradeoffs of Large Scale Learning". In Sra, Suvrit; Nowozin, Sebastian; Wright, Stephen J. (eds.). Optimization for Machine Learning. Cambridge: MIT Press. pp. 351–368. ISBN 978-0-262-01646-9
2012
Cited alongside, same era.
Krizhevsky, Alex & Sutskever, Ilya & Hinton, Geoffrey. (2012). ImageNet Classification with Deep Convolutional Neural Networks. Neural Information Processing Systems. 25. 10.1145/3065386
2012
Cited alongside, same era.
T. Fux and D. Jouvet, "Evaluation of PNCC and extended spectral subtraction methods for robust speech recognition," 2015 23rd European Signal Processing Conference (EUSIPCO), Nice, 2015, pp. 1416-1420. doi: 10.1109/EUSIPCO.2015.7362617
2015
Later among the works it cites.
Müller, Meinard (2015). Fundamentals of Music Processing. Springer. doi:10.1007/978-3-319-21945-5. ISBN 978-3-319-21944-8
2015
Later among the works it cites.
Shaw, A., Kumar, R., & Saxena, S. (2016). Emotion Recognition and Classification in Speech using Artificial Neural Networks
2016
Later among the works it cites.
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yamashita, Y. (2013). A review of paralinguistic information processing for natural speech communication
2013
Cited alongside, same era.
Diederik P. Kingma and Jimmy Ba (2014), Adam: A Method for Stochastic Optimization, arXiv, eprint=1412.6980
2014
Cited alongside, same era.
2014
Cited alongside, same era.
Chenchah, Farah, and Zied Lachiri. "Acoustic emotion recognition using linear and nonlinear cepstral coefficients." International Journal of Advanced Computer Science and Applications 6.11 (2015): 135-138
2015
Cited alongside, same era.
Lyons, J. Mel Frequency Cepstral Coefficient (MFCC) tutorial
Cited in the paper.
Livingstone, S. R., & Russo, F. A. (2018). The Ryerson Audio-Visual Database of Emotional Speech and Song (RAVDESS): A dynamic, multimodal set of facial and vocal expressions in North American English
2018
Later among the works it cites.
Srinivas Parthasarathy and Ivan Tashev (2018), Convolutional Neural Network Techniques For Speech Emotion Recognition, Microsoft Research
2018
Later among the works it cites.
Tomas G. S. (2019). Speech Emotion Recognition Using Convolutional Neural Networks. Technical University of Berlin. Retrieved from
2019
Closest in time.
United Nations Educational, Scientific, and Cultural Organization. (2019). I’d blush if I could: closing gender divides in digital skills through education (Programme Document GEN/2019/EQUALS/1 REV 2). Retrieved from
2019
Closest in time.