Fetching the paper…
Reading the bibliography…
One of the challenges in Speech Emotion Recognition (SER) "in the wild" is the large mismatch between training and test data (e.g.
L. R. Brody, “Gender differences in emotional development: A review of theories and research,”
1985
Earlier work this paper cites.
R. Caruna, “Multitask learning: A knowledge-based source of inductive bias,” in
1993
Earlier work this paper cites.
R. Caruana, “Multitask learning,” in
1998
Earlier work this paper cites.
L. Prechelt, “Automatic early stopping using cross validation: quantifying the criteria,”
1998
Earlier work this paper cites.
J. Baxter, “A model of inductive bias learning,”
2000
Earlier work this paper cites.
M. Liberman, K. Davis, M. Grossman, N. Martey, and J. Bell, “Emotional prosody speech and transcripts,”
2002
Earlier work this paper cites.
A. Batliner, C. Hacker, S. Steidl, E. Nöth, S. D’Arcy, M. J. Russell, and M. Wong, “You stupid tin box-children interacting with the aibo robot: A cross-linguistic emotional speech corpus.” in
2004
Earlier work this paper cites.
T. Vogt and E. André, “Comparing feature sets for acted and spontaneous speech in view of automatic emotion recognition,” in
2005
Earlier work this paper cites.
F. Burkhardt, A. Paeschke, M. Rolfes, W. F. Sendlmeier, and B. Weiss, “A database of german emotional speech.” in
2005
Earlier work this paper cites.
D. Ververidis and C. Kotropoulos, “Emotional speech recognition: Resources, features, and methods,”
2006
Earlier work this paper cites.
O. Martin, I. Kotsia, B. Macq, and I. Pitas, “The enterface’05 audio-visual emotion database,” in
2006
Earlier work this paper cites.
A. Evgeniou and M. Pontil, “Multi-task feature learning,”
2007
Cited alongside, same era.
C. Busso, M. Bulut, C.-C. Lee, A. Kazemzadeh, E. Mower, S. Kim, J. N. Chang, S. Lee, and S. S. Narayanan, “Iemocap: Interactive emotional dyadic motion capture database,”
2008
Cited alongside, same era.
L. v. d. Maaten and G. Hinton, “Visualizing data using t-sne,”
2008
Cited alongside, same era.
B. Schuller, B. Vlasenko, F. Eyben, M. Wollmer, A. Stuhlsatz, A. Wendemuth, and G. Rigoll, “Cross-corpus acoustic emotion recognition: variances and strategies,”
2010
Cited alongside, same era.
B. Schuller, Z. Zhang, F. Weninger, G. Rigoll
2011
Cited alongside, same era.
Z. Zhang, F. Weninger, M. Wollmer, and B. Schuller, “Unsupervised learning in cross-corpus acoustic emotion recognition,” in
J. Deng, Z. Zhang, F. Eyben, and B. Schuller, “Autoencoder-based unsupervised domain adaptation for speech emotion recognition,”
2014
Later among the works it cites.
J. Deng, Z. Zhang, and B. Schuller, “Linked source and target domain subspace feature transfer learning–exemplified by speech emotion recognition,” in
2014
Later among the works it cites.
Z. Zhang, P. Luo, C. C. Loy, and X. Tang, “Facial landmark detection by deep multi-task learning,” in
2014
Later among the works it cites.
I. T. Kun Han, Dong Yu, “Speech emotion recognition using deep neural network and extreme learning machine,” in
2014
Later among the works it cites.
D. Kingma and J. Ba, “Adam: A method for stochastic optimization,”
2014
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2011
Cited alongside, same era.
J. D. Gibbons and S. Chakraborti,
2011
Cited alongside, same era.
F. Eyben, M. Wöllmer, and B. Schuller, “A multitask approach to continuous five-dimensional affect sensing in natural speech,”
2012
Cited alongside, same era.
A.-r. Mohamed, G. E. Dahl, and G. Hinton, “Acoustic modeling using deep belief networks,”
2012
Cited alongside, same era.
M. L. Seltzer and J. Droppo, “Multi-task learning in deep neural networks for improved phoneme recognition,” in
2013
Cited alongside, same era.
Y. Bengio, A. Courville, and P. Vincent, “Representation learning: A review and new perspectives,”
2013
Cited alongside, same era.
N. Srivastava, G. E. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov, “Dropout: a simple way to prevent neural networks from overfitting.”
2014
Later among the works it cites.
2015
Later among the works it cites.
J. Lee and I. Tashev, “High-level feature representation using recurrent neural network for speech emotion recognition,” in
2015
Later among the works it cites.
B. Jou and S.-F. Chang, “Deep cross residual learning for multitask visual recognition,” in
2016
Later among the works it cites.
B. Zhang, E. M. Provost, and G. Essi, “Cross-corpus acoustic emotion recognition from singing and speaking: A multi-task learning approach,” in
2016
Later among the works it cites.
G. Trigeorgis, F. Ringeval, R. Brueckner, E. Marchi, M. A. Nicolaou, S. Zafeiriou
2016
Later among the works it cites.