Fetching the paper…
Reading the bibliography…
In this work we design a neural network for recognizing emotions in speech, using the IEMOCAP dataset.
L. Lee and R. Rose, “A frequency warping approach to speaker normalization,”
1998
Earlier work this paper cites.
C. Busso, M. Bulut, C.-C. Lee, A. Kazemzadeh, E. Mower, S. Kim, J. N. Chang, S. Lee, and S. S. Narayanan, “Iemocap: interactive emotional dyadic motion capture database,”
2008
Earlier work this paper cites.
Y. Kim, H. Lee, and E. Mower Provost, “Deep learning for robust feature generation in audiovisual emotion recognition,” in
2013
Earlier work this paper cites.
N. Jaitly and G. E. Hinton, “Vocal tract length perturbation (VTLP) improves speech recognition,” in
2013
Earlier work this paper cites.
X. Cui, V. Goel, and B. Kingsbury, “Data augmentation for deep neural network acoustic modeling,” in
2014
Earlier work this paper cites.
D. Amodei and etc., “Deep speech 2: End-to-end speech recognition in english and mandarin,” in
2015
Earlier work this paper cites.
J. Lee and I. Tashev, “High-level feature representation using recurrent neural network for speech emotion recognition,” in
2015
Cited alongside, same era.
T. Sainath, O. Vinyals, A. Senior, and H. Sak, “Convolutional long short-term memory, fully connected deep neural networks,” in
2015
Cited alongside, same era.
2015
Cited alongside, same era.
I. Medennikov, A. Prudnikov, and A. Zatvornitskiy, “Improving english conversational telephone speech recognition,” in
2016
Cited alongside, same era.
G. Saon, T. Sercu, S. Rennie, and H.-K. J. Kuo, “The ibm 2016 english conversational telephone speech recognition system,” in
2016
Cited alongside, same era.
T. Cooijmans, N. Ballas, C. Laurent, C. Gulcehre, and A. Courville, “Recurrent batch normalization,”
2016
Later among the works it cites.
J. Ba, R. Kiros, and G. E. Hinton, “Layer normalization,”
2016
Later among the works it cites.
H. Harutyunyan and E. Sanogh, “Khosk’its’ lezvi chanach’um khory usuts’man met’vodnerov, BS thesis,” 2016
2016
Later among the works it cites.
2017
Later among the works it cites.
A. Satt, S. Rozenberg, and R. Hoory, “Efficient emotion recognition from speech using deep learning on spectrograms,” in
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
C. Laurent, G. Pereyra, P. Brakel, Y. Zhang, and Y. Bengio, “Batch normalized recurrent neural networks,” in
2016
Cited alongside, same era.