Fetching the paper…
Reading the bibliography…
Acoustic emotion recognition aims to categorize the affective state of the speaker and is still a difficult task for machine learning models.
DARPA TIMIT acoustic-phonetic continuous speech corpus CD-ROM. NIST speech disc 1-1.1
J. S. Garofolo, L. F. Lamel, W. M. Fisher, J. G. Fiscus, and D. S. Pallett. 1993 · 1993
Earlier work this paper cites.
Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks
Alex Graves, Santiago Fernandez, Faustino Gomez, and Jürgen Schmidhuber. 2006 · 2006
Earlier work this paper cites.
IEMOCAP: interactive emotional dyadic motion capture database
Carlos Busso, Murtaza Bulut, Chi-Chun Lee, Abe Kazemzadeh, Emily Mower, Samuel Kim, Jeannette N. Chang, Sungbok Lee, and Shrikanth S. Narayanan. 2008 · 2008
Earlier work this paper cites.
Acoustic emotion recognition: A benchmark comparison of performances
Björn Schuller, Bogdan Vlasenko, Florian Eyben, Gerhard Rigoll, and Andreas Wendemuth. 2009 · 2009
Earlier work this paper cites.
Deep learning of representations for unsupervised and transfer learning
Yoshua Bengio. 2012 · 2012
Earlier work this paper cites.
ImageNet Classification with Deep Convolutional Neural Networks
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E Hinton. 2012 · 2012
Earlier work this paper cites.
Recent developments in opensmile, the munich open-source multimedia feature extractor
Florian Eyben, Felix Weninger, Florian Gross, and Björn Schuller. 2013 · 2013
Earlier work this paper cites.
The interspeech 2013 computational paralinguistics challenge: social signals, conflict, emotion, autism
Björn Schuller, Stefan Steidl, Anton Batliner, Alessandro Vinciarelli, Klaus Scherer, Fabien Ringeval, Mohamed Chetouani, Felix Weninger, Florian Eyben, Erik Marchi, et al. 2013 · 2013
Earlier work this paper cites.
Deep Speech: Scaling up end-to-end speech recognition
Awni Hannun, Carl Case, Jared Casper, Bryan Catanzaro, and et al. 2014 · 2014
Cited alongside, same era.
Adam: A Method for Stochastic Optimization
Diederik P. Kingma and Jimmy Ba. 2014 · 2014
Cited alongside, same era.
Enhancing the ted-lium corpus with selected data for language modeling and more ted talks
Anthony Rousseau, Paul Deléglise, and Yannick Estève. 2014 · 2014
Cited alongside, same era.
Very deep convolutional networks for large-scale image recognition
Karen Simonyan and Andrew Zisserman. 2014 · 2014
Cited alongside, same era.
Sequence to sequence learning with neural networks
Ilya Sutskever, Oriol Vinyals, and Quoc V. Le. 2014 · 2014
Cited alongside, same era.
Beyond Fine Tuning: A Modular Approach to Learning on Small Data
Ark Anderson, Kyle Shaffer, Artem Yankov, Court Corley, and Nathan Hodas. 2016 · 2016
Later among the works it cites.
On the Correlation and Transferability of Features Between Automatic Speech Recognition and Speech Emotion Recognition
Haytham M. Fayek, Margaret Lech, and Lawrence Cavedon. 2016 · 2016
Later among the works it cites.
Representation Learning for Speech Emotion Recognition
Sayan Ghosh, Eugene Laksana, Louis-Philippe Morency, and Stefan Scherer. 2016 · 2016
Later among the works it cites.
Attention Assisted Discovery of Sub-Utterance Structure in Speech Emotion Recognition
Che-Wei Huang and Shrikanth S. Narayanan. 2016 · 2016
Later among the works it cites.
Andrei A. Rusu, Neil C. Rabinowitz, Guillaume Desjardins, Hubert Soyer, James Kirkpatrick, Koray Kavukcuoglu, Razvan Pascanu, and Raia Hadsell. 2016 · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
How transferable are features in deep neural networks?
Jason Yosinski, Jeff Clune, Yoshua Bengio, and Hod Lipson. 2014 · 2014
Cited alongside, same era.
Deep speech 2: End-to-end speech recognition in English and Mandarin
Dario Amodei, Rishita Anubhai, Eric Battenberg, Carl Case, Jared Casper, and et al. 2015 · 2015
Cited alongside, same era.
Librispeech: an ASR corpus based on public domain audio books
Vassil Panayotov, Guoguo Chen, Daniel Povey, and Sanjeev Khudanpur. 2015 · 2015
Cited alongside, same era.
Later among the works it cites.
Adieu features? End-to-end speech emotion recognition using a deep convolutional recurrent network
George Trigeorgis, Fabien Ringeval, Raymond Brueckner, Erik Marchi, Mihalis A. Nicolaou, Björn Schuller, and Stefanos Zafeiriou. 2016 · 2016
Later among the works it cites.
Emotion recognition from speech with recurrent neural networks
Vladimir Chernykh, Grigoriy Sterling, and Pavel Prihodko. 2017 · 2017
Later among the works it cites.
Evaluating deep learning architectures for Speech Emotion Recognition
Haytham M. Fayek, Margaret Lech, and Lawrence Cavedon. 2017 · 2017
Later among the works it cites.