Fetching the paper…
Reading the bibliography…
Different from the emotion recognition in individual utterances, we propose a multimodal learning framework using relation and dependencies among the utterances for conversational emotion analysis.
C. Busso, M. Bulut, C.-C. Lee, A. Kazemzadeh, E. Mower, S. Kim, J. N. Chang, S. Lee, and S. S. Narayanan, “Iemocap: Interactive emotional dyadic motion capture database,”
2008
Earlier work this paper cites.
M. Paleari, R. Benmokhtar, and B. Huet, “Evidence theory-based multimodal emotion recognition,” in
2009
Earlier work this paper cites.
C.-C. Lee, C. Busso, S. Lee, and S. S. Narayanan, “Modeling mutual influence of interlocutor emotion states in dyadic spoken interactions,” in
2009
Earlier work this paper cites.
A. Metallinou, S. Lee, and S. Narayanan, “Decision level combination of multiple modalities for recognition and analysis of emotional expression,” in
2010
Earlier work this paper cites.
F. Eyben, M. Wöllmer, and B. Schuller, “Opensmile: the munich versatile and fast open-source audio feature extractor,” in
2010
Earlier work this paper cites.
A. F. Martin and C. S. Greenberg, “The nist 2010 speaker recognition evaluation,” in
2010
Earlier work this paper cites.
G. A. Ramirez, T. Baltrušaitis, and L.-P. Morency, “Modeling latent discriminative dynamic of multi-dimensional affective signals,” in
2011
Earlier work this paper cites.
D. Jiang, Y. Cui, X. Zhang, P. Fan, I. Ganzalez, and H. Sahli, “Audio visual emotion recognition based on triple-stream dynamic bayesian network models,” in
2011
Earlier work this paper cites.
J. J. Gross and L. Feldman Barrett, “Emotion generation and emotion regulation: One or two depends on your point of view,”
2011
Earlier work this paper cites.
D. Povey, A. Ghoshal, G. Boulianne, L. Burget, O. Glembek, N. Goel, M. Hannemann, P. Motlicek, Y. Qian, P. Schwarz
2011
Earlier work this paper cites.
F. Eyben, S. Petridis, B. Schuller, and M. Pantic, “Audiovisual vocal outburst classification in noisy acoustic conditions,” in
2012
Earlier work this paper cites.
B. Schuller, M. Valster, F. Eyben, R. Cowie, and M. Pantic, “Avec 2012: the continuous audio/visual emotion challenge,” in
2012
Earlier work this paper cites.
V. Rozgić, S. Ananthakrishnan, S. Saleem, R. Kumar, and R. Prasad, “Ensemble of svm trees for multimodal emotion recognition,” in
2012
Cited alongside, same era.
C.-H. Wu, J.-C. Lin, and W.-L. Wei, “Two-level hierarchical alignment for semi-coupled hmm-based audiovisual emotion recognition with temporal course,”
2013
Cited alongside, same era.
2013
Cited alongside, same era.
C.-H. Wu, J.-C. Lin, and W.-L. Wei, “Survey on audiovisual emotion recognition: databases, features, and data fusion strategies,”
2014
Cited alongside, same era.
A. Vanzo, D. Croce, and R. Basili, “A context-based model for sentiment analysis in twitter,” in
2014
Cited alongside, same era.
S. Poria, E. Cambria, D. Hazarika, N. Majumder, A. Zadeh, and L.-P. Morency, “Context-dependent sentiment analysis in user-generated videos,” in
2017
Later among the works it cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” in
2017
Later among the works it cites.
R. Zhang, A. Ando, S. Kobashikawa, and Y. Aono, “Interaction and transition model for speech emotion recognition in dialogue.” in
2017
Later among the works it cites.
S. Chen, Q. Jin, J. Zhao, and S. Wang, “Multimodal multi-task learning for dimensional and continuous emotion recognition,” in
2017
Later among the works it cites.
D. Snyder, D. Garcia-Romero, D. Povey, and S. Khudanpur, “Deep neural network embeddings for text-independent speaker verification.” in
2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2014
Cited alongside, same era.
2014
Cited alongside, same era.
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov, “Dropout: a simple way to prevent neural networks from overfitting,”
2014
Cited alongside, same era.
2015
Cited alongside, same era.
Q. Jin, C. Li, S. Chen, and H. Wu, “Speech emotion recognition with acoustic and lexical features,” in
2015
Cited alongside, same era.
S. Chen and Q. Jin, “Multi-modal conditional attention fusion for dimensional emotion prediction,” in
2016
Cited alongside, same era.
S. Poria, E. Cambria, D. Hazarika, N. Mazumder, A. Zadeh, and L.-P. Morency, “Multi-level multiple attentions for contextual multimodal sentiment analysis,” in
2017
Cited alongside, same era.
S. O. Sadjadi, T. Kheyrkhah, A. Tong, C. S. Greenberg, D. A. Reynolds, E. Singer, L. P. Mason, and J. Hernandez-Cordero, “The 2016 nist speaker recognition evaluation.” in
2017
Later among the works it cites.
2018
Later among the works it cites.
J. Zhao, R. Li, S. Chen, and Q. Jin, “Multi-modal multi-cultural dimensional continues emotion recognition in dyadic interactions,” in
2018
Later among the works it cites.
D. Snyder, D. Garcia-Romero, G. Sell, D. Povey, and S. Khudanpur, “X-vectors: Robust dnn embeddings for speaker recognition,” in
2018
Later among the works it cites.
2018
Later among the works it cites.
R. Li, Z. Wu, J. Jia, J. Li, W. Chen, and H. Meng, “Inferring user emotive state changes in realistic human-computer conversational dialogs,” in
2018
Later among the works it cites.