Fetching the paper…
Reading the bibliography…
Emotion represents an essential aspect of human speech that is manifested in speech prosody.
S. Shimojo and L. Shams, “Sensory modalities are not separate modalities: plasticity and interactions,”
2001
Earlier work this paper cites.
L. F. Barrett, “Solving the emotion paradox: Categorization and the experience of emotion,”
2006
Earlier work this paper cites.
C. Busso, M. Bulut, C.-C. Lee, A. Kazemzadeh, E. Mower, S. Kim, J. N. Chang, S. Lee, and S. S. Narayanan, “IEMOCAP: Interactive emotional dyadic motion capture database,”
2008
Earlier work this paper cites.
F. Eyben, M. Wöllmer, and B. Schuller, “Opensmile: the munich versatile and fast open-source audio feature extractor,” in
2010
Earlier work this paper cites.
S. Ji, W. Xu, M. Yang, and K. Yu, “3d convolutional neural networks for human action recognition,”
2012
Earlier work this paper cites.
2012
Earlier work this paper cites.
V. Pérez-Rosas, R. Mihalcea, and L.-P. Morency, “Utterance-level multimodal sentiment analysis,” in
2013
Earlier work this paper cites.
M. Wöllmer, F. Weninger, T. Knaup, B. Schuller, C. Sun, K. Sagae, and L.-P. Morency, “Youtube movie reviews: Sentiment analysis in an audio-visual context,”
2013
Earlier work this paper cites.
B. Schuller, S. Steidl, A. Batliner, A. Vinciarelli, K. Scherer, F. Ringeval, M. Chetouani, F. Weninger, F. Eyben, E. Marchi
2013
Earlier work this paper cites.
2013
Earlier work this paper cites.
A. Karpathy, G. Toderici, S. Shetty, T. Leung, R. Sukthankar, and L. Fei-Fei, “Large-scale video classification with convolutional neural networks,” in
2014
Earlier work this paper cites.
S. Poria, E. Cambria, and A. Gelbukh, “Deep convolutional neural network textual features and multiple kernel learning for utterance-level multimodal sentiment analysis,” in
2015
Cited alongside, same era.
M. Sreeshakthy and J. Preethi, “Classification of human emotion from deap eeg signal using hybrid improved neural networks with cuckoo search,”
2016
Cited alongside, same era.
S. Poria, E. Cambria, N. Howard, G.-B. Huang, and A. Hussain, “Fusing audio, visual and textual clues for sentiment analysis from multimodal content,”
2016
Cited alongside, same era.
J. Xue, Z. Luo, K. Eguchi, T. Takiguchi, and T. Omoto, “A bayesian nonparametric multimodal data modeling framework for video emotion recognition,” in
2017
Cited alongside, same era.
S. Poria, E. Cambria, R. Bajpai, and A. Hussain, “A review of affective computing: From unimodal analysis to multimodal fusion,”
2019
Later among the works it cites.
M. S. Hossain and G. Muhammad, “Emotion recognition using deep learning approach from audio–visual emotional big data,”
2019
Later among the works it cites.
S. Tripathi and H. Beigi, “Multi-modal emotion recognition on iemocap dataset using deep learning,”
2019
Later among the works it cites.
Y. Wang, Y. Shen, Z. Liu, P. P. Liang, A. Zadeh, and L.-P. Morency, “Words can shift: Dynamically adjusting word representations using nonverbal behaviors,” in
2019
Later among the works it cites.
J. Sebastian and P. Pierucci, “Fusion techniques for utterance-level emotion recognition combining speech and transcripts,” in
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2017
Cited alongside, same era.
E. Cambria, D. Hazarika, S. Poria, A. Hussain, and R. Subramanyam, “Benchmarking multimodal sentiment analysis,” in
2017
Cited alongside, same era.
S. Poria, E. Cambria, D. Hazarika, N. Majumder, A. Zadeh, and L.-P. Morency, “Context-dependent sentiment analysis in user-generated videos,” in
2017
Cited alongside, same era.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. Kaiser, and I. Polosukhin, “Attention is all you need,” in
2017
Cited alongside, same era.
Z. Liu, Y. Shen, V. B. Lakshminarasimhan, P. P. Liang, A. B. Zadeh, and L.-P. Morency, “Efficient low-rank multimodal fusion with modality-specific factors,” in
2018
Cited alongside, same era.
S. Poria, N. Majumder, D. Hazarika, E. Cambria, A. Gelbukh, and A. Hussain, “Multimodal sentiment analysis: Addressing key issues and setting up the baselines,”
2018
Cited alongside, same era.
2019
Later among the works it cites.
E. Georgiou, C. Papaioannou, and A. Potamianos, “Deep hierarchical fusion with application in sentiment analysis,”
2019
Later among the works it cites.
Y.-H. H. Tsai, S. Bai, P. P. Liang, J. Z. Kolter, L.-P. Morency, and R. Salakhutdinov, “Multimodal transformer for unaligned multimodal language sequences,” in
2019
Later among the works it cites.
H. Le, D. Sahoo, N. Chen, and S. Hoi, “Multimodal transformer networks for end-to-end video-grounded dialogue systems,” in
2019
Later among the works it cites.
Y. H. Tsai, P. P. Liang, A. Zadeh, L. Morency, and R. Salakhutdinov, “Learning factorized multimodal representations,” in
2019
Later among the works it cites.