Fetching the paper…
Reading the bibliography…
Multimodal Sentiment Analysis is an active area of research that leverages multimodal signals for affective understanding of user-generated videos.
VisualBERT: A Simple and Performant Baseline for Vision and Language
Liunian Harold Li, Mark Yatskar, Da Yin, Cho-Jui Hsieh, and Kai-Wei Chang. 2019 · 1908
Earlier work this paper cites.
VL-BERT: Pre-training of Generic Visual-Linguistic Representations
Weijie Su, Xizhou Zhu, Yue Cao, Bin Li, Lewei Lu, Furu Wei, and Jifeng Dai. 2019 · 1908
Earlier work this paper cites.
Zhongkai Sun, Prathusha K. Sarma, William A. Sethares, and Yingyu Liang. 2019 · 1911
Earlier work this paper cites.
Joint Robust Voicing Detection and Pitch Estimation Based on Residual Harmonics. In INTERSPEECH 2011, 12th Annual Conference of the International Speech Communication Association . ISCA, Florence, Italy, 1973–1976
Thomas Drugman and Abeer Alwan. 2011 · 1976
Earlier work this paper cites.
From utterance to text: The bias of language in speech and writing
David Olson. 1977 · 1977
Earlier work this paper cites.
What the face reveals: Basic and applied studies of spontaneous expression using the Facial Action Coding System (FACS)
Rosenberg Ekman. 1997 · 1997
Earlier work this paper cites.
Long Short-Term Memory
Sepp Hochreiter and Jürgen Schmidhuber. 1997 · 1997
Earlier work this paper cites.
Unsupervised Domain Clusters in Pretrained Language Models
Roee Aharoni and Yoav Goldberg. 2020 · 2004
Earlier work this paper cites.
Beneath the Tip of the Iceberg: Current Challenges and New Directions in Sentiment Analysis Research
Soujanya Poria, Devamanyu Hazarika, Navonil Majumder, and Rada Mihalcea. 2020 · 2005
Earlier work this paper cites.
Visualizing data using t-SNE
Laurens van der Maaten and Geoffrey Hinton. 2008 · 2008
Earlier work this paper cites.
Factorized Orthogonal Latent Spaces. In Proceedings of the Thirteenth International Conference on Artificial Intelligence and Statistics, AISTATS 2010 (JMLR Proceedings, Vol. 9) . JMLR.org, Sardinia, Italy, 701–708
Mathieu Salzmann, Carl Henrik Ek, Raquel Urtasun, and Trevor Darrell. 2010 · 2010
Earlier work this paper cites.
Detection of Glottal Closure Instants From Speech Signals: A Quantitative Review
Thomas Drugman, Mark R. P. Thomas, Jón Guðnason, Patrick A. Naylor, and Thierry Dutoit. 2012 · 2011
Earlier work this paper cites.
Multi-View Latent Variable Discriminative Models for Action Recognition. In Proceedings of the 2012 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (CVPR ’12) . IEEE Computer Society, USA, 2120–2127
Randall Davis. 2012 · 2012
Earlier work this paper cites.
Multimodal Sentiment Analysis. In Proceedings of the 3rd Workshop in Computational Approaches to Subjectivity and Sentiment Analysis (Jeju, Republic of Korea) (WASSA ’12) . Association for Computational Linguistics, USA, 1
Rada Mihalcea. 2012 · 2012
Earlier work this paper cites.
Deep Canonical Correlation Analysis. In Proceedings of the 30th International Conference on Machine Learning, ICML 2013 (JMLR Workshop and Conference Proceedings, Vol. 28) . JMLR.org, Atlanta, GA, USA, 1247–1255
Galen Andrew, Raman Arora, Jeff A. Bilmes, and Karen Livescu. 2013 · 2013
Earlier work this paper cites.
COVAREP - A collaborative voice analysis repository for speech technologies. In IEEE International Conference on Acoustics, Speech and Signal Processing, ICASSP 2014 . IEEE, Florence, Italy, 960–964
Gilles Degottex, John Kane, Thomas Drugman, Tuomo Raitio, and Stefan Scherer. 2014 · 2014
Earlier work this paper cites.
GloVe: Global Vectors for Word Representation. In Proceedings of the 2014 Conference on Empirical Methods in Natural Language Processing (EMNLP) . Association for Computational Linguistics, Doha, Qatar, 1532–1543
Jeffrey Pennington, Richard Socher, and Christopher Manning. 2014 · 2014
Earlier work this paper cites.
Deep Convolutional Neural Network Textual Features and Multiple Kernel Learning for Utterance-level Multimodal Sentiment Analysis. In Proceedings of the 2015 Conference on Empirical Methods in Natural Language Processing . Association for Computational Linguistics, Lisbon, Portugal, 2539–2544
Soujanya Poria, Erik Cambria, and Alexander Gelbukh. 2015 · 2015
Earlier work this paper cites.
Simultaneous Deep Transfer Across Domains and Tasks. In Proceedings of the 2015 IEEE International Conference on Computer Vision (ICCV) (ICCV ’15) . IEEE Computer Society, USA, 4068–4076
Eric Tzeng, Judy Hoffman, Trevor Darrell, and Kate Saenko. 2015 · 2015
Earlier work this paper cites.
OpenFace: An open source facial behavior analysis toolkit. In 2016 IEEE Winter Conference on Applications of Computer Vision, WACV 2016 . IEEE Computer Society, Lake Placid, NY, USA, 1–10
Tadas Baltrusaitis, Peter Robinson, and Louis-Philippe Morency. 2016 · 2016
Earlier work this paper cites.
Domain Separation Networks. In Advances in Neural Information Processing Systems 29 . Curran Associates, Inc., Barcelona, Spain, 343–351
Konstantinos Bousmalis, George Trigeorgis, Nathan Silberman, Dilip Krishnan, and Dumitru Erhan. 2016 · 2016
Cited alongside, same era.
Multimodal Compact Bilinear Pooling for Visual Question Answering and Visual Grounding. In Proceedings of the 2016 Conference on Empirical Methods in Natural Language Processing . Association for Computational Linguistics, Austin, Texas, 457–468
Akira Fukui, Dong Huk Park, Daylen Yang, Anna Rohrbach, Trevor Darrell, and Marcus Rohrbach. 2016 · 2016
Cited alongside, same era.
Image-Text Multi-Modal Representation Learning by Adversarial Backpropagation
Gwangbeen Park and Woobin Im. 2016 · 2016
Cited alongside, same era.
Extending Long Short-Term Memory for Multi-View Structured Learning. In Computer Vision - ECCV 2016 - 14th European Conference, Proceedings, Part VII (Lecture Notes in Computer Science, Vol. 9911) . Springer, Amsterdam, The Netherlands, 338–353
Shyam Sundar Rajagopalan, Louis-Philippe Morency, Tadas Baltrusaitis, and Roland Goecke. 2016 · 2016
Seq2Seq2Sentiment: Multimodal Sequence to Sequence Models for Sentiment Analysis. In Proceedings of Grand Challenge and Workshop on Human Multimodal Language (Challenge-HML) . Association for Computational Linguistics, Melbourne, Australia, 53–63
Hai Pham, Thomas Manzini, Paul Pu Liang, and Barnabás Poczós. 2018 · 2018
Later among the works it cites.
Strong Baselines for Neural Semi-Supervised Learning under Domain Shift. In Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics, ACL 2018, Volume 1: Long Papers . Association for Computational Linguistics, Melbourne, Australia, 1044–1054
Sebastian Ruder and Barbara Plank. 2018 · 2018
Later among the works it cites.
Multimodal Language Analysis in the Wild: CMU-MOSEI Dataset and Interpretable Dynamic Fusion Graph. In Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics, ACL 2018, Volume 1: Long Papers . Association for Computational Linguistics, Melbourne, Australia, 2236–2246
Amir Zadeh, Paul Pu Liang, Soujanya Poria, Erik Cambria, and Louis-Philippe Morency. 2018b · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Deep Variational Canonical Correlation Analysis
Weiran Wang, Honglak Lee, and Karen Livescu. 2016 · 2016
Cited alongside, same era.
Multimodal Sentiment Intensity Analysis in Videos: Facial Gestures and Verbal Messages
Amir Zadeh, Rowan Zellers, Eli Pincus, and Louis-Philippe Morency. 2016 · 2016
Cited alongside, same era.
Multimodal Sentiment Analysis with Word-Level Fusion and Reinforcement Learning. In Proceedings of the 19th ACM International Conference on Multimodal Interaction (Glasgow, UK) (ICMI ’17) . Association for Computing Machinery, New York, NY, USA, 163–171
Minghai Chen, Sen Wang, Paul Pu Liang, Tadas Baltrušaitis, Amir Zadeh, and Louis-Philippe Morency. 2017 · 2017
Cited alongside, same era.
Attribute-Enhanced Face Recognition with Neural Tensor Fusion Networks. In IEEE International Conference on Computer Vision, ICCV 2017 . IEEE Computer Society, Venice, Italy, 3764–3773
Guosheng Hu, Yang Hua, Yang Yuan, Zhihong Zhang, Zheng Lu, Sankha S. Mukherjee, Timothy M. Hospedales, Neil Martin Robertson, and Yongxin Yang. 2017 · 2017
Cited alongside, same era.
Adversarial Multi-task Learning for Text Classification. In Proceedings of the 55th Annual Meeting of the Association for Computational Linguistics, ACL 2017, Volume 1: Long Papers . Association for Computational Linguistics, Vancouver, Canada, 1–10
Pengfei Liu, Xipeng Qiu, and Xuanjing Huang. 2017 · 2017
Cited alongside, same era.
A review of affective computing: From unimodal analysis to multimodal fusion
Soujanya Poria, Erik Cambria, Rajiv Bajpai, and Amir Hussain. 2017a · 2017
Cited alongside, same era.
Multi-level Multiple Attentions for Contextual Multimodal Sentiment Analysis. In 2017 IEEE International Conference on Data Mining, ICDM 2017 . IEEE Computer Society, New Orleans, LA, USA, 1033–1038
Soujanya Poria, Erik Cambria, Devamanyu Hazarika, Navonil Majumder, Amir Zadeh, and Louis-Philippe Morency. 2017b · 2017
Cited alongside, same era.
Attention is All you Need. In Advances in Neural Information Processing Systems 30 . Curran Associates, Inc., Long Beach, CA, USA, 5998–6008
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Cited alongside, same era.
Multi-task Learning for Multi-modal Emotion Recognition and Sentiment Analysis. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, NAACL-HLT 2019, Volume 1 (Long and Short Papers) . Association for Computational Linguistics, Minneapolis, MN, USA, 370–379
Md. Shad Akhtar, Dushyant Singh Chauhan, Deepanway Ghosal, Soujanya Poria, Asif Ekbal, and Pushpak Bhattacharyya. 2019 · 2019
Later among the works it cites.
Context-aware Interactive Attention for Multi-modal Sentiment and Emotion Analysis. In Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP) . Association for Computational Linguistics, Hong Kong, China, 5647–5657
Dushyant Singh Chauhan, Md Shad Akhtar, Asif Ekbal, and Pushpak Bhattacharyya. 2019 · 2019
Later among the works it cites.
Complementary Fusion of Multi-Features and Multi-Modalities in Sentiment Analysis
Feiyang Chen, Ziqian Luo, Yanyan Xu, and Dengfeng Ke. 2019 · 2019
Later among the works it cites.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding. In Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers) . Association for Computational Linguistics, Minneapolis, Minnesota, 4171–4186
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Later among the works it cites.
Deep Multimodal Representation Learning: A Survey
Wenzhong Guo, Jianwen Wang, and Shiping Wang. 2019 · 2019
Later among the works it cites.
Supervised Multimodal Bitransformers for Classifying Images and Text. In Visually Grounded Interaction and Language (ViGIL), NeurIPS 2019 Workshop . Curran Associates, Inc., Vancouver, Canada
Douwe Kiela, Suvrat Bhooshan, Hamed Firooz, and Davide Testuggine. 2019 · 2019
Later among the works it cites.
ViLBERT: Pretraining Task-Agnostic Visiolinguistic Representations for Vision-and-Language Tasks. In Advances in Neural Information Processing Systems 32 . Curran Associates, Inc., Vancouver, BC, Canada, 13–23
Jiasen Lu, Dhruv Batra, Devi Parikh, and Stefan Lee. 2019 · 2019
Later among the works it cites.
Divide, Conquer and Combine: Hierarchical Feature Fusion Network with Local and Global Perspectives for Multimodal Affective Computing. In Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics . Association for Computational Linguistics, Florence, Italy, 481–492
Sijie Mai, Haifeng Hu, and Songlong Xing. 2019 · 2019
Later among the works it cites.
Locally Confined Modality Fusion Network With a Global Perspective for Multimodal Human Affective Computing
Sijie Mai, Songlong Xing, and Haifeng Hu. 2020b · 2019
Later among the works it cites.
CM-GANs: Cross-Modal Generative Adversarial Networks for Common Representation Learning
Yuxin Peng and Jinwei Qi. 2019 · 2019
Later among the works it cites.
Found in Translation: Learning Robust Joint Representations by Cyclic Translations between Modalities. In The Thirty-Third AAAI Conference on Artificial Intelligence . AAAI Press, Honolulu, Hawaii, 6892–6899
Hai Pham, Paul Pu Liang, Thomas Manzini, Louis-Philippe Morency, and Barnabás Póczos. 2019 · 2019
Later among the works it cites.
Learning Factorized Multimodal Representations. In 7th International Conference on Learning Representations, ICLR 2019 . OpenReview.net, New Orleans, LA, USA
Yao-Hung Hubert Tsai, Paul Pu Liang, Amir Zadeh, Louis-Philippe Morency, and Ruslan Salakhutdinov. 2019b · 2019
Later among the works it cites.
Words Can Shift: Dynamically Adjusting Word Representations Using Nonverbal Behaviors. In The Thirty-Third AAAI Conference on Artificial Intelligence . AAAI Press, Honolulu, Hawaii, 7216–7223
Yansen Wang, Ying Shen, Zhun Liu, Paul Pu Liang, Amir Zadeh, and Louis-Philippe Morency. 2019 · 2019
Later among the works it cites.
Modality to Modality Translation: An Adversarial Representation Learning and Graph Fusion Network for Multimodal Fusion. In The Thirty-Fourth AAAI Conference on Artificial Intelligence, AAAI 2020 . AAAI Press, New York, NY, USA, 164–172
Sijie Mai, Haifeng Hu, and Songlong Xing. 2020a · 2020
Closest in time.
Multimodal Sentiment Analysis Based on Multi-Head Attention Mechanism. In Proceedings of the 4th International Conference on Machine Learning and Soft Computing (Haiphong City, Viet Nam) (ICMLSC 2020) . Association for Computing Machinery, New York, NY, USA, 34–39
Chen Xi, Guanming Lu, and Jingjie Yan. 2020 · 2020
Closest in time.
UR-FUNNY: A Multimodal Language Dataset for Understanding Humor. In Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP) . Association for Computational Linguistics, Hong Kong, China, 2046–2056
Md Kamrul Hasan, Wasifur Rahman, AmirAli Bagher Zadeh, Jianyuan Zhong, Md Iftekhar Tanveer, Louis-Philippe Morency, and Mohammed (Ehsan) Hoque. 2019 · 2056
Closest in time.