Fetching the paper…
Reading the bibliography…
Multimodal emotion recognition aims to recognize emotions for each utterance of multiple modalities, which has received increasing attention for its application in human-machine interaction.
Roberta: A robustly optimized BERT pretraining approach
Yinhan Liu, Myle Ott, Naman Goyal, Jingfei Du, Mandar Joshi, Danqi Chen, Omer Levy, Mike Lewis, Luke Zettlemoyer, and Veselin Stoyanov. 2019 · 1907
Earlier work this paper cites.
A new status index derived from sociometric analysis
Leo Katz. 1953 · 1953
Earlier work this paper cites.
Contributions to probability and statistics
Howard Levene et al. 1960 · 1960
Earlier work this paper cites.
An analysis of variance test for normality (complete samples)
Samuel Sanford Shapiro and Martin B Wilk. 1965 · 1965
Earlier work this paper cites.
Levene’s test for relative variation
Brian B Schultz. 1985 · 1985
Earlier work this paper cites.
Controlling the false discovery rate: a practical and powerful approach to multiple testing
Yoav Benjamini and Yosef Hochberg. 1995 · 1995
Earlier work this paper cites.
Convolutional networks for images, speech, and time series
Yann LeCun, Yoshua Bengio, et al. 1995 · 1995
Earlier work this paper cites.
Communication goes multimodal
Sarah Partan and Peter Marler. 1999 · 1999
Earlier work this paper cites.
A re-examination of text categorization methods
Yiming Yang, Xin Liu, and et al. 1999 · 1999
Earlier work this paper cites.
IEMOCAP: interactive emotional dyadic motion capture database
Carlos Busso, Murtaza Bulut, Chi-Chun Lee, Abe Kazemzadeh, Emily Mower, Samuel Kim, Jeannette N. Chang, Sungbok Lee, and Shrikanth S. Narayanan. 2008 · 2008
Earlier work this paper cites.
The graph neural network model
Franco Scarselli, Marco Gori, Ah Chung Tsoi, Markus Hagenbuchner, and Gabriele Monfardini. 2009 · 2009
Earlier work this paper cites.
Opensmile: The munich versatile and fast open-source audio feature extractor
Florian Eyben, Martin Wöllmer, and Björn Schuller. 2010 · 2010
Earlier work this paper cites.
Convolutional neural networks for sentence classification
Yoon Kim. 2014 · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba. 2015 · 2015
Earlier work this paper cites.
Multimodal sentiment intensity analysis in videos: Facial gestures and verbal messages
Amir Zadeh, Rowan Zellers, Eli Pincus, and Louis-Philippe Morency. 2016 · 2016
Earlier work this paper cites.
Inductive representation learning on large graphs
William L. Hamilton, Zhitao Ying, and Jure Leskovec. 2017 · 2017
Earlier work this paper cites.
Densely connected convolutional networks
Gao Huang, Zhuang Liu, Laurens van der Maaten, and Kilian Q. Weinberger. 2017 · 2017
Earlier work this paper cites.
Semi-supervised classification with graph convolutional networks
Thomas N. Kipf and Max Welling. 2017 · 2017
Earlier work this paper cites.
Tensor fusion network for multimodal sentiment analysis
Amir Zadeh, Minghai Chen, Soujanya Poria, Erik Cambria, and Louis-Philippe Morency. 2017 · 2017
Earlier work this paper cites.
Openface 2.0: Facial behavior analysis toolkit
Tadas Baltrusaitis, Amir Zadeh, Yao Chong Lim, and Louis-Philippe Morency. 2018 · 2018
Earlier work this paper cites.
The hitchhiker’s guide to testing statistical significance in natural language processing
Rotem Dror, Gili Baumer, Segev Shlomov, and Roi Reichart. 2018 · 2018
Earlier work this paper cites.
Adafactor: Adaptive learning rates with sublinear memory cost
Noam Shazeer and Mitchell Stern. 2018 · 2018
Earlier work this paper cites.
Multimodal language analysis in the wild: CMU-MOSEI dataset and interpretable dynamic fusion graph
Amir Zadeh, Paul Pu Liang, Soujanya Poria, Erik Cambria, and Louis-Philippe Morency. 2018 · 2018
Earlier work this paper cites.
BERT: pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Earlier work this paper cites.
DialogueGCN: A graph convolutional neural network for emotion recognition in conversation
Deepanway Ghosal, Navonil Majumder, Soujanya Poria, Niyati Chhaya, and Alexander Gelbukh. 2019 · 2019
Earlier work this paper cites.
Dialoguernn: An attentive RNN for emotion detection in conversations
Navonil Majumder, Soujanya Poria, Devamanyu Hazarika, Rada Mihalcea, Alexander F. Gelbukh, and Erik Cambria. 2019 · 2019
Earlier work this paper cites.
Librosa based assessment tool for music information retrieval systems
Preeth Raguraman, Mohan Ramasundaram, and Midhula Vijayan. 2019 · 2019
Earlier work this paper cites.
Sentence-BERT: Sentence embeddings using Siamese BERT-networks
Nils Reimers and Iryna Gurevych. 2019 · 2019
Cited alongside, same era.
Words can shift: Dynamically adjusting word representations using nonverbal behaviors
Yansen Wang, Ying Shen, Zhun Liu, Paul Pu Liang, Amir Zadeh, and Louis-Philippe Morency. 2019 · 2019
Cited alongside, same era.
MMGCN: multi-modal graph convolution network for personalized recommendation of micro-video
Yinwei Wei, Xiang Wang, Liqiang Nie, Xiangnan He, Richang Hong, and Tat-Seng Chua. 2019 · 2019
Cited alongside, same era.
A transformer-based joint-encoding for emotion recognition and sentiment analysis
Jean-Benoit Delbrouck, Noé Tits, Mathilde Brousmiche, and Stéphane Dupont. 2020 · 2020
Cited alongside, same era.
COSMIC: COmmonSense knowledge for eMotion identification in conversations
Deepanway Ghosal, Navonil Majumder, Alexander Gelbukh, Rada Mihalcea, and Soujanya Poria. 2020 · 2020
Cited alongside, same era.
A contextual attention network for multimodal emotion recognition in conversation
Tana Wang, Yaqing Hou, Dongsheng Zhou, and Qiang Zhang. 2021 · 2021
Later among the works it cites.
Multimodal fusion with co-attention networks for fake news detection
Yang Wu, Pengwei Zhan, Yunjian Zhang, LiMing Wang, and Zhen Xu. 2021 · 2021
Later among the works it cites.
Infogcl: Information-aware graph contrastive learning
Dongkuan Xu, Wei Cheng, Dongsheng Luo, Haifeng Chen, and Xiang Zhang. 2021 · 2021
Later among the works it cites.
MTAG: modal-temporal attention graph for unaligned human multimodal language sequences
Jianing Yang, Yongxin Wang, Ruitao Yi, Yuying Zhu, Azaan Rehman, Amir Zadeh, Soujanya Poria, and Louis-Philippe Morency. 2021 · 2021
Later among the works it cites.
Wenmeng Yu, Hua Xu, Ziqi Yuan, and Jiele Wu. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Misa: Modality-invariant and -specific representations for multimodal sentiment analysis
Devamanyu Hazarika, Roger Zimmermann, and Soujanya Poria. 2020 · 2020
Cited alongside, same era.
Relation-aware graph attention networks with relational position encodings for emotion recognition in conversations
Taichi Ishiwatari, Yuki Yasuda, Taro Miyazaki, and Jun Goto. 2020 · 2020
Cited alongside, same era.
Real-time emotion recognition via attention gated hierarchical memory network
Wenxiang Jiao, Michael R. Lyu, and Irwin King. 2020 · 2020
Cited alongside, same era.
Multistage fusion with forget gate for multimodal summarization in open-domain videos
Nayu Liu, Xian Sun, Hongfeng Yu, Wenkai Zhang, and Guangluan Xu. 2020 · 2020
Cited alongside, same era.
Influence of multiple hypothesis testing on reproducibility in neuroimaging research: A simulation study and python-based software
Tuomas Puoliväli, Satu Palva, and J. Matias Palva. 2020 · 2020
Cited alongside, same era.
Integrating multimodal information in large pretrained transformers
Wasifur Rahman, Md. Kamrul Hasan, Sangwu Lee, AmirAli Bagher Zadeh, Chengfeng Mao, Louis-Philippe Morency, and Mohammed E. Hoque. 2020 · 2020
Cited alongside, same era.
Summarize before aggregate: A global-to-local heterogeneous graph inference network for conversational emotion recognition
Dongming Sheng, Dong Wang, Ying Shen, Haitao Zheng, and Haozhuang Liu. 2020 · 2020
Cited alongside, same era.
Contrastive self-supervised learning for graph classification
Jiaqi Zeng and Pengtao Xie. 2021 · 2021
Later among the works it cites.
Multi-modal multi-label emotion recognition with heterogeneous hierarchical message passing
Dong Zhang, Xincheng Ju, Wei Zhang, Junhui Li, Shoushan Li, Qiaoming Zhu, and Guodong Zhou. 2021 · 2021
Later among the works it cites.
3massiv: Multilingual, multimodal and multi-aspect dataset of social media short videos
Vikram Gupta, Trisha Mittal, Puneet Mathur, Vaibhav Mishra, Mayank Maheshwari, Aniket Bera, Debdoot Mukherjee, and Dinesh Manocha. 2022 · 2022
Later among the works it cites.
COGMEN: COntextualized GNN based multimodal emotion recognitioN
Abhinav Joshi, Ashwani Bhat, Ayush Jain, Atin Singh, and Ashutosh Modi. 2022 · 2022
Later among the works it cites.
Multimodal dialogue state tracking
Hung Le, Nancy Chen, and Steven Hoi. 2022 · 2022
Later among the works it cites.
Enhanced knowledge selection for grounded dialogues via document semantic graphs
Sha Li, Madhi Namazifar, Di Jin, MOHIT BANSAL, Heng Ji, Yang Liu, and Dilek Hakkani-Tur. 2022b · 2022
Later among the works it cites.
Modular and parameter-efficient multimodal fusion with prompting
Sheng Liang, Mengjie Zhao, and Hinrich Schuetze. 2022 · 2022
Later among the works it cites.
Vision-language pre-training for multimodal aspect-based sentiment analysis
Yan Ling, Jianfei Yu, and Rui Xia. 2022 · 2022
Later among the works it cites.
M-SENA: An integrated platform for multimodal sentiment analysis
Huisheng Mao, Ziqi Yuan, Hua Xu, Wenmeng Yu, Yihe Liu, and Kai Gao. 2022 · 2022
Later among the works it cites.
Is cross-attention preferable to self-attention for multi-modal emotion recognition?
Vandana Rajan, Alessio Brutti, and Andrea Cavallaro. 2022 · 2022
Later among the works it cites.
Multimodal fusion via cortical network inspired losses
Shiv Shankar. 2022 · 2022
Later among the works it cites.
Sentiment and emotion-aware multi-modal complaint identification
Apoorva Singh, Soumyodeep Dey, Anamitra Singha, and Sriparna Saha. 2022 · 2022
Later among the works it cites.
Rumor detection on social media with graph adversarial contrastive learning
Tiening Sun, Zhong Qian, Sujun Dong, Peifeng Li, and Qiaoming Zhu. 2022 · 2022
Later among the works it cites.
Temporality- and frequency-aware graph contrastive learning for temporal network
Shiyin Tan, Jingyi You, and Dongyuan Li. 2022 · 2022
Later among the works it cites.
Modern question answering datasets and benchmarks: A survey
Zhen Wang. 2022 · 2022
Later among the works it cites.
JPG - jointly learn to align: Automated disease prediction and radiology report generation
Jingyi You, Dongyuan Li, Manabu Okumura, and Kenji Suzuki. 2022 · 2022
Later among the works it cites.
COSTA: covariance-preserving feature augmentation for graph contrastive learning
Yifei Zhang, Hao Zhu, Zixing Song, Piotr Koniusz, and Irwin King. 2022 · 2022
Later among the works it cites.
Summary-oriented vision modeling for multimodal abstractive summarization
Yunlong Liang, Fandong Meng, Jinan Xu, Jiaan Wang, Yufeng Chen, and Jie Zhou. 2023 · 2023
Closest in time.
e-health CSIRO at radsum23: Adapting a chest x-ray report generator to multimodal radiology report summarisation
Aaron Nicolson, Jason Dowling, and Bevan Koopman. 2023 · 2023
Closest in time.
Retrieving multimodal prompts for generative visual question answering
Timothy Ossowski and Junjie Hu. 2023 · 2023
Closest in time.
Emp: Emotion-guided multi-modal fusion and contrastive learning for personality traits recognition
Yusong Wang, Dongyuan Li, Kotaro Funakoshi, and Manabu Okumura. 2023 · 2023
Closest in time.
Bidirectional transformer reranker for grammatical error correction
Ying Zhang, Hidetaka Kamigaito, and Manabu Okumura. 2023 · 2023
Closest in time.