Fetching the paper…
Reading the bibliography…
In the field of multimodal sentiment analysis (MSA), a few studies have leveraged the inherent modality correlation information stored in samples for self-supervised learning.
Lxmert: Learning cross-modality encoder representations from transformers
Hao Tan and Mohit Bansal. 2019 · 1908
Earlier work this paper cites.
From utterance to text: The bias of language in speech and writing
David Olson. 1977 · 1977
Earlier work this paper cites.
Multimodal routing: Improving local and global interpretability of multimodal language analysis
Yao-Hung Hubert Tsai, Martin Q. Ma, Muqiao Yang, Ruslan Salakhutdinov, and Louis-Philippe Morency. 2020 · 2001
Earlier work this paper cites.
Aman Shenoy and Ashish Sardana. 2020 · 2002
Earlier work this paper cites.
Cobra: Contrastive bi-modal representation algorithm
Vishaal Udandarao, Abhishek Maiti, Deepak Srivatsav, S. R. Vyalla, Yifang Yin, and R. Shah. 2020 · 2005
Earlier work this paper cites.
Curriculum learning
Yoshua Bengio, Jérôme Louradour, Ronan Collobert, and Jason Weston. 2009 · 2009
Earlier work this paper cites.
Distance metric learning for large margin nearest neighbor classification
Kilian Q Weinberger and Lawrence K Saul. 2009 · 2009
Earlier work this paper cites.
Youtube movie reviews: Sentiment analysis in an audio-visual context
Martin Wollmer, Felix Weninger, Tobias Knaup, Bjorn Schuller, Congkai Sun, Kenji Sagae, and Louis Philippe Morency. 2013 · 2013
Earlier work this paper cites.
Covarep: A collaborative voice analysis repository for speech technologies
Gilles Degottex, John Kane, Thomas Drugman, Tuomo Raitio, and Stefan Scherer. 2014 · 2014
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P Kingma and Jimmy Ba. 2015 · 2015
Earlier work this paper cites.
Deep multimodal fusion for persuasiveness prediction
Behnaz Nojavanasghari, Deepak Gopinath, Jayanth Koushik, Tadas Baltrušaitis, and Louis-Philippe Morency. 2016 · 2016
Earlier work this paper cites.
Convolutional mkl based multimodal emotion recognition and sentiment analysis
Soujanya Poria, Iti Chaturvedi, Erik Cambria, and Amir Hussain. 2016 · 2016
Earlier work this paper cites.
Multimodal sentiment intensity analysis in videos: Facial gestures and verbal messages
Amir Zadeh, Rowan Zellers, Eli Pincus, and Louis Philippe Morency. 2016 · 2016
Earlier work this paper cites.
Curriculum learning for facial expression recognition
Liangke Gui, Tadas Baltrušaitis, and Louis-Philippe Morency. 2017 · 2017
Earlier work this paper cites.
The kinetics human action video dataset
Will Kay, Joao Carreira, Karen Simonyan, Brian Zhang, Chloe Hillier, Sudheendra Vijayanarasimhan, Fabio Viola, Tim Green, Trevor Back, Paul Natsev, et al. 2017 · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
Tensor fusion network for multimodal sentiment analysis
Amir Zadeh, Minghai Chen, Soujanya Poria, Erik Cambria, and Louis Philippe Morency. 2017 · 2017
Earlier work this paper cites.
Audio-visual fusion for sentiment classification using cross-modal autoencoder
Sri Harsha Dumpala, Imran Sheikh, Rupayan Chakraborty, and Sunil Kumar Kopparapu. 2019 · 2018
Earlier work this paper cites.
Investigating audio, visual, and text fusion methods for end-to-end automatic personality prediction
Onno Kampman, Elham J. Barezi, Dario Bertero, and Pascale Fung. 2018 · 2018
Cited alongside, same era.
Efficient low-rank multimodal fusion with modality-specific factors
Zhun Liu, Ying Shen, Paul Pu Liang, Amir Zadeh, and Louis Philippe Morency. 2018 · 2018
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
On the power of curriculum learning in training deep networks
Guy Hacohen and Daphna Weinshall. 2019 · 2019
Cited alongside, same era.
Learning representations from imperfect time series data via tensor rank regularization
Paul Pu Liang, Zhun Liu, Yao-Hung Hubert Tsai, Qibin Zhao, Ruslan Salakhutdinov, and Louis-Philippe Morency. 2019 · 2019
Cited alongside, same era.
Noise estimation using density estimation for self-supervised multimodal learning
Elad Amrani, Rami Ben-Ari, Daniel Rotman, and Alex Bronstein. 2021 · 2021
Later among the works it cites.
Geometric multimodal deep learning with multi-scaled graph wavelet convolutional network
Maysam Behmanesh, Peyman Adibi, Mohammad Saeed Ehsani, and Jocelyn Chanussot. 2021 · 2021
Later among the works it cites.
A variational information bottleneck approach to multi-omics data integration
Changhee Lee and Mihaela Schaar. 2021 · 2021
Later among the works it cites.
Quantum-inspired multimodal fusion for video sentiment analysis
Qiuchi Li, Dimitris Gkoumas, Christina Lioma, and Massimo Melucci. 2021 · 2021
Later among the works it cites.
Competence-based multimodal curriculum learning for medical report generation
Fenglin Liu, Shen Ge, and Xian Wu. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks
Jiasen Lu, Dhruv Batra, Devi Parikh, and Stefan Lee. 2019 · 2019
Cited alongside, same era.
Multimodal transformer for unaligned multimodal language sequences
Yao-Hung Hubert Tsai, Shaojie Bai, Paul Pu Liang, J. Zico Kolter, Louis-Philippe Morency, and Ruslan Salakhutdinov. 2019 · 2019
Cited alongside, same era.
Words can shift: Dynamically adjusting word representations using nonverbal behaviors
Yansen Wang, Ying Shen, Zhun Liu, Paul Pu Liang, Amir Zadeh, and Louis-Philippe Morency. 2019 · 2019
Cited alongside, same era.
Xlnet: Generalized autoregressive pretraining for language understanding
Z. Yang, Zihang Dai, Yiming Yang, J. Carbonell, R. Salakhutdinov, and Quoc V. Le. 2019 · 2019
Cited alongside, same era.
Large-scale adversarial training for vision-and-language representation learning
Zhe Gan, Yen-Chun Chen, Linjie Li, Chen Zhu, Yu Cheng, and Jingjing Liu. 2020 · 2020
Cited alongside, same era.
Misa: Modality-invariant and -specific representations for multimodal sentiment analysis
Devamanyu Hazarika, R. Zimmermann, and Soujanya Poria. 2020 · 2020
Cited alongside, same era.
Modality to modality translation: An adversarial representation learning and graph fusion network for multimodal fusion
Sijie Mai, Haifeng Hu, and Songlong Xing. 2020 · 2020
Cited alongside, same era.
Sijie Mai, Ying Zeng, Shuangjia Zheng, and Haifeng Hu. 2021 · 2021
Later among the works it cites.
A survey on curriculum learning
Xin Wang, Yudong Chen, and Wenwu Zhu. 2021 · 2021
Later among the works it cites.
Graph capsule aggregation for unaligned multimodal sequences
Jianfeng Wu, Sijie Mai, and Haifeng Hu. 2021 · 2021
Later among the works it cites.
Transformer-based feature reconstruction network for robust multimodal sentiment analysis
Ziqi Yuan, Wei Li, Hua Xu, and Wenmeng Yu. 2021 · 2021
Later among the works it cites.
Multimodal affective states recognition based on multiscale cnns and biologically inspired decision fusion model
Yuxuan Zhao, Xinyan Cao, Jinlong Lin, Dunshan Yu, and Xixin Cao. 2021 · 2021
Later among the works it cites.
Robust contrastive learning against noisy views
Ching-Yao Chuang, R Devon Hjelm, Xin Wang, Vibhav Vineet, Neel Joshi, Antonio Torralba, Stefanie Jegelka, and Yale Song. 2022 · 2022
Closest in time.
Excavating multimodal correlation for representation learning
Sijie Mai, Ya Sun, Ying Zeng, and Haifeng Hu. 2022 · 2022
Closest in time.
Tricolo: Trimodal contrastive loss for fine-grained text to shape retrieval
Yue Ruan, Han-Hung Lee, Ke Zhang, and Angel X Chang. 2022 · 2022
Closest in time.
Multimodal fusion via cortical network inspired losses
Shiv Shankar. 2022 · 2022
Closest in time.
Curriculum learning: A survey
Petru Soviany, Radu Tudor Ionescu, Paolo Rota, and Nicu Sebe. 2022 · 2022
Closest in time.
Learning to learn better unimodal representations via adaptive multimodal meta-learning
Ya Sun, Sijie Mai, and Haifeng Hu. 2022 · 2022
Closest in time.
Hybrid curriculum learning for emotion recognition in conversation
Lin Yang, Yi Shen, Yue Mao, and Longjun Cai. 2022 · 2022
Closest in time.