Fetching the paper…
Reading the bibliography…
Compared to traditional sentiment analysis, which only considers text, multimodal sentiment analysis needs to consider emotional signals from multimodal sources simultaneously and is therefore more consistent with the way how humans process sentiment in real-world scenarios.
Hatzivassiloglou V, McKeown KR. Predicting the semantic orientation of adjectives. In
1997
Earlier work this paper cites.
Paul Mc Kevitt. MultiModal Semantic Representation[C]// Tilburg University
2003
Earlier work this paper cites.
Tian Y L, Kanade T, Cohn J F. Facial expression analysis[M]//Handbook of face recognition. Springer, New York, NY, 2005: 247-275
2005
Earlier work this paper cites.
Feng J, Lin M, Shang L, Gao X. Autonomous Aspect-Image Instruction a2II: Q-Former Guided Multimodal Sentiment Classification. InProceedings of the 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation (LREC-COLING 2024) 2024 May (pp. 1996-2005)
2005
Earlier work this paper cites.
LeCun Y, Chopra S, Hadsell R, Ranzato M, Huang F. A tutorial on energy-based learning. Predicting structured data. 2006 Aug 19;1(0)
2006
Earlier work this paper cites.
Cheang H S, Pell M D. The Sound of Sarcasm[J]. Speech Communication, 2008, 50(5): 366-381
2008
Earlier work this paper cites.
Busso C, Bulut M, Lee CC, Kazemzadeh A, Mower E, Kim S, Chang JN, Lee S, Narayanan SS. IEMOCAP: Interactive emotional dyadic motion capture database. Language resources and evaluation. 2008 Dec;42:335-59
2008
Earlier work this paper cites.
M. Thelwall, K. Buckley, G. Paltoglou, D. Cai, and A. Kappas, ”Sentiment strength detection in short informal text,” Journal of the American Society for Information Science and Technology, vol. 61, pp. 2544-2558, 2010
2010
Earlier work this paper cites.
Xiao-Wei Wang, Dan Nie, and Bao-Liang Lu. Eeg-based emotion recognition using frequency domain features and support vector machines. In International conference on neural information processing, pages 734-743. Springer, 2011
2011
Earlier work this paper cites.
Jennifer Woodland and Daniel Voyer. Context and intonation in the perception of sarcasm. Metaphor and Symbol. 2011, 26(3):227-239
2011
Earlier work this paper cites.
Morency LP, Mihalcea R, Doshi P. Towards multimodal sentiment analysis: Harvesting opinions from the web. InProceedings of the 13th international conference on multimodal interfaces 2011 Nov 14 (pp. 169-176)
2011
Earlier work this paper cites.
Morency LP, Mihalcea R, Doshi P. Towards multimodal sentiment analysis: Harvesting opinions from the web. InProceedings of the 13th international conference on multimodal interfaces 2011 Nov 14 (pp. 169-176)
2011
Earlier work this paper cites.
Richard Socher, Alex Perelygin, Jean Wu, Jason Chuang, Christopher D. Manning, Andrew Y. Ng, and Christopher Potts. 2013. Recursive deep models for semantic compositionality over a sentiment treebank. In
2013
Earlier work this paper cites.
Fatemeh Bahari and Amin Janghorbani. Eeg-based emotion recognition using recurrence plot analysis and k nearest neighbor classifier. In
2013
Earlier work this paper cites.
Riloff E, Qadir A, Surve P, et al. Sarcasm as Contrast between a Positive Sentiment and Negative Situation[C]//
2013
Earlier work this paper cites.
Mesnil G, He X, Deng L, et al. Investigation of recurrent-neural-network architectures and learning methods for spoken language understanding[C]//Interspeech. 2013: 3771-3775
2013
Earlier work this paper cites.
Mitchell M, Aguilar J, Wilson T, et al. Open domain targeted sentiment[C]//Proceedings of the 2013 Conference on Empirical Methods in Natural Language Processing. 2013: 1643-1654
2013
Earlier work this paper cites.
D. Borth, R. Ji, T. Chen, T. Breuel, and S.-F. Chang, ”Large-scale visual sentiment ontology and detectors using adjective noun pairs,” in Proceedings of the 21st ACM international conference on Multimedia, 2013, pp. 223-232
2013
Earlier work this paper cites.
2014
Earlier work this paper cites.
Sutskever I, Vinyals O, Le Q V. Sequence to sequence learning with neural networks[J]. Advances in neural information processing systems, 2014, 27
2014
Earlier work this paper cites.
Degottex G, Kane J, Drugman T, et al. COVAREP-A collaborative voice analysis repository for speech technologies[C]//2014 ieee international conference on acoustics, speech and signal processing (icassp). IEEE, 2014: 960-964
2014
Earlier work this paper cites.
M. Wang, D. Cao, L. Li, S. Li, and R. Ji, ”Microblog sentiment analysis based on cross-media bag-of-words model,” in Proceedings of international conference on internet multimedia computing and service, 2014, p. 76
2014
Earlier work this paper cites.
2015
Earlier work this paper cites.
Liu P, Joty S, Meng H. Fine-grained opinion mining with recurrent neural networks and word embeddings[C]//Proceedings of the 2015 conference on empirical methods in natural language processing. 2015: 1433-1443
2015
Earlier work this paper cites.
Vinyals O, Toshev A, Bengio S, et al. Show and tell: A neural image caption generator[C]//Proceedings of the IEEE conference on computer vision and pattern recognition. 2015: 3156-3164
2015
Earlier work this paper cites.
G. Cai and B. Xia, ”Convolutional Neural Networks for Multimedia Sentiment Analysis,” in National CCF Conference on Natural Language Processing and Chinese Computing, 2015, pp. 159-167
2015
Earlier work this paper cites.
Cao D, Ji R, Lin D, et al. A Cross-media Public Sentiment Analysis System for Microblog[J]
2016
Earlier work this paper cites.
You Q, Cao L, Jin H, et al. Robust Visual-textual Sentiment Analysis: When Attention Meets Tree-structured Recursive Neural Networks[C]// In
2016
Earlier work this paper cites.
Cao D, Ji R, Lin D, et al. A Cross-media Public Sentiment Analysis System for Microblog[J]
2016
Earlier work this paper cites.
Amos B, Ludwiczuk B, Satyanarayanan M. Openface: A general-purpose face recognition library with mobile applications[J]. CMU School of Computer Science, 2016, 6(2): 20
2016
Earlier work this paper cites.
Chen, Long, et al. ”SCA-CNN: Spatial and Channel-wise Attention in Convolutional Networks for Image Captioning.” (2016):6298-6306
2016
Earlier work this paper cites.
T. Niu, S. A. Zhu, L. Pang and A. El Saddik, Sentiment Analysis on Multi-view Social Data, MultiMedia Modeling (MMM), pp: 15-27, Miami, 2016
2016
Earlier work this paper cites.
Niu T, Zhu S, Pang L, El Saddik A. Sentiment analysis on multi-view social data. InMultiMedia Modeling: 22nd International Conference, MMM 2016, Miami, FL, USA, January 4-6, 2016, Proceedings, Part II 22 2016 (pp. 15-27). Springer International Publishing
2016
Earlier work this paper cites.
Y. Yu, H. Lin, J. Meng, and Z. Zhao, ”Visual and Textual Sentiment Analysis of a Microblog Using Deep Convolutional Neural Networks,” Algorithms, vol. 9, p. 41, 2016
2016
Earlier work this paper cites.
Schifanella R, De Juan P, Tetreault J, Cao L. Detecting sarcasm in multimodal social platforms. InProceedings of the 24th ACM international conference on Multimedia 2016 Oct 1 (pp. 1136-1145)
2016
Earlier work this paper cites.
Zadeh A, Chen M, Poria S, et al. Tensor Fusion Network for Multimodal Sentiment Analysis[C] // In
2017
Earlier work this paper cites.
Chen M, Wang S, Liang P P, et al. Multimodal Sentiment Analysis with WordLevel Fusion and Reinforcement Learning // In
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
Vaswani A, Shazeer N, Parmar N, et al. Attention is all you need[J]
2017
Earlier work this paper cites.
Soleymani M, Garcia D, Jou B, Schuller B, Chang SF, Pantic M. A survey of multimodal sentiment analysis. Image and Vision Computing. 2017 Sep 1;65:3-14
2017
Earlier work this paper cites.
Xu N. Analyzing multimodal public sentiment based on hierarchical semantic attentional network. In2017 IEEE international conference on intelligence and security informatics (ISI) 2017 Jul 22 (pp. 152-154). IEEE
2017
Earlier work this paper cites.
Xu N, Mao W. Multisentinet: A deep semantic network for multimodal sentiment analysis. InProceedings of the 2017 ACM on Conference on Information and Knowledge Management 2017 Nov 6 (pp. 2399-2402)
2017
Earlier work this paper cites.
Chen M, Wang S, Liang PP, Baltrušaitis T, Zadeh A, Morency LP. Multimodal sentiment analysis with word-level fusion and reinforcement learning. InProceedings of the 19th ACM international conference on multimodal interaction 2017 Nov 3 (pp. 163-171)
2017
Earlier work this paper cites.
Speer R, Chin J, Havasi C. Conceptnet 5.5: An open multilingual graph of general knowledge. InProceedings of the AAAI conference on artificial intelligence 2017 Feb 12 (Vol. 31, No. 1)
2017
Earlier work this paper cites.
Tadas Baltrus aitis, Chaitanya Ahuja, and Louis-Philippe Morency. Multimodal machine learning: A survey and taxonomy
2018
Earlier work this paper cites.
Yi Tay, Anh Tuan Luu, Siu Cheung Hui, and Jian Su. Reasoning with sarcasm by reading in between. Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics. 2018: 1010-1020
2018
Earlier work this paper cites.
P. Anderson, X. He, C. Buehler, D. Teney, M. Johnson, S. Gould, and L. Zhang. Bottom-up and top-down attention for image captioning and visual question answering. In CVPR, volume 3, page 6, 2018
2018
Earlier work this paper cites.
Zhou H, Huang M, Zhang T, et al. Emotional chatting machine: Emotional conversation generation with internal and external memory[C]//Proceedings of the AAAI Conference on Artificial Intelligence. 2018, 32(1)
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
Sharma P, Ding N, Goodman S, Soricut R. Conceptual captions: A cleaned, hypernymed, image alt-text dataset for automatic image captioning. InProceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) 2018 Jul (pp. 2556-2565)
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
Xu N, Mao W, Chen G. A co-memory network for multimodal sentiment analysis. InThe 41st international ACM SIGIR conference on research & development in information retrieval 2018 Jun 27 (pp. 929-932)
2018
Earlier work this paper cites.
Q. Zhang, J. Fu, X. Liu, X. Huang, Adaptive co-attention network for named entity recognition in tweets, in: Proceedings of the AAAI Conference on Artificial Intelligence, Vol. 32, 2018
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
Anderson P, He X, Buehler C, Teney D, Johnson M, Gould S, Zhang L. Bottom-up and top-down attention for image captioning and visual question answering. InProceedings of the IEEE conference on computer vision and pattern recognition 2018 (pp. 6077-6086)
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
Zadeh A, Liang PP, Mazumder N, Poria S, Cambria E, Morency LP. Memory fusion network for multi-view sequential learning. InProceedings of the AAAI conference on artificial intelligence 2018 Apr 27 (Vol. 32, No. 1)
2018
Earlier work this paper cites.
Truong Q T, Lauw H W. Vistanet: Visual Aspect Attention Network for Multimodal Sentiment Analysis[C]// In
2019
Earlier work this paper cites.
Wang Y, Shen Y, Liu Z, et al. Words can shift: Dynamically adjusting word representations using nonverbal behaviors[C]// In
2019
Earlier work this paper cites.
Tsai Y H H, Bai S, Liang P P, et al. Multimodal transformer for unaligned multimodal language sequences[C]//
2019
Earlier work this paper cites.
Xiong T, Zhang P, Zhu H, et al. Sarcasm Detection with Self-matching Networks and Low-Rank Bilinear Pooling[C]// The World Wide Web Conference. New York: ACM, 2019: 2115-2124
2019
Earlier work this paper cites.
Santiago Castro, Devamanyu Hazarika, Veronica PerezRosas, Roger Zimmermann, Rada Mihalcea, and Soujanya Poria. Towards multimodal sarcasm detection (an obviously perfect paper). Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics. 2019: 4619-4629
2019
Earlier work this paper cites.
Yitao Cai, Huiyu Cai, and Xiaojun Wan. Multimodal sarcasm detection in twitter with hierarchical fusion model. Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics. 2019: 2506-2515
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
Zeng B, Yang H, Xu R, et al. Lcf: A local context focus mechanism for aspect-based sentiment classification[J]. Applied Sciences, 2019, 9(16): 3389
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
Y. Wu, A. Kirillov, F. Massa, W.-Y. Lo, and R. Girshick, ”Detectron2,”https://github.com/facebookresearch/ detectron2, 2019
2019
Earlier work this paper cites.
Tsai Y H H, Bai S, Liang P P, et al. Multimodal transformer for unaligned multimodal language sequences[C] //Proceedings of the conference. Association for Computational Linguistics. Meeting. NIH Public Access, 2019, 2019: 6558
2019
Earlier work this paper cites.
Nan Xu, Wenji Mao, Guandan Chen. Multi-interactive memory network for aspect based multimodal sentiment analysis. Proceedings of the AAAI Conference on Artificial Intelligence. 2019: 371-378
2019
Earlier work this paper cites.
Yu J, Jiang J, Xia R. Entity-sensitive attention and fusion network for entity-level multimodal sentiment classification[J]. IEEE/ACM Transactions on Audio, Speech, and Language Processing, 2019, 28: 429-439
2019
Earlier work this paper cites.
Yu J, Jiang J. Adapting BERT for target-oriented multimodal sentiment classification[C]. IJCAI, 2019
2019
Cited alongside, same era.
2019
Cited alongside, same era.
2019
Cited alongside, same era.
Yu J, Jiang J. Adapting BERT for target-oriented multimodal sentiment classification[C]
2019
Cited alongside, same era.
Yang H, Zhao Y, Qin B. Face-Sensitive Image-to-Emotional-Text Cross-modal Translation for Multimodal Aspect-based Sentiment Analysis[C]// In
2022
Later among the works it cites.
Lu P, Mishra S, Xia T, Qiu L, Chang KW, Zhu SC, Tafjord O, Clark P, Kalyan A. Learn to explain: Multimodal reasoning via thought chains for science question answering. Advances in Neural Information Processing Systems. 2022 Dec 6;35:2507-21
2022
Later among the works it cites.
Ramamoorthy S, Gunti N, Mishra S, Suryavardan S, Reganti A, Patwa P, DaS A, Chakraborty T, Sheth A, Ekbal A, Ahuja C. Memotion 2: Dataset on sentiment and emotion analysis of memes. InProceedings of De-Factify: Workshop on Multimodal Fact Checking and Hate Speech Detection, CEUR 2022
2022
Later among the works it cites.
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2019
Cited alongside, same era.
Y. Wu, A. Kirillov, F. Massa, W.-Y. Lo, and R. Girshick, “Detectron2,” https://github.com/facebookresearch/ detectron2, 2019
2019
Cited alongside, same era.
Cai Y, Cai H, Wan X. Multi-modal sarcasm detection in twitter with hierarchical fusion model. InProceedings of the 57th annual meeting of the association for computational linguistics 2019 Jul (pp. 2506-2515)
2019
Cited alongside, same era.
2019
Cited alongside, same era.
Cai Y, Cai H, Wan X. Multi-modal sarcasm detection in twitter with hierarchical fusion model. InProceedings of the 57th annual meeting of the association for computational linguistics 2019 Jul (pp. 2506-2515)
2019
Cited alongside, same era.
Ma D, Li S, Wu F, Xie X, Wang H. Exploring sequence-to-sequence learning in aspect term extraction. InProceedings of the 57th annual meeting of the association for computational linguistics 2019 Jul (pp. 3538-3547)
2019
Cited alongside, same era.
2019
Cited alongside, same era.
2019
Cited alongside, same era.
Tong Zhu, Leida Li, Jufeng Yang, Sicheng Zhao, Xiao Xiao, Multimodal emotion classification with multi-level semantic reasoning network, IEEE Trans. Multim. (2022) http://dx.doi.org/10.1109/TMM.2022.3214989, Early Access
2022
Later among the works it cites.
Jia A, He Y, Zhang Y, Uprety S, Song D, Lioma C. Beyond emotion: A multi-modal dataset for human desire understanding. InProceedings of the 2022 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies 2022 Jul (pp. 1512-1522)
2022
Later among the works it cites.
Ouyang L, Wu J, Jiang X, Almeida D, Wainwright C, Mishkin P, Zhang C, Agarwal S, Slama K, Ray A, Schulman J. Training language models to follow instructions with human feedback. Advances in neural information processing systems. 2022 Dec 6;35:27730-44
2022
Later among the works it cites.
2022
Later among the works it cites.
Yang L, Na JC, Yu J. Cross-modal multitask transformer for end-to-end multimodal aspect-based sentiment analysis. Information Processing & Management. 2022 Sep 1;59(5):103038
2022
Later among the works it cites.
Yu Y, Zhang D, Li S. Unified multi-modal pre-training for few-shot sentiment analysis with prompt-based learning. InProceedings of the 30th ACM International Conference on Multimedia 2022 Oct 10 (pp. 189-198)
2022
Later among the works it cites.
Yu Y, Zhang D. Few-shot multi-modal sentiment analysis with prompt-based vision-aware language modeling. In2022 IEEE International Conference on Multimedia and Expo (ICME) 2022 Jul 18 (pp. 1-6). IEEE
2022
Later among the works it cites.
Liu Y, Yuan Z, Mao H, Liang Z, Yang W, Qiu Y, Cheng T, Li X, Xu H, Gao K. Make acoustic and visual cues matter: CH-SIMS v2. 0 dataset and AV-Mixup consistent module. InProceedings of the 2022 International Conference on Multimodal Interaction 2022 Nov 7 (pp. 247-258)
2022
Later among the works it cites.
2022
Later among the works it cites.
Mai S, Zeng Y, Zheng S, Hu H. Hybrid contrastive learning of tri-modal representation for multimodal sentiment analysis. IEEE Transactions on Affective Computing. 2022 May 3
2022
Later among the works it cites.
2022
Later among the works it cites.
D. Tomás, R. Ortega-Bueno, G. Zhang, P. Rosso, R. Schifanella, Transformer-based models for multimodal irony detection, J. Ambient Intell. Humaniz. Comput. (2022) 1–12
2022
Later among the works it cites.
B. Liang, C. Lou, X. Li, M. Yang, L. Gui, Y. He, W. Pei, R. Xu, Multi-modal sarcasm detection via cross-modal graph convolutional network, in: Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), 2022, pp. 1767–1777
2022
Later among the works it cites.
OpenAI. 2023. GPT-4 technical report
2023
Later among the works it cites.
Xiang Deng, Vasilisa Bashlovkina, Feng Han, Simon Baumgartner, and Michael Bendersky. 2023. Llms to the moon? reddit market sentiment analysis with LLMs. In
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
Liu H, Li C, Wu Q, et al. Visual instruction tuning[J]. arXiv preprint arXiv:2304.08485, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
Meta A I. Introducing LLaMA: A foundational, 65-billion-parameter large language model[J]. Meta AI. https://ai. facebook. com/blog/large-language-model-llama-meta-ai, 2023
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
Girdhar R, El-Nouby A, Liu Z, Singh M, Alwala KV, Joulin A, Misra I. Imagebind: One embedding space to bind them all. InProceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition 2023 (pp. 15180-15190)
2023
Later among the works it cites.
2023
Later among the works it cites.
Li J, Li D, Savarese S, Hoi S. Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models. InInternational conference on machine learning 2023 Jul 3 (pp. 19730-19742). PMLR
2023
Later among the works it cites.
2023
Later among the works it cites.
Chiang WL, Li Z, Lin Z, Sheng Y, Wu Z, Zhang H, Zheng L, Zhuang S, Zhuang Y, Gonzalez JE, Stoica I. Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality. See https://vicuna. lmsys. org (accessed 14 April 2023). 2023 Mar;2(3):6
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
Lian Z, Sun H, Sun L, Chen K, Xu M, Wang K, Xu K, He Y, Li Y, Zhao J, Liu Y. Mer 2023: Multi-label learning, modality robustness, and semi-supervised learning. InProceedings of the 31st ACM International Conference on Multimedia 2023 Oct 26 (pp. 9610-9614)
2023
Later among the works it cites.
2023
Later among the works it cites.
Yue T, Mao R, Wang H, Hu Z, Cambria E. KnowleNet: Knowledge fusion network for multimodal sarcasm detection. Information Fusion. 2023 Dec 1;100:101921
2023
Later among the works it cites.
2023
Later among the works it cites.
Leveraging Generative Large Language Models with Visual Instruction and Demonstration Retrieval for Multimodal Sarcasm Detection. openreview Dec.2023. https://openreview.net/forum?id=_98UHlfKejb
2023
Later among the works it cites.
Zheng Z, Zhang Z, Wang Z, Fu R, Liu M, Wang Z, Qin B. Decompose, Prioritize, and Eliminate: Dynamically Integrating Diverse Representations for Multimodal Named Entity Recognition. InProceedings of the 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation (LREC-COLING 2024) 2024 May (pp. 4498-4508)
2024
Closest in time.
Liu H, Li C, Wu Q, Lee YJ. Visual instruction tuning. Advances in neural information processing systems. 2024 Feb 13;36
2024
Closest in time.
Zhao H, Yang M, Bai X, Liu H. A Survey on Multimodal Aspect-Based Sentiment Analysis. IEEE Access. 2024 Jan 16
2024
Closest in time.
2024
Closest in time.
Peng T, Li Z, Wang P, Zhang L, Zhao H. A Novel Energy Based Model Mechanism for Multi-Modal Aspect-Based Sentiment Analysis. InProceedings of the AAAI Conference on Artificial Intelligence 2024 Mar 24 (Vol. 38, No. 17, pp. 18869-18878)
2024
Closest in time.
Dai W, Li J, Li D, Tiong AM, Zhao J, Wang W, Li B, Fung PN, Hoi S. Instructblip: Towards general-purpose vision-language models with instruction tuning. Advances in Neural Information Processing Systems. 2024 Feb 13;36
2024
Closest in time.
Xiao L, Wu X, Xu J, Li W, Jin C, He L. Atlantis: Aesthetic-oriented multiple granularities fusion network for joint multimodal aspect-based sentiment analysis. Information Fusion. 2024 Feb 15:102304
2024
Closest in time.
2024
Closest in time.
Li Y, Ding H, Lin Y, Feng X, Chang L. Multi-level textual-visual alignment and fusion network for multimodal aspect-based sentiment analysis. Artificial Intelligence Review. 2024 Apr;57(4):1-26
2024
Closest in time.
Yang L, Wang Z, Li Z, Na JC, Yu J. An empirical study of Multimodal Entity-Based Sentiment Analysis with ChatGPT: Improving in-context learning via entity-aware contrastive learning. Information Processing & Management. 2024 Jul 1;61(4):103724
2024
Closest in time.
2024
Closest in time.
Huang J, Pu Y, Zhou D, Shi H, Zhao Z, Xu D, Cao J. Multimodal Sentiment Analysis Based on 3D Stereoscopic Attention. InICASSP 2024-2024 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) 2024 Apr 14 (pp. 11151-11155). IEEE
2024
Closest in time.
H. Liu, B. Yang, Z. Yu, A multi-view interactive approach for multimodal sarcasm detection in social internet of things with knowledge enhancement, Appl. Sci. 14 (5) (2024) 2146
2024
Closest in time.
H. Fu, H. Liu, H. Wang, L. Xu, J. Lin, D. Jiang, Multi-modal sarcasm detection with sentiment word embedding, Electronics 13 (5) (2024) 855
2024
Closest in time.
Yi G, Fan C, Zhu K, Lv Z, Liang S, Wen Z, Pei G, Li T, Tao J. Vlp2msa: expanding vision-language pre-training to multimodal sentiment analysis. Knowledge-Based Systems. 2024 Jan 11;283:111136
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.
Dettmers T, Pagnoni A, Holtzman A, Zettlemoyer L. Qlora: Efficient finetuning of quantized llms. Advances in Neural Information Processing Systems. 2024 Feb 13;36
2024
Closest in time.
Zhang Z, Peng L, Pang T, Han J, Zhao H, Schuller BW. Refashioning emotion recognition modelling: The advent of generalised large models. IEEE Transactions on Computational Social Systems. 2024 May 30
2024
Closest in time.
Peng L, Zhang Z, Pang T, Han J, Zhao H, Chen H, Schuller BW. Customising General Large Language Models for Specialised Emotion Recognition Tasks. InICASSP 2024-2024 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP) 2024 Apr 14 (pp. 11326-11330). IEEE
2024
Closest in time.
Geetha AV, Mala T, Priyanka D, Uma E. Multimodal Emotion Recognition with deep learning: advancements, challenges, and future directions. Information Fusion. 2024 May 1;105:102218
2024
Closest in time.
2024
Closest in time.
Xu K, Ba J, Kiros R, et al. Show, attend and tell: Neural image caption generation with visual attention[C]//International conference on machine learning. PMLR, 2015: 2048-2057
2057
Closest in time.
Li L, Chen Y C, Cheng Y, et al. Hero: Hierarchical Encoder for Video+ Language Omni-representation Pre-training[C]// In
2065
Closest in time.