Fetching the paper…
Reading the bibliography…
Recently, GPT-4 with Vision (GPT-4V) has demonstrated remarkable visual capabilities across various tasks, but its performance in emotion recognition has not been fully evaluated.
P. Ekman, W. V. Friesen, Facial action coding system, Environmental Psychology & Nonverbal Behavior (1978)
1978
Earlier work this paper cites.
H. Hermansky, Perceptual linear predictive (plp) analysis of speech, the Journal of the Acoustical Society of America 87 (4) (1990) 1738–1752
1990
Earlier work this paper cites.
P. EKMAN, Strong evidence for universals in facial expressions: a reply to russell’s mistaken critique, Psychological bulletin 115 (2) (1994) 268–287
1994
Earlier work this paper cites.
R. W. Picard, E. Vyzas, J. Healey, Toward machine emotional intelligence: Analysis of affective physiological state, IEEE Transactions on Pattern Analysis and Machine Intelligence 23 (10) (2001) 1175–1191
2001
Earlier work this paper cites.
N. Sebe, I. Cohen, T. Gevers, T. S. Huang, Multimodal approaches for emotion recognition: a survey, in: Internet Imaging VI, Vol. 5670, SPIE, 2005, pp. 56–67
2005
Earlier work this paper cites.
O. Martin, I. Kotsia, B. Macq, I. Pitas, The enterface’05 audio-visual emotion database, in: Proceedings of the 22nd International Conference on Data Engineering Workshops, IEEE, 2006, pp. 8–8
2006
Earlier work this paper cites.
P. Lucey, J. F. Cohn, T. Kanade, J. Saragih, Z. Ambadar, I. Matthews, The extended cohn-kanade dataset (ck+): A complete dataset for action unit and emotion-specified expression, in: IEEE Computer Society Conference on Computer Vision and Pattern Recognition Workshops, IEEE, 2010, pp. 94–101
2010
Earlier work this paper cites.
D. Borth, R. Ji, T. Chen, T. Breuel, S.-F. Chang, Large-scale visual sentiment ontology and detectors using adjective noun pairs, in: Proceedings of the 21st ACM International Conference on Multimedia, 2013, pp. 223–232
2013
Earlier work this paper cites.
W.-J. Yan, Q. Wu, Y.-J. Liu, S.-J. Wang, X. Fu, Casme database: A dataset of spontaneous micro-expressions collected from neutralized faces, in: Proceedings of the 10th IEEE International Conference and Workshops on Automatic Face and Gesture Recognition (FG), IEEE, 2013, pp. 1–7
2013
Earlier work this paper cites.
C.-H. Wu, J.-C. Lin, W.-L. Wei, Survey on audiovisual emotion recognition: databases, features, and data fusion strategies, APSIPA Transactions on Signal and Information Processing 3 (2014)
2014
Earlier work this paper cites.
W.-J. Yan, X. Li, S.-J. Wang, G. Zhao, Y.-J. Liu, Y.-H. Chen, X. Fu, Casme ii: An improved spontaneous micro-expression database and the baseline evaluation, PLoS ONE 9 (1) (2014) e86041–e86041
2014
Earlier work this paper cites.
S. Zhao, Y. Gao, X. Jiang, H. Yao, T.-S. Chua, X. Sun, Exploring principles-of-art features for image emotion recognition, in: Proceedings of the 22nd ACM international conference on Multimedia, 2014, pp. 47–56
2014
Earlier work this paper cites.
2014
Earlier work this paper cites.
Q. You, J. Luo, H. Jin, J. Yang, Robust image sentiment analysis using progressively trained and domain transferred deep networks, in: Proceedings of the Twenty-Ninth AAAI Conference on Artificial Intelligence, 2015, pp. 381–388
2015
Earlier work this paper cites.
A. Dhall, O. Ramana Murthy, R. Goecke, J. Joshi, T. Gedeon, Video and image based emotion recognition challenges in the wild: Emotiw 2015, in: Proceedings of the 2015 ACM on International Conference on Multimodal Interaction, 2015, pp. 423–426
2015
Earlier work this paper cites.
K. Simonyan, A. Zisserman, Very deep convolutional networks for large-scale image recognition, in: Proceedings of the International Conference on Learning Representations, ICLR, 2015, pp. 1–14
2015
Earlier work this paper cites.
Y. Wang, J. See, R. C.-W. Phan, Y.-H. Oh, Lbp with six intersection points: Reducing redundant information in lbp-top for micro-expression recognition, in: Computer Vision–ACCV 2014: 12th Asian Conference on Computer Vision, Singapore, Singapore, November 1-5, 2014, Revised Selected Papers, Part I 12, Springer, 2015, pp. 525–537
2015
Earlier work this paper cites.
G. Cai, B. Xia, Convolutional neural networks for multimedia sentiment analysis, in: Natural Language Processing and Chinese Computing: 4th CCF Conference, NLPCC 2015, Nanchang, China, October 9-13, 2015, Proceedings 4, Springer, 2015, pp. 159–167
2015
Earlier work this paper cites.
D. Tran, L. Bourdev, R. Fergus, L. Torresani, M. Paluri, Learning spatiotemporal features with 3d convolutional networks, in: Proceedings of the IEEE international conference on computer vision, 2015, pp. 4489–4497
2015
Earlier work this paper cites.
T. Niu, S. Zhu, L. Pang, A. El Saddik, Sentiment analysis on multi-view social data, in: MultiMedia Modeling: 22nd International Conference, MMM 2016, Miami, FL, USA, January 4-6, 2016, Proceedings, Part II 22, Springer, 2016, pp. 15–27
2016
Earlier work this paper cites.
Q. You, J. Luo, H. Jin, J. Yang, Building a large scale dataset for image emotion recognition: the fine print and the benchmark, in: Proceedings of the Thirtieth AAAI Conference on Artificial Intelligence, 2016, pp. 308–314
2016
Earlier work this paper cites.
A. K. Davison, C. Lansley, N. Costen, K. Tan, M. H. Yap, Samm: A spontaneous micro-facial movement dataset, IEEE Transactions on Affective Computing 9 (1) (2016) 116–129
2016
Earlier work this paper cites.
E. Barsoum, C. Zhang, C. C. Ferrer, Z. Zhang, Training deep networks for facial expression recognition with crowd-sourced label distribution, in: Proceedings of the 18th ACM International Conference on Multimodal Interaction, 2016, pp. 279–283
2016
Earlier work this paper cites.
Y. Yu, H. Lin, J. Meng, Z. Zhao, Visual and textual sentiment analysis of a microblog using deep convolutional neural networks, Algorithms 9 (41) (2016) 1–11
2016
Earlier work this paper cites.
Q. You, H. Jin, J. Luo, Visual sentiment analysis by attending on local image regions, in: Proceedings of the Thirty-First AAAI Conference on Artificial Intelligence, 2017, pp. 231–237
2017
Earlier work this paper cites.
S. Li, W. Deng, J. Du, Reliable crowdsourcing and deep locality-preserving learning for expression recognition in the wild, in: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, 2017, pp. 2852–2861
2017
Earlier work this paper cites.
A. Mollahosseini, B. Hasani, M. H. Mahoor, Affectnet: A database for facial expression, valence, and arousal computing in the wild, IEEE Transactions on Affective Computing 10 (1) (2017) 18–31
2017
Earlier work this paper cites.
A. Zadeh, M. Chen, S. Poria, E. Cambria, L.-P. Morency, Tensor fusion network for multimodal sentiment analysis, in: Proceedings of the Conference on Empirical Methods in Natural Language Processing, 2017, pp. 1103–1114
2017
Earlier work this paper cites.
N. Xu, W. Mao, Multisentinet: A deep semantic network for multimodal sentiment analysis, in: Proceedings of the 2017 ACM on Conference on Information and Knowledge Management, 2017, pp. 2399–2402
2017
Earlier work this paper cites.
Z. Meng, P. Liu, J. Cai, S. Han, Y. Tong, Identity-aware convolutional neural network for facial expression recognition, in: 2017 12th IEEE International Conference on Automatic Face & Gesture Recognition (FG 2017), IEEE, 2017, pp. 558–565
2017
Earlier work this paper cites.
S. R. Livingstone, F. A. Russo, The ryerson audio-visual database of emotional speech and song (ravdess): A dynamic, multimodal set of facial and vocal expressions in north american english, PloS One 13 (5) (2018) e0196391
2018
Earlier work this paper cites.
J. Yang, D. She, M. Sun, M.-M. Cheng, P. L. Rosin, L. Wang, Visual sentiment prediction based on automatic discovery of affective regions, IEEE Transactions on Multimedia 20 (9) (2018) 2513–2525
2018
Earlier work this paper cites.
Y. Li, X. Huang, G. Zhao, Can micro-expression be recognized based on single apex frame?, in: 2018 25th IEEE International Conference on Image Processing (ICIP), IEEE, 2018, pp. 3094–3098
2018
Earlier work this paper cites.
N. Xu, W. Mao, G. Chen, A co-memory network for multimodal sentiment analysis, in: The 41st international ACM SIGIR conference on research & development in information retrieval, 2018, pp. 929–932
2018
Earlier work this paper cites.
A. Zadeh, P. P. Liang, N. Mazumder, S. Poria, E. Cambria, L.-P. Morency, Memory fusion network for multi-view sequential learning, in: Proceedings of the AAAI Conference on Artificial Intelligence, 2018, pp. 5634–5641
2018
Cited alongside, same era.
H. Yang, Z. Zhang, L. Yin, Identity-adaptive facial expression recognition through expression regeneration using conditional generative adversarial networks, in: 2018 13th IEEE International Conference on Automatic Face & Gesture Recognition (FG 2018), IEEE, 2018, pp. 294–301
2018
Cited alongside, same era.
Z. Xia, X. Hong, X. Gao, X. Feng, G. Zhao, Spatiotemporal recurrent convolutional networks for recognizing spontaneous micro-expressions, IEEE Transactions on Multimedia 22 (3) (2019) 626–640
2019
Cited alongside, same era.
B. Song, K. Li, Y. Zong, J. Zhu, W. Zheng, J. Shi, L. Zhao, Recognizing spontaneous micro-expression using a three-stream convolutional neural network, IEEE Access 7 (2019) 184537–184551
2019
Y. Li, J. Wei, Y. Liu, J. Kauttonen, G. Zhao, Deep learning for micro-expression recognition: A survey, IEEE Transactions on Affective Computing (2022)
2022
Later among the works it cites.
Y. Wang, Y. Sun, Y. Huang, Z. Liu, S. Gao, W. Zhang, W. Ge, W. Zhang, Ferv39k: A large-scale multi-scene dataset for facial expression recognition in videos, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2022, pp. 20922–20931
2022
Later among the works it cites.
T. Zhu, L. Li, J. Yang, S. Zhao, H. Liu, J. Qian, Multimodal sentiment analysis with image-text interaction network, IEEE Transactions on Multimedia (2022)
2022
Later among the works it cites.
J. Jiang, W. Deng, Disentangling identity and pose for facial expression recognition, IEEE Transactions on Affective Computing 13 (4) (2022) 1868–1878
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Y.-H. H. Tsai, P. P. Liang, A. Zadeh, L.-P. Morency, R. Salakhutdinov, Learning factorized multimodal representations, in: Proceedings of the 7th International Conference on Learning Representations, 2019, pp. 1–20
2019
Cited alongside, same era.
Y.-H. H. Tsai, S. Bai, P. P. Liang, J. Z. Kolter, L.-P. Morency, R. Salakhutdinov, Multimodal transformer for unaligned multimodal language sequences, in: Proceedings of the 57th Conference of the Association for Computational Linguistics, 2019, pp. 6558–6569
2019
Cited alongside, same era.
M. Bai, W. Xie, L. Shen, Disentangled feature based adversarial learning for facial expression recognition, in: 2019 IEEE International Conference on Image Processing (ICIP), IEEE, 2019, pp. 31–35
2019
Cited alongside, same era.
K. Yan, W. Zheng, T. Zhang, Y. Zong, C. Tang, C. Lu, Z. Cui, Cross-domain facial expression recognition based on transductive deep transfer learning, IEEE Access 7 (2019) 108906–108915
2019
Cited alongside, same era.
X. Liu, B. V. Kumar, P. Jia, J. You, Hard negative generation for identity-disentangled facial expression recognition, Pattern Recognition 88 (2019) 1–12
2019
Cited alongside, same era.
E. Ghaleb, M. Popa, S. Asteriadis, Multimodal and temporal perception of audio-visual cues for emotion recognition, in: 2019 8th International Conference on Affective Computing and Intelligent Interaction (ACII), IEEE, 2019, pp. 552–558
2019
Cited alongside, same era.
X. Pan, G. Ying, G. Chen, H. Li, W. Li, A deep spatial and temporal aggregation framework for video-based facial expression recognition, IEEE Access 7 (2019) 48807–48815
2019
Cited alongside, same era.
D. Meng, X. Peng, K. Wang, Y. Qiao, Frame attention networks for facial expression recognition in videos, in: 2019 IEEE international conference on image processing (ICIP), IEEE, 2019, pp. 3866–3870
2019
Cited alongside, same era.
Y. Wang, Y. Sun, W. Song, S. Gao, Y. Huang, Z. Chen, W. Ge, W. Zhang, Dpcnet: Dual path multi-excitation collaborative network for facial expression representation learning in videos, in: Proceedings of the 30th ACM International Conference on Multimedia, 2022, pp. 101–110
2022
Later among the works it cites.
Y. Zhang, C. Wang, X. Ling, W. Deng, Learn from all: Erasing attention consistency for noisy label facial expression recognition, in: European Conference on Computer Vision, Springer, 2022, pp. 418–434
2022
Later among the works it cites.
2022
Later among the works it cites.
R. Zhao, T. Liu, Z. Huang, D. P. Lun, K.-M. Lam, Spatial-temporal graphs plus transformers for geometry-guided facial expression recognition, IEEE Transactions on Affective Computing (2022)
2022
Later among the works it cites.
R. Mao, Q. Liu, K. He, W. Li, E. Cambria, The biases of pre-trained language models: An empirical study on prompt-based sentiment analysis and emotion detection, IEEE Transactions on Affective Computing (2022)
2022
Later among the works it cites.
J. Yang, Q. Huang, T. Ding, D. Lischinski, D. Cohen-Or, H. Huang, Emoset: A large-scale visual emotion dataset with rich attributes, in: Proceedings of the IEEE/CVF International Conference on Computer Vision, 2023, pp. 20383–20394
2023
Closest in time.
2023
Closest in time.
Z. Lian, H. Sun, L. Sun, K. Chen, M. Xu, K. Wang, K. Xu, Y. He, Y. Li, J. Zhao, et al., Mer 2023: Multi-label learning, modality robustness, and semi-supervised learning, in: Proceedings of the 31st ACM International Conference on Multimedia, 2023, pp. 9610–9614
2023
Closest in time.
2023
Closest in time.
H. Liu, C. Li, Q. Wu, Y. J. Lee, Visual instruction tuning, arXiv preprint arXiv:2304.08485 (2023)
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
X. Li, T. Zhang, Y. Dubois, R. Taori, I. Gulrajani, C. Guestrin, P. Liang, T. B. Hashimoto, Alpacaeval: An automatic evaluator of instruction-following models, https://github.com/tatsu-lab/alpaca_eval (2023)
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
Z. Wen, W. Lin, T. Wang, G. Xu, Distract your attention: Multi-head cross attention network for facial expression recognition, Biomimetics 8 (2) (2023) 199
2023
Closest in time.
C. Zheng, M. Mendieta, C. Chen, Poster: A pyramid cross-fusion transformer network for facial expression recognition, in: Proceedings of the IEEE/CVF International Conference on Computer Vision, 2023, pp. 3146–3155
2023
Closest in time.
2023
Closest in time.
H. Wang, B. Li, S. Wu, S. Shen, F. Liu, S. Ding, A. Zhou, Rethinking the learning paradigm for dynamic facial expression recognition, in: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2023, pp. 17958–17968
2023
Closest in time.
H. Li, H. Niu, Z. Zhu, F. Zhao, Intensity-aware loss for dynamic facial expression recognition in the wild, in: Proceedings of the AAAI Conference on Artificial Intelligence, 2023, pp. 67–75
2023
Closest in time.
2023
Closest in time.
L. Sun, Z. Lian, B. Liu, J. Tao, Mae-dfer: Efficient masked autoencoder for self-supervised dynamic facial expression recognition, in: Proceedings of the 31st ACM International Conference on Multimedia, 2023, pp. 6110–6121
2023
Closest in time.
2023
Closest in time.
2024
Closest in time.
P. Lu, H. Bansal, T. Xia, J. Liu, C. Li, H. Hajishirzi, H. Cheng, K.-W. Chang, M. Galley, J. Gao, Mathvista: Evaluating mathematical reasoning of foundation models in visual contexts, in: Proceedings of the International Conference on Learning Representations, ICLR, 2024, pp. 1–116
2024
Closest in time.
H. Li, N. Wang, X. Ding, X. Yang, X. Gao, Adaptively learning facial expression representation via cf labels and distillation, IEEE Transactions on Image Processing 30 (2021) 2016–2028
2028
Closest in time.