Fetching the paper…
Reading the bibliography…
Dynamic facial expression recognition (DFER) in the wild is still hindered by data limitations, e.g., insufficient quantity and diversity of pose, occlusion and illumination, as well as the inherent ambiguity of facial expressions.
J. L. Fleiss, “Measuring nominal scale agreement among many raters.” Psychological bulletin , vol. 76, no. 5, p. 378, 1971
1971
Earlier work this paper cites.
L. Van der Maaten and G. Hinton, “Visualizing data using t-sne.” Journal of machine learning research , vol. 9, no. 11, 2008
2008
Earlier work this paper cites.
M. Liu, S. Li, S. Shan, and X. Chen, “Au-aware deep networks for facial expression recognition,” in 2013 10th IEEE international conference and workshops on automatic face and gesture recognition (FG) . IEEE, 2013, pp. 1–6
2013
Earlier work this paper cites.
D. Tran, L. D. Bourdev, R. Fergus, L. Torresani, and M. Paluri, “Learning spatiotemporal features with 3d convolutional networks,” 2015 IEEE International Conference on Computer Vision (ICCV) , pp. 4489–4497, 2014
2014
Earlier work this paper cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , pp. 770–778, 2015
2015
Earlier work this paper cites.
H. Jung, S. Lee, J. Yim, S. Park, and J. Kim, “Joint fine-tuning in deep neural networks for facial expression recognition,” 2015 IEEE International Conference on Computer Vision (ICCV) , pp. 2983–2991, 2015
2015
Earlier work this paper cites.
O. Russakovsky, J. Deng, H. Su, J. Krause, S. Satheesh, S. Ma, Z. Huang, A. Karpathy, A. Khosla, M. Bernstein et al. , “Imagenet large scale visual recognition challenge,” International journal of computer vision , vol. 115, pp. 211–252, 2015
2015
Earlier work this paper cites.
E. Barsoum, C. Zhang, C. Canton-Ferrer, and Z. Zhang, “Training deep networks for facial expression recognition with crowd-sourced label distribution,” Proceedings of the 18th ACM International Conference on Multimodal Interaction , 2016
2016
Earlier work this paper cites.
Y. Fan, X. Lu, D. Li, and Y. Liu, “Video-based emotion recognition using cnn-rnn and c3d hybrid networks,” Proceedings of the 18th ACM International Conference on Multimodal Interaction , 2016
2016
Earlier work this paper cites.
Z. Liu, M. Wu, W. Cao, L. Chen, J.-P. Xu, R. Zhang, M. Zhou, and J.-W. Mao, “A facial expression emotion recognition based human-robot interaction system,” IEEE/CAA Journal of Automatica Sinica , vol. 4, pp. 668–676, 2017
2017
Earlier work this paper cites.
S. Li, W. Deng, and J. Du, “Reliable crowdsourcing and deep locality-preserving learning for expression recognition in the wild,” 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , pp. 2584–2593, 2017
2017
Earlier work this paper cites.
A. Mollahosseini, B. Hasani, and M. H. Mahoor, “Affectnet: A database for facial expression, valence, and arousal computing in the wild,” IEEE Transactions on Affective Computing , vol. 10, pp. 18–31, 2017
2017
Earlier work this paper cites.
J. Carreira and A. Zisserman, “Quo vadis, action recognition? a new model and the kinetics dataset,” in proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , 2017, pp. 6299–6308
2017
Earlier work this paper cites.
D. Tran, H. Wang, L. Torresani, J. Ray, Y. LeCun, and M. Paluri, “A closer look at spatiotemporal convolutions for action recognition,” in Proceedings of the IEEE conference on Computer Vision and Pattern Recognition , 2018, pp. 6450–6459
2018
Earlier work this paper cites.
J. Chen, Z. Chen, Z. Chi, and H. Fu, “Facial expression recognition in video with multiple feature fusion,” IEEE Transactions on Affective Computing , vol. 9, no. 1, pp. 38–50, 2018
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
M. Sandler, A. G. Howard, M. Zhu, A. Zhmoginov, and L.-C. Chen, “Mobilenetv2: Inverted residuals and linear bottlenecks,” 2018 IEEE/CVF Conference on Computer Vision and Pattern Recognition , pp. 4510–4520, 2018
2018
Earlier work this paper cites.
T. Wilhelm, “Towards facial expression analysis in a driver assistance system,” 2019 14th IEEE International Conference on Automatic Face & Gesture Recognition (FG 2019) , pp. 1–4, 2019
2019
Earlier work this paper cites.
J. Kossaifi, A. Toisoul, A. Bulat, Y. Panagakis, T. M. Hospedales, and M. Pantic, “Factorized higher-order cnns with an application to spatio-temporal emotion estimation,” 2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , pp. 6059–6068, 2019
2019
Earlier work this paper cites.
N. Houlsby, A. Giurgiu, S. Jastrzebski, B. Morrone, Q. De Laroussilhe, A. Gesmundo, M. Attariyan, and S. Gelly, “Parameter-efficient transfer learning for nlp,” in International Conference on Machine Learning . PMLR, 2019, pp. 2790–2799
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
X. Jiang, Y. Zong, W. Zheng, C. Tang, W. Xia, C. Lu, and J. Liu, “Dfew: A large-scale database for recognizing dynamic facial expressions in the wild,” in Proceedings of the 28th ACM International Conference on Multimedia , 2020, pp. 2881–2889
2020
Earlier work this paper cites.
L. Sun, Z. Lian, J. Tao, B. Liu, and M. Niu, “Multi-modal continuous dimensional emotion recognition using recurrent neural network and self-attention mechanism,” Proceedings of the 1st International on Multimodal Sentiment Analysis in Real-life Media Challenge and Workshop , 2020
2020
Earlier work this paper cites.
L. Liang, C. Lang, Y. Li, S. Feng, and J. Zhao, “Fine-grained facial expression recognition in the wild,” IEEE Transactions on Information Forensics and Security , vol. 16, pp. 482–494, 2020
2020
Earlier work this paper cites.
A. Dosovitskiy, L. Beyer, A. Kolesnikov, D. Weissenborn, X. Zhai, T. Unterthiner, M. Dehghani, M. Minderer, G. Heigold, S. Gelly et al. , “An image is worth 16x16 words: Transformers for image recognition at scale,” International Conference on Learning Representations , 2020
2020
Cited alongside, same era.
J. Deng, J. Guo, E. Ververas, I. Kotsia, and S. Zafeiriou, “Retinaface: Single-shot multi-level face localisation in the wild,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2020, pp. 5203–5212
2020
Cited alongside, same era.
S. I. Serengil and A. Ozpinar, “Lightface: A hybrid deep face recognition framework,” in 2020 Innovations in Intelligent Systems and Applications Conference (ASYU) . IEEE, 2020, pp. 23–27
2020
Cited alongside, same era.
K. Wang, X. Peng, J. Yang, S. Lu, and Y. Qiao, “Suppressing uncertainties for large-scale facial expression recognition,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2020, pp. 6897–6906
K. He, X. Chen, S. Xie, Y. Li, P. Dollár, and R. Girshick, “Masked autoencoders are scalable vision learners,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2022, pp. 16 000–16 009
2022
Later among the works it cites.
Y. Zhang, C. Wang, X. Ling, and W. Deng, “Learn from all: Erasing attention consistency for noisy label facial expression recognition,” in Computer Vision–ECCV 2022: 17th European Conference, Tel Aviv, Israel, October 23–27, 2022, Proceedings, Part XXVI . Springer, 2022, pp. 418–434
2022
Later among the works it cites.
F. Xue, Q. Wang, Z. Tan, Z. Ma, and G. Guo, “Vision transformer with attentive pooling for robust facial expression recognition,” IEEE Transactions on Affective Computing , 2022
2022
Later among the works it cites.
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2020
Cited alongside, same era.
Z. Zhao and Q. Liu, “Former-dfer: Dynamic facial expression recognition transformer,” Proceedings of the 29th ACM International Conference on Multimedia , 2021
2021
Cited alongside, same era.
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark et al. , “Learning transferable visual models from natural language supervision,” in International conference on machine learning . PMLR, 2021, pp. 8748–8763
2021
Cited alongside, same era.
Y. Chen and J. Joo, “Understanding and mitigating annotation bias in facial expression recognition,” in Proceedings of the IEEE/CVF International Conference on Computer Vision , 2021, pp. 14 980–14 991
2021
Cited alongside, same era.
2021
Cited alongside, same era.
2021
Cited alongside, same era.
2021
Cited alongside, same era.
J. She, Y. Hu, H. Shi, J. Wang, Q. Shen, and T. Mei, “Dive into ambiguity: Latent distribution mining and pairwise uncertainty estimation for facial expression recognition,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2021, pp. 6248–6257
2021
Cited alongside, same era.
2021
Cited alongside, same era.
Z. Tong, Y. Song, J. Wang, and L. Wang, “Videomae: Masked autoencoders are data-efficient learners for self-supervised video pre-training,” Advances in neural information processing systems , vol. 35, pp. 10 078–10 093, 2022
2022
Later among the works it cites.
Y. Liu, W. Wang, C. Feng, H. Zhang, Z. Chen, and Y. Zhan, “Expression snippet transformer for robust video-based facial expression recognition,” Pattern Recognition , vol. 138, p. 109368, 2023
2023
Closest in time.
H. Li, H. Niu, Z. Zhu, and F. Zhao, “Intensity-aware loss for dynamic facial expression recognition in the wild,” in Proceedings of the AAAI Conference on Artificial Intelligence , vol. 37, no. 1, 2023, pp. 67–75
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
L. Sun, Z. Lian, B. Liu, and J. Tao, “Mae-dfer: Efficient masked autoencoder for self-supervised dynamic facial expression recognition,” pp. 6110–6121, 2023
2023
Closest in time.
J. Li, Y. Chen, X. Zhang, J. Nie, Z. Li, Y. Yu, Y. Zhang, R. Hong, and M. Wang, “Multimodal feature extraction and fusion for emotional reaction intensity estimation and expression classification in videos with transformers,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 5837–5843
2023
Closest in time.
Z. Cai, S. Ghosh, K. Stefanov, A. Dhall, J. Cai, H. Rezatofighi, R. Haffari, and M. Hayat, “Marlin: Masked autoencoder for facial video representation learning,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , June 2023, pp. 1493–1504
2023
Closest in time.
G. Chen, X. Liu, G. Wang, K. Zhang, P. H. Torr, X.-P. Zhang, and Y. Tang, “Tem-adapter: Adapting image-text pretraining for video question answer,” in Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) , October 2023, pp. 13 945–13 955
2023
Closest in time.
R. Belmonte, B. Allaert, P. Tirilly, I. M. Bilasco, C. Djeraba, and N. Sebe, “Impact of facial landmark localization on facial expression recognition,” IEEE Transactions on Affective Computing , vol. 14, no. 2, pp. 1267–1279, 2023
2023
Closest in time.
C. Zheng, M. Mendieta, and C. Chen, “Poster: A pyramid cross-fusion transformer network for facial expression recognition,” 2023
2023
Closest in time.
J. Li, Y. Chen, X. Zhang, J. Nie, Z. Li, Y. Yu, Y. Zhang, R. Hong, and M. Wang, “Multimodal feature extraction and fusion for emotional reaction intensity estimation and expression classification in videos with transformers,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) Workshops , June 2023, pp. 5838–5844
2023
Closest in time.
D. Kollias, P. Tzirakis, A. Baird, A. Cowen, and S. Zafeiriou, “Abaw: Valence-arousal estimation, expression recognition, action unit detection & emotional reaction intensity estimation challenges,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) Workshops , June 2023, pp. 5889–5898
2023
Closest in time.
Z. Qing, S. Zhang, Z. Huang, X. Wang, Y. Wang, Y. Lv, C. Gao, and N. Sang, “Mar: Masked autoencoders for efficient action recognition,” IEEE Transactions on Multimedia , 2023
2023
Closest in time.
J. Zhu, S. Lai, X. Chen, D. Wang, and H. Lu, “Visual prompt multi-modal tracking,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 9516–9526
2023
Closest in time.
K. Sohn, H. Chang, J. Lezama, L. Polania, H. Zhang, Y. Hao, I. Essa, and L. Jiang, “Visual prompt tuning for generative transfer learning,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 19 840–19 851
2023
Closest in time.
H. Wang, B. Li, S. Wu, S. Shen, F. Liu, S. Ding, and A. Zhou, “Rethinking the learning paradigm for dynamic facial expression recognition,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, pp. 17 958–17 968
2023
Closest in time.
2023
Closest in time.
N. Le, K. Nguyen, Q. Tran, E. Tjiputra, B. Le, and A. Nguyen, “Uncertainty-aware label distribution learning for facial expression recognition,” in Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision , 2023, pp. 6088–6097
2023
Closest in time.
F. Ma, B. Sun, and S. Li, “Transformer-augmented network with online label correction for facial expression recognition,” IEEE Transactions on Affective Computing , 2023
2023
Closest in time.