Fetching the paper…
Reading the bibliography…
Facial action unit (AU) detection is challenging due to the difficulty in capturing correlated information from subtle and dynamic AUs.
J. Pearl, Probabilistic reasoning in intelligent systems: Networks of plausible inference . Morgan Kaufmann, 1988
1988
Earlier work this paper cites.
P. Ekman and E. L. Rosenberg, What the face reveals: Basic and applied studies of spontaneous expression using the Facial Action Coding System (FACS) . Oxford University Press, USA, 1997
1997
Earlier work this paper cites.
D. G. Lowe, “Object recognition from local scale-invariant features,” in IEEE International Conference on Computer Vision . IEEE, 1999, pp. 1150–1157
1999
Earlier work this paper cites.
T. F. Cootes, G. J. Edwards, and C. J. Taylor, “Active appearance models,” IEEE Transactions on Pattern Analysis and Machine Intelligence , vol. 23, no. 6, pp. 681–685, 2001
2001
Earlier work this paper cites.
Y. Tong and Q. Ji, “Learning bayesian networks with qualitative constraints,” in IEEE Conference on Computer Vision and Pattern Recognition . IEEE, 2008, pp. 1–8
2008
Earlier work this paper cites.
R.-E. Fan, K.-W. Chang, C.-J. Hsieh, X.-R. Wang, and C.-J. Lin, “Liblinear: A library for large linear classification,” Journal of Machine Learning Research , vol. 9, no. Aug, pp. 1871–1874, 2008
2008
Earlier work this paper cites.
D. E. King, “Dlib-ml: A machine learning toolkit,” Journal of Machine Learning Research , vol. 10, pp. 1755–1758, 2009
2009
Earlier work this paper cites.
V. Nair and G. E. Hinton, “Rectified linear units improve restricted boltzmann machines,” in International Conference on Machine Learning , 2010, pp. 807–814
2010
Earlier work this paper cites.
P. Krähenbühl and V. Koltun, “Efficient inference in fully connected crfs with gaussian edge potentials,” in Advances in Neural Information Processing Systems , 2011, pp. 109–117
2011
Earlier work this paper cites.
A. Krizhevsky, I. Sutskever, and G. E. Hinton, “Imagenet classification with deep convolutional neural networks,” in Advances in Neural Information Processing Systems . Curran Associates, Inc., 2012, pp. 1097–1105
2012
Earlier work this paper cites.
Y. Li, S. Wang, Y. Zhao, and Q. Ji, “Simultaneous facial feature tracking and facial expression recognition,” IEEE Transactions on Image Processing , vol. 22, no. 7, pp. 2559–2573, 2013
2013
Earlier work this paper cites.
Z. Wang, S. Wang, and Q. Ji, “Capturing complex spatio-temporal relations among facial muscles for facial expression recognition,” in IEEE Conference on Computer Vision and Pattern Recognition . IEEE, 2013, pp. 3422–3429
2013
Earlier work this paper cites.
Z. Wang, Y. Li, S. Wang, and Q. Ji, “Capturing global semantic relationships for facial action unit recognition,” in IEEE International Conference on Computer Vision . IEEE, 2013, pp. 3304–3311
2013
Earlier work this paper cites.
S. M. Mavadati, M. H. Mahoor, K. Bartlett, P. Trinh, and J. F. Cohn, “Disfa: A spontaneous facial action intensity database,” IEEE Transactions on Affective Computing , vol. 4, no. 2, pp. 151–160, 2013
2013
Earlier work this paper cites.
X. Xiong and F. De la Torre, “Supervised descent method and its applications to face alignment,” in IEEE Conference on Computer Vision and Pattern Recognition . IEEE, 2013, pp. 532–539
2013
Earlier work this paper cites.
I. Sutskever, J. Martens, G. Dahl, and G. Hinton, “On the importance of initialization and momentum in deep learning,” in International Conference on Machine Learning , 2013, pp. 1139–1147
2013
Earlier work this paper cites.
X. Zhang, L. Yin, J. F. Cohn, S. Canavan, M. Reale, A. Horowitz, P. Liu, and J. M. Girard, “Bp4d-spontaneous: A high-resolution spontaneous 3d dynamic facial expression database,” Image and Vision Computing , vol. 32, no. 10, pp. 692–706, 2014
2014
Earlier work this paper cites.
K. Cho, B. van Merriënboer, D. Bahdanau, and Y. Bengio, “On the properties of neural machine translation: Encoder–decoder approaches,” in Conference on Empirical Methods in Natural Language Processing Workshops . Association for Computational Linguistics, 2014, pp. 103–111
2014
Earlier work this paper cites.
M. Lin, Q. Chen, and S. Yan, “Network in network,” in International Conference on Learning Representations , 2014
2014
Earlier work this paper cites.
V. Kazemi and J. Sullivan, “One millisecond face alignment with an ensemble of regression trees,” in IEEE Conference on Computer Vision and Pattern Recognition . IEEE, 2014, pp. 1867–1874
2014
Earlier work this paper cites.
C. Cao, X. Liu, Y. Yang, Y. Yu, J. Wang, Z. Wang, Y. Huang, L. Wang, C. Huang, W. Xu, D. Ramanan, and T. S. Huang, “Look and think twice: Capturing top-down visual attention with feedback convolutional neural networks,” in IEEE International Conference on Computer Vision . IEEE, 2015, pp. 2956–2964
2015
Earlier work this paper cites.
D. P. Kingma and J. Ba, “Adam: A method for stochastic optimization,” in International Conference on Learning Representations , 2015
2015
Earlier work this paper cites.
K. Simonyan and A. Zisserman, “Very deep convolutional networks for large-scale image recognition,” in International Conference on Learning Representations , 2015
2015
Cited alongside, same era.
S. Zheng, S. Jayasumana, B. Romera-Paredes, V. Vineet, Z. Su, D. Du, C. Huang, and P. H. Torr, “Conditional random fields as recurrent neural networks,” in IEEE International Conference on Computer Vision . IEEE, 2015, pp. 1529–1537
2015
Cited alongside, same era.
K. Zhao, W.-S. Chu, F. De la Torre, J. F. Cohn, and H. Zhang, “Joint patch and multi-label learning for facial action unit and holistic expression recognition,” IEEE Transactions on Image Processing , vol. 25, no. 8, pp. 3931–3946, 2016
2016
Cited alongside, same era.
K. Zhao, W.-S. Chu, and H. Zhang, “Deep region and multi-label learning for facial action unit detection,” in IEEE Conference on Computer Vision and Pattern Recognition . IEEE, 2016, pp. 3391–3399
2016
Cited alongside, same era.
G. Li, X. Zhu, Y. Zeng, Q. Wang, and L. Lin, “Semantic relationships guided representation learning for facial action unit recognition,” in AAAI Conference on Artificial Intelligence , vol. 33, no. 01, 2019, pp. 8594–8601
2019
Later among the works it cites.
C. Ma, L. Chen, and J. Yong, “Au r-cnn: Encoding expert prior knowledge into r-cnn for action unit detection,” Neurocomputing , vol. 355, pp. 35–47, 2019
2019
Later among the works it cites.
J. P. Robinson, Y. Li, N. Zhang, Y. Fu, and S. Tulyakov, “Laplace landmark localization,” in IEEE International Conference on Computer Vision . IEEE, 2019, pp. 10 103–10 112
2019
Later among the works it cites.
D. Kollias and S. Zafeiriou, “Expression, affect, action unit recognition: Aff-wild2, multi-task learning and arcface,” in British Machine Vision Conference . BMVA Press, 2019, p. 297
2019
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Kuen, Z. Wang, and G. Wang, “Recurrent attentional networks for saliency detection,” in IEEE Conference on Computer Vision and Pattern Recognition . IEEE, 2016, pp. 3668–3677
2016
Cited alongside, same era.
M. Defferrard, X. Bresson, and P. Vandergheynst, “Convolutional neural networks on graphs with fast localized spectral filtering,” in Advances in Neural Information Processing Systems . Curran Associates, Inc., 2016, pp. 3837–3845
2016
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in IEEE Conference on Computer Vision and Pattern Recognition . IEEE, 2016, pp. 770–778
2016
Cited alongside, same era.
C. Szegedy, V. Vanhoucke, S. Ioffe, J. Shlens, and Z. Wojna, “Rethinking the inception architecture for computer vision,” in IEEE Conference on Computer Vision and Pattern Recognition . IEEE, 2016, pp. 2818–2826
2016
Cited alongside, same era.
W. Li, F. Abtahi, and Z. Zhu, “Action unit detection with region adaptation, multi-labeling learning and optimal temporal fusing,” in IEEE Conference on Computer Vision and Pattern Recognition . IEEE, 2017, pp. 6766–6775
2017
Cited alongside, same era.
M. Pedersoli, T. Lucas, C. Schmid, and J. J. Verbeek, “Areas of attention for image captioning,” in IEEE International Conference on Computer Vision . IEEE, 2017, pp. 1242–1250
2017
Cited alongside, same era.
W.-S. Chu, F. De la Torre, and J. F. Cohn, “Learning spatial and temporal cues for multi-label facial action unit detection,” in IEEE International Conference on Automatic Face and Gesture Recognition . IEEE, 2017, pp. 25–32
2017
Cited alongside, same era.
J. He, D. Li, B. Yang, S. Cao, B. Sun, and L. Yu, “Multi view facial action unit detection based on cnn and blstm-rnn,” in IEEE International Conference on Automatic Face and Gesture Recognition . IEEE, 2017, pp. 848–853
2017
Cited alongside, same era.
N. Sankaran, D. D. Mohan, S. Setlur, V. Govindaraju, and D. Fedorishin, “Representation learning through cross-modality supervision,” in IEEE International Conference on Automatic Face and Gesture Recognition . IEEE, 2019, pp. 1–8
2019
Later among the works it cites.
A. Paszke, S. Gross, F. Massa, A. Lerer, J. Bradbury, G. Chanan, T. Killeen, Z. Lin, N. Gimelshein, L. Antiga, A. Desmaison, A. Kopf, E. Yang, Z. DeVito, M. Raison, A. Tejani, S. Chilamkurthy, B. Steiner, L. Fang, J. Bai, and S. Chintala, “Pytorch: An imperative style, high-performance deep learning library,” in Advances in Neural Information Processing Systems . Curran Associates, Inc., 2019, pp. 8024–8035
2019
Later among the works it cites.
Y. Li, J. Zeng, S. Shan, and X. Chen, “Self-supervised representation learning from videos for facial action unit detection,” in IEEE Conference on Computer Vision and Pattern Recognition . IEEE, 2019, pp. 10 924–10 933
2019
Later among the works it cites.
Z. Liu, J. Dong, C. Zhang, L. Wang, and J. Dang, “Relation modeling with graph convolutional networks for facial action unit detection,” in International Conference on Multimedia Modeling . Springer, 2020, pp. 489–501
2020
Closest in time.
I. O. Ertugrul, J. F. Cohn, L. A. Jeni, Z. Zhang, L. Yin, and Q. Ji, “Crossing domains for au coding: Perspectives, approaches, and measures,” IEEE Transactions on Biometrics, Behavior, and Identity Science , vol. 2, no. 2, pp. 158–171, 2020
2020
Closest in time.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin, “Attention is all you need,” in Advances in Neural Information Processing Systems . Curran Associates, Inc., 2017, pp. 5998–6008
2020
Closest in time.
Z. Shao, Z. Liu, J. Cai, and L. Ma, “Jâa-net: Joint facial action unit detection and face alignment via adaptive attention,” International Journal of Computer Vision , vol. 129, no. 2, pp. 321–340, 2021
2021
Closest in time.
T. Song, L. Chen, W. Zheng, and Q. Ji, “Uncertain graph neural networks for facial action unit detection,” in AAAI Conference on Artificial Intelligence , vol. 35, no. 7, 2021, pp. 5993–6001
2021
Closest in time.
Y. Fan and Z. Lin, “G2rl: geometry-guided representation learning for facial action unit intensity estimation,” in International Joint Conference on Artificial Intelligence , 2021, pp. 731–737
2021
Closest in time.
T. Song, Z. Cui, W. Zheng, and Q. Ji, “Hybrid message passing with performance-driven structures for facial action unit detection,” in IEEE Conference on Computer Vision and Pattern Recognition . IEEE, 2021, pp. 6267–6276
2021
Closest in time.
——, “Analysing affective behavior in the second abaw2 competition,” in IEEE International Conference on Computer Vision Workshops . IEEE, 2021, pp. 3652–3660
2021
Closest in time.
G. M. Jacob and B. Stenger, “Facial action unit detection with transformers,” in IEEE Conference on Computer Vision and Pattern Recognition . IEEE, 2021, pp. 7680–7689
2021
Closest in time.
J. Yan, J. Wang, Q. Li, C. Wang, and S. Pu, “Self-supervised regional and temporal auxiliary tasks for facial action unit recognition,” in ACM International Conference on Multimedia , 2021, pp. 1038–1046
2021
Closest in time.
Z. Li, X. Deng, X. Li, and L. Yin, “Integrating semantic and temporal relationships in facial action unit detection,” in ACM International Conference on Multimedia , 2021, pp. 5519–5527
2021
Closest in time.
W. Zhang, Z. Guo, K. Chen, L. Li, Z. Zhang, Y. Ding, R. Wu, T. Lv, and C. Fan, “Prior aided streaming network for multi-task affective analysis,” in IEEE International Conference on Computer Vision Workshops . IEEE, 2021, pp. 3539–3549
2021
Closest in time.
Z. Shao, Z. Liu, J. Cai, Y. Wu, and L. Ma, “Facial action unit detection using attention and relation learning,” IEEE Transactions on Affective Computing , vol. 13, no. 3, pp. 1274–1289, 2022
2022
Closest in time.
Y. Chen, G. Song, Z. Shao, J. Cai, T.-J. Cham, and J. Zheng, “Geoconv: Geodesic guided convolution for facial action unit recognition,” Pattern Recognition , vol. 122, p. 108355, 2022
2022
Closest in time.
Y. Chang and S. Wang, “Knowledge-driven self-supervised representation learning for facial action unit recognition,” in IEEE Conference on Computer Vision and Pattern Recognition . IEEE, 2022, pp. 20 417–20 426
2022
Closest in time.