Fetching the paper…
Reading the bibliography…
Synthesising 3D facial motion from speech is a crucial problem manifesting in a multitude of applications such as computer games and movies.
Using hmms and anns for mapping acoustic to visual speech
G. Salvi · 1999
Earlier work this paper cites.
Hmm-based text-to-audio-visual speech synthesis
S. Sako, K. Tokuda, T. Masuko, T. Kobayashi, and T. Kitamura · 2000
Earlier work this paper cites.
Web-based database for facial expression analysis
M. Pantic, M. Valstar, R. Rademaker, and L. Maat · 2005
Earlier work this paper cites.
A 3d facial expression database for facial behavior research
L. Yin, X. Wei, Y. Sun, J. Wang, and M. J. Rosato · 2006
Earlier work this paper cites.
Optimal step nonrigid icp algorithms for surface registration
B. Amberg, S. Romdhani, and T. Vetter · 2007
Earlier work this paper cites.
Assembling an expressive facial animation system
A. Wang, M. Emmi, and P. Faloutsos · 2007
Earlier work this paper cites.
Realistic mouth-synching for speech-driven talking face using articulatory modelling
L. Xie and Z.-Q. Liu · 2007
Earlier work this paper cites.
Consistency of trace norm minimization
F. R. Bach · 2008
Earlier work this paper cites.
A high-resolution 3d dynamic facial expression database
L. Yin, X. Chen, Y. Sun, T. Worm, and M. Reale · 2008
Earlier work this paper cites.
Face active appearance modeling and speech acoustic information to recover articulation
A. Katsamanis, G. Papandreou, and P. Maragos · 2009
Earlier work this paper cites.
3d morphable face models revisited
A. Patel and W. A. Smith · 2009
Earlier work this paper cites.
Synface: speech-driven facial animation for virtual speech-reading support
G. Salvi, J. Beskow, S. Al Moubayed, and B. Granström · 2009
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
X. Glorot and Y. Bengio · 2010
Earlier work this paper cites.
Point set registration: Coherent point drift
A. Myronenko and X. Song · 2010
Earlier work this paper cites.
A facs valid 3d dynamic action unit database with applications to 3d dynamic morphable facial modeling
D. Cosker, E. Krumhuber, and A. Hilton · 2011
Cited alongside, same era.
Dynamic units of visual speech
S. L. Taylor, M. Mahler, B.-J. Theobald, and I. Matthews · 2012
Cited alongside, same era.
Generalized time warping for multi-modal alignment of human motion
F. Zhou and F. De la Torre · 2012
Cited alongside, same era.
Deep canonical correlation analysis
G. Andrew, R. Arora, J. Bilmes, and K. Livescu · 2013
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
D. Bahdanau, K. Cho, and Y. Bengio · 2014
Cited alongside, same era.
Joint face detection and alignment using multitask cascaded convolutional networks
K. Zhang, Z. Zhang, Z. Li, and Y. Qiao · 2016
Later among the works it cites.
Lip reading sentences in the wild
J. S. Chung, A. Senior, O. Vinyals, and A. Zisserman · 2017
Later among the works it cites.
Speech segmentation with a neural encoder model of working memory
M. Elsner and C. Shain · 2017
Later among the works it cites.
Audio-driven facial animation by joint end-to-end learning of pose and emotion
T. Karras, T. Aila, S. Laine, A. Herva, and J. Lehtinen · 2017
Later among the works it cites.
Online and linear-time attention by enforcing monotonic alignments
C. Raffel, M.-T. Luong, P. J. Liu, R. J. Weiss, and D. Eck · 2017
Later among the works it cites.
Synthesizing obama: learning lip sync from audio
S. Suwajanakorn, S. M. Seitz, and I. Kemelmacher-Shlizerman · 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
D. Kingma and J. Ba · 2014
Cited alongside, same era.
Photo-real talking head with deep bidirectional lstm
B. Fan, L. Wang, F. K. Soong, and L. Xie · 2015
Cited alongside, same era.
Delving deep into rectifiers: Surpassing human-level performance on imagenet classification
K. He, X. Zhang, S. Ren, and J. Sun · 2015
Cited alongside, same era.
Effective approaches to attention-based neural machine translation
M.-T. Luong, H. Pham, and C. D. Manning · 2015
Cited alongside, same era.
4d cardiff conversation database (4d ccdb): A 4d database of natural, dyadic conversations
A. D. Marshall, P. L. Rosin, J. Vandeventer, and A. Aubrey · 2015
Cited alongside, same era.
Audiovisual speech synthesis: An overview of the state-of-the-art
W. Mattheyses and W. Verhelst · 2015
Cited alongside, same era.
A 3d morphable model learnt from 10,000 faces
J. Booth, A. Roussos, S. Zafeiriou, A. Ponniah, and D. Dunaway · 2016
Cited alongside, same era.
Later among the works it cites.
A deep learning approach for generalized speech animation
S. Taylor, T. Kim, Y. Yue, M. Mahler, J. Krahe, A. G. Rodriguez, J. Hodgins, and I. Matthews · 2017
Later among the works it cites.
The 3d menpo facial landmark tracking challenge
S. Zafeiriou, G. G. Chrysos, A. Roussos, E. Ververas, J. Deng, and G. Trigeorgis · 2017
Later among the works it cites.
Large scale 3d morphable models
J. Booth, A. Roussos, A. Ponniah, D. Dunaway, and S. Zafeiriou · 2018
Later among the works it cites.
4dfab: A large scale 4d database for facial expression analysis and biometric applications
S. Cheng, I. Kotsia, M. Pantic, and S. Zafeiriou · 2018
Later among the works it cites.
End-to-end learning for 3d facial animation from speech
H. X. Pham, Y. Wang, and V. Pavlovic · 2018
Later among the works it cites.
Deep canonical time warping for simultaneous alignment and representation learning of sequences
G. Trigeorgis, M. A. Nicolaou, B. W. Schuller, and S. Zafeiriou · 2018
Later among the works it cites.
Visemenet: Audio-driven animator-centric speech animation
Y. Zhou, Z. Xu, C. Landreth, E. Kalogerakis, S. Maji, and K. Singh · 2018
Later among the works it cites.