Fetching the paper…
Reading the bibliography…
The recent state of the art on monocular 3D face reconstruction from image data has made some impressive advancements, thanks to the advent of Deep Learning.
Darpa timit acoustic-phonetic continous speech corpus cd-rom. nist speech disc 1-1.1
John S Garofolo, Lori F Lamel, William M Fisher, Jonathan G Fiscus, and David S Pallett · 1993
Earlier work this paper cites.
3d lip shapes from video: A combined physical–statistical model
Sumit Basu, Nuria Oliver, and Alex Pentland · 1998
Earlier work this paper cites.
3d modeling and tracking of human lip motions
Sumit Basu, Nuria Oliver, and Alex Pentland · 1998
Earlier work this paper cites.
A morphable model for the synthesis of 3d faces
Volker Blanz and Thomas Vetter · 1999
Earlier work this paper cites.
Lip animation based on observed 3D speech dynamics
Gregor A. Kalberer and Luc J. Van Gool · 2000
Earlier work this paper cites.
Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks
Alex Graves, Santiago Fernández, Faustino Gomez, and Jürgen Schmidhuber · 2006
Earlier work this paper cites.
A 3d face model for pose and illumination invariant face recognition
Pascal Paysan, Reinhard Knothe, Brian Amberg, Sami Romdhani, and Thomas Vetter · 2009
Earlier work this paper cites.
Real-time mouth tracking and 3d reconstruction
Jie Cheng and Peisen Huang · 2010
Earlier work this paper cites.
Inverse rendering of faces with a 3d morphable model
Oswald Aldrian and William AP Smith · 2012
Earlier work this paper cites.
The uncanny valley [from the field]
Masahiro Mori, Karl F MacDorman, and Norri Kageki · 2012
Earlier work this paper cites.
Facewarehouse: A 3d facial expression database for visual computing
Chen Cao, Yanlin Weng, Shun Zhou, Yiying Tong, and Kun Zhou · 2013
Earlier work this paper cites.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Earlier work this paper cites.
Amazon Polly. Developer Guide., 2015
2015
Earlier work this paper cites.
Real-time high-fidelity facial performance capture
Chen Cao, Derek Bradley, Kun Zhou, and Thabo Beeler · 2015
Earlier work this paper cites.
Tcd-timit: An audio-visual corpus of continuous speech
Naomi Harte and Eoin Gillen · 2015
Earlier work this paper cites.
Real-time expression transfer for facial reenactment
Justus Thies, Michael Zollhöfer, Matthias Nießner, Levi Valgaerts, Marc Stamminger, and Christian Theobalt · 2015
Earlier work this paper cites.
Reconstruction of personalized 3d face rigs from monocular video
Pablo Garrido, Michael Zollhöfer, Dan Casas, Levi Valgaerts, Kiran Varanasi, Patrick Pérez, and Christian Theobalt · 2016
Earlier work this paper cites.
Corrective 3d reconstruction of lips from monocular video
Pablo Garrido, Michael Zollhöfer, Chenglei Wu, Derek Bradley, Patrick Pérez, Thabo Beeler, and Christian Theobalt · 2016
Earlier work this paper cites.
Real-time 3d face fitting and texture fusion on in-the-wild videos
Patrik Huber, Philipp Kopp, William Christmas, Matthias Rätsch, and Josef Kittler · 2016
Earlier work this paper cites.
Large-pose face alignment via cnn-based dense 3d model fitting
Amin Jourabloo and Xiaoming Liu · 2016
Earlier work this paper cites.
300 faces in-the-wild challenge: Database and results
Christos Sagonas, Epameinondas Antonakos, Georgios Tzimiropoulos, Stefanos Zafeiriou, and Maja Pantic · 2016
Earlier work this paper cites.
Real-time facial segmentation and performance capture from rgb input
Shunsuke Saito, Tianye Li, and Hao Li · 2016
Cited alongside, same era.
Facevr: Real-time facial reenactment and eye gaze control in virtual reality
Justus Thies, Michael Zollhöfer, Marc Stamminger, Christian Theobalt, and Matthias Nießner · 2016
Cited alongside, same era.
How far are we from solving the 2d & 3d face alignment problem?(and a dataset of 230,000 3d facial landmarks)
Adrian Bulat and Georgios Tzimiropoulos · 2017
Cited alongside, same era.
Large pose 3d face reconstruction from a single image via direct volumetric cnn regression
Aaron S Jackson, Adrian Bulat, Vasileios Argyriou, and Georgios Tzimiropoulos · 2017
Cited alongside, same era.
Learning a model of facial shape and expression from 4D scans
Tianye Li, Timo Bolkart, Michael. J. Black, Hao Li, and Javier Romero · 2017
Cited alongside, same era.
Meshgan: Non-linear 3d morphable models of faces
Shiyang Cheng, Michael Bronstein, Yuxiang Zhou, Irene Kotsia, Maja Pantic, and Stefanos Zafeiriou · 2019
Later among the works it cites.
Ganfit: Generative adversarial network fitting for high fidelity 3d face reconstruction
Baris Gecer, Stylianos Ploumpis, Irene Kotsia, and Stefanos Zafeiriou · 2019
Later among the works it cites.
Learning to regress 3d face shape and expression from an image without 3d supervision
Soubhik Sanyal, Timo Bolkart, Haiwen Feng, and Michael Black · 2019
Later among the works it cites.
3d morphable face models—past, present, and future
Bernhard Egger, William A. P. Smith, Ayush Tewari, Stefanie Wuhrer, Michael Zollhoefer, Thabo Beeler, Florian Bernard, Timo Bolkart, Adam Kortylewski, Sami Romdhani, Christian Theobalt, Volker Blanz, and Thomas Vetter · 2020
Later among the works it cites.
Towards fast, accurate and stable 3d dense face alignment
Jianzhu Guo, Xiangyu Zhu, Yang Yang, Fan Yang, Zhen Lei, and Stan Z Li · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
MoFA: Model-based Deep Convolutional Face Autoencoder for Unsupervised Monocular Reconstruction
Ayush Tewari, Michael Zollöfer, Hyeongwoo Kim, Pablo Garrido, Florian Bernard, Patrick Perez, and Christian Theobalt · 2017
Cited alongside, same era.
Face alignment in full pose range: A 3d total solution
Xiangyu Zhu, Xiaoming Liu, Zhen Lei, and Stan Z Li · 2017
Cited alongside, same era.
Lrs3-ted: a large-scale dataset for visual speech recognition
Triantafyllos Afouras, Joon Son Chung, and Andrew Zisserman · 2018
Cited alongside, same era.
Threat of adversarial attacks on deep learning in computer vision: A survey
Naveed Akhtar and Ajmal Mian · 2018
Cited alongside, same era.
Modeling facial geometry using compositional vaes
Timur Bagautdinov, Chenglei Wu, Jason Saragih, Pascal Fua, and Yaser Sheikh · 2018
Cited alongside, same era.
Large scale 3d morphable models
James Booth, Anastasios Roussos, Allan Ponniah, David Dunaway, and Stefanos Zafeiriou · 2018
Cited alongside, same era.
3d reconstruction of “in-the-wild” faces in images and videos
James Booth, Anastasios Roussos, Evangelos Ververas, Epameinondas Antonakos, Stylianos Ploumpis, Yannis Panagakis, and Stefanos Zafeiriou · 2018
Cited alongside, same era.
The role of plausibility in the experience of spatial presence in virtual environments
Matthias Hofer, Tilo Hartmann, Allison Eden, Rabindra Ratan, and Lindsay Hahn · 2020
Later among the works it cites.
Emotion recognition in immersive virtual reality: From statistics to affective computing
Javier Marín-Morales, Carmen Llinares, Jaime Guixeres, and Mariano Alcañiz · 2020
Later among the works it cites.
Mead: A large-scale audio-visual dataset for emotional talking-face generation
Kaisiyuan Wang, Qianyi Wu, Linsen Song, Zhuoqian Yang, Wayne Wu, Chen Qian, Ran He, Yu Qiao, and Chen Change Loy · 2020
Later among the works it cites.
Facescape: a large-scale high quality 3d face dataset and detailed riggable 3d face prediction
Haotian Yang, Hao Zhu, Yanru Wang, Mingkai Huang, Qiu Shen, Ruigang Yang, and Xun Cao · 2020
Later among the works it cites.
Learning an animatable detailed 3d face model from in-the-wild images
Yao Feng, Haiwen Feng, Michael J Black, and Timo Bolkart · 2021
Later among the works it cites.
Neural head avatars from monocular rgb videos
Philip-William Grassal, Malte Prinzler, Titus Leistner, Carsten Rother, Matthias Nießner, and Justus Thies · 2021
Later among the works it cites.
Emoca: Emotion driven monocular face capture and animation
Radek Daněček, Michael J Black, and Timo Bolkart · 2022
Closest in time.
Danceconv: Dance motion generation with convolutional networks
Kosmas Kritsis, Aggelos Gkiokas, Aggelos Pikrakis, and Vassilis Katsouros · 2022
Closest in time.
Visual Speech Recognition for Multiple Languages in the Wild
Pingchuan Ma, Stavros Petridis, and Maja Pantic · 2022
Closest in time.
Dad-3dheads: A large-scale dense, accurate and diverse dataset for 3d head alignment from a single image
Tetiana Martyniuk, Orest Kupyn, Yana Kurlyak, Igor Krashenyi, Jiři Matas, and Viktoriia Sharmanska · 2022
Closest in time.
Learning audio-visual speech representation by masked multimodal cluster prediction
Bowen Shi, Wei-Ning Hsu, Kushal Lakhotia, and Abdelrahman Mohamed · 2022
Closest in time.
Robust self-supervised audio-visual speech recognition
Bowen Shi, Wei-Ning Hsu, and Abdelrahman Mohamed · 2022
Closest in time.
The effect of virtual human rendering style on user perceptions of visual cues
Jacob Stuart, Karen Aul, Anita Stephen, Michael D Bumbach, and Benjamin Lok · 2022
Closest in time.
Faceverse: a fine-grained and detail-controllable 3d face morphable model from a hybrid dataset
Lizhen Wang, Zhiyua Chen, Tao Yu, Chenguang Ma, Liang Li, and Yebin Liu · 2022
Closest in time.
Towards metrical reconstruction of human faces
Wojciech Zielonka, Timo Bolkart, and Justus Thies · 2022
Closest in time.