Fetching the paper…
Reading the bibliography…
Head avatars animated by visual signals have gained popularity, particularly in cross-driving synthesis where the driver differs from the animated character, a challenging but highly practical approach.
Web-based database for facial expression analysis
Maja Pantic, Michel Valstar, Ron Rademaker, and Ludo Maat · 2005
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, K. Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
The extended cohn-kanade dataset (ck+): A complete dataset for action unit and emotion-specified expression
Patrick Lucey, Jeffrey F Cohn, Takeo Kanade, Jason Saragih, Zara Ambadar, and Iain Matthews · 2010
Earlier work this paper cites.
Crema-d: Crowd-sourced emotional multimodal actors dataset
Houwei Cao, David G Cooper, Michael K Keutmann, Ruben C Gur, Ani Nenkova, and Ragini Verma · 2014
Earlier work this paper cites.
Surrey audio-visual expressed emotion (savee) database
Philip Jackson and SJUoSG Haq · 2014
Earlier work this paper cites.
Deep face recognition
Omkar M. Parkhi, Andrea Vedaldi, and Andrew Zisserman · 2015
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
Karen Simonyan and Andrew Zisserman · 2015
Earlier work this paper cites.
Out of time: automated lip sync in the wild
Joon Son Chung and Andrew Zisserman · 2017
Earlier work this paper cites.
Improved training of wasserstein gans, 2017
Ishaan Gulrajani, Faruk Ahmed, Martin Arjovsky, Vincent Dumoulin, and Aaron Courville · 2017
Earlier work this paper cites.
Gans trained by a two time-scale update rule converge to a local nash equilibrium
Martin Heusel, Hubert Ramsauer, Thomas Unterthiner, Bernhard Nessler, and Sepp Hochreiter · 2017
Earlier work this paper cites.
Unpaired image-to-image translation using cycle-consistent adversarial networks
Jun-Yan Zhu, Taesung Park, Phillip Isola, and Alexei A. Efros · 2017
Earlier work this paper cites.
Vggface2: A dataset for recognising faces across pose and age, 2018
Qiong Cao, Li Shen, Weidi Xie, Omkar M. Parkhi, and Andrew Zisserman · 2018
Earlier work this paper cites.
Voxceleb2: Deep speaker recognition
Joon Son Chung, Arsha Nagrani, and Andrew Zisserman · 2018
Earlier work this paper cites.
Rt-gene: Real-time eye gaze estimation in natural environments
Tobias Fischer, Hyung Jin Chang, and Y. Demiris · 2018
Earlier work this paper cites.
Image-to-image translation with conditional adversarial networks, 2018
Phillip Isola, Jun-Yan Zhu, Tinghui Zhou, and Alexei A. Efros · 2018
Earlier work this paper cites.
The ryerson audio-visual database of emotional speech and song (ravdess): A dynamic, multimodal set of facial and vocal expressions in north american english
Steven R Livingstone and Frank A Russo · 2018
Cited alongside, same era.
Bisenet: Bilateral segmentation network for real-time semantic segmentation, 2018
Changqian Yu, Jingbo Wang, Chao Peng, Changxin Gao, Gang Yu, and Nong Sang · 2018
Cited alongside, same era.
Decoupled weight decay regularization
Ilya Loshchilov and Frank Hutter · 2019
Cited alongside, same era.
First order motion model for image animation
Aliaksandr Siarohin, Stéphane Lathuilière, S. Tulyakov, Elisa Ricci, and N. Sebe · 2019
Cited alongside, same era.
Few-shot adversarial learning of realistic neural talking head models, 2019
Egor Zakharov, Aliaksandra Shysheya, Egor Burkov, and Victor Lempitsky · 2019
Cited alongside, same era.
Pose-controllable talking face generation by implicitly modularized audio-visual representation, 2021
Hang Zhou, Yasheng Sun, Wayne Wu, Chen Change Loy, Xiaogang Wang, and Ziwei Liu · 2021
Later among the works it cites.
Neural head avatars from monocular rgb videos, 2022
Philip-William Grassal, Malte Prinzler, Titus Leistner, Carsten Rother, Matthias Nießner, and Justus Thies · 2022
Later among the works it cites.
Realistic one-shot mesh-based head avatars, 2022
Taras Khakhulin, Vanessa Sklyarova, Victor Lempitsky, and Egor Zakharov · 2022
Later among the works it cites.
Understanding collapse in non-contrastive siamese representation learning, 2022
Alexander C. Li, Alexei A. Efros, and Deepak Pathak · 2022
Later among the works it cites.
Robust speech recognition via large-scale weak supervision, 2022
Alec Radford, Jong Wook Kim, Tao Xu, Greg Brockman, Christine McLeavey, and Ilya Sutskever · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
On the continuity of rotation representations in neural networks
Yi Zhou, Connelly Barnes, Jingwan Lu, Jimei Yang, and Hao Li · 2019
Cited alongside, same era.
Neural head reenactment with latent pose descriptors
Egor Burkov, Igor Pasechnik, Artur Grigorev, and Victor Lempitsky · 2020
Cited alongside, same era.
Dynamic neural radiance fields for monocular 4d facial avatar reconstruction, 2020
Guy Gafni, Justus Thies, Michael Zollhöfer, and Matthias Nießner · 2020
Cited alongside, same era.
A lip sync expert is all you need for speech to lip generation in the wild
KR Prajwal, Rudrabha Mukhopadhyay, Vinay P Namboodiri, and CV Jawahar · 2020
Cited alongside, same era.
Mead: A large-scale audio-visual dataset for emotional talking-face generation
Kaisiyuan Wang, Qianyi Wu, Linsen Song, Zhuoqian Yang, Wayne Wu, Chen Qian, Ran He, Yu Qiao, and Chen Change Loy · 2020
Cited alongside, same era.
Fast bi-layer neural synthesis of one-shot realistic head avatars, 2020
Egor Zakharov, Aleksei Ivakhnenko, Aliaksandra Shysheya, and Victor Lempitsky · 2020
Cited alongside, same era.
MakeltTalk
Yang Zhou, Xintong Han, Eli Shechtman, Jose Echevarria, Evangelos Kalogerakis, and Dingzeyu Li · 2020
Cited alongside, same era.
Fei Yin, Yong Zhang, Xiaodong Cun, Mingdeng Cao, Yanbo Fan, Xuan Wang, Qingyan Bai, Baoyuan Wu, Jue Wang, and Yujiu Yang · 2022
Later among the works it cites.
I m avatar: Implicit morphable head avatars from videos, 2022
Yufeng Zheng, Victoria Fernández Abrevaya, Marcel C. Bühler, Xu Chen, Michael J. Black, and Otmar Hilliges · 2022
Later among the works it cites.
Hyperreenact: One-shot reenactment via jointly learning to refine and retarget faces, 2023
Stella Bounareli, Christos Tzelepis, Vasileios Argyriou, Ioannis Patras, and Georgios Tzimiropoulos · 2023
Later among the works it cites.
Megaportraits: One-shot megapixel neural head avatars, 2023
Nikita Drobyshev, Jenya Chelishev, Taras Khakhulin, Aleksei Ivakhnenko, Victor Lempitsky, and Egor Zakharov · 2023
Later among the works it cites.
Generalizable one-shot neural head avatar, 2023
Xueting Li, Shalini De Mello, Sifei Liu, Koki Nagano, Umar Iqbal, and Jan Kautz · 2023
Later among the works it cites.
Unsupervised volumetric animation
Aliaksandr Siarohin, Willi Menapace, Ivan Skorokhodov, Kyle Olszewski, Hsin-Ying Lee, Jian Ren, Menglei Chai, and Sergey Tulyakov · 2023
Later among the works it cites.
Diffused heads: Diffusion models beat gans on talking-face generation, 2023
Michał Stypułkowski, Konstantinos Vougioukas, Sen He, Maciej Zieba, Stavros Petridis, and Maja Pantic · 2023
Later among the works it cites.
Are 3d face shapes expressive enough for recognising continuous emotions and action unit intensities?, 2023
Mani Kumar Tellamekala, Ömer Sümer, Björn W. Schuller, Elisabeth André, Timo Giesbrecht, and Michel Valstar · 2023
Later among the works it cites.
AvatarMAV: Fast 3d head avatar reconstruction using motion-aware neural voxels
Yuelang Xu, Lizhen Wang, Xiaochen Zhao, Hongwen Zhang, and Yebin Liu · 2023
Later among the works it cites.
Nofa: Nerf-based one-shot facial avatar reconstruction, 2023
Wangbo Yu, Yanbo Fan, Yong Zhang, Xuan Wang, Fei Yin, Yunpeng Bai, Yan-Pei Cao, Ying Shan, Yang Wu, Zhongqian Sun, and Baoyuan Wu · 2023
Later among the works it cites.