Fetching the paper…
Reading the bibliography…
We propose GaussianTalker, a novel framework for real-time generation of pose-controllable talking heads.
Facial Action Coding System: Manual
Paul Ekman and Wallace V. Friesen. 1978 · 1978
Earlier work this paper cites.
Surface splatting. In Proceedings of the 28th annual conference on Computer graphics and interactive techniques . 371–378
Matthias Zwicker, Hanspeter Pfister, Jeroen Van Baar, and Markus Gross. 2001 · 2001
Earlier work this paper cites.
How far are we from solving the 2d & 3d face alignment problem?(and a dataset of 230,000 3d facial landmarks). In Proceedings of the IEEE international conference on computer vision
Adrian Bulat and Georgios Tzimiropoulos. 2017 · 2017
Earlier work this paper cites.
GANs Trained by a Two Time-Scale Update Rule Converge to a Local Nash Equilibrium. In NeurIPS
Martin Heusel, Hubert Ramsauer, Thomas Unterthiner, Bernhard Nessler, and Sepp Hochreiter. 2017 · 2017
Earlier work this paper cites.
Learning a Model of Facial Shape and Expression from 4D Scans
Tianye Li, Timo Bolkart, Michael J. Black, Hao Li, and Javier Romero. 2017 · 2017
Earlier work this paper cites.
Synthesizing obama: learning lip sync from audio
Supasorn Suwajanakorn, Steven M Seitz, and Ira Kemelmacher-Shlizerman. 2017 · 2017
Earlier work this paper cites.
X2Face: A Network for Controlling Face Generation Using Images, Audio, and Pose Codes. In Computer Vision–ECCV 2018: 15th European Conference, Munich, Germany, September 8-14, 2018, Proceedings, Part XIII 15 . Springer, 690–706
Olivia Wiles, A Sophia Koepke, and Andrew Zisserman. 2018 · 2018
Earlier work this paper cites.
The Unreasonable Effectiveness of Deep Features as a Perceptual Metric. In Proceedings of the IEEE conference on computer vision and pattern recognition . 586–595
Richard Zhang, Phillip Isola, Alexei A Efros, Eli Shechtman, and Oliver Wang. 2018 · 2018
Earlier work this paper cites.
Hierarchical Cross-Modal Talking Face Generation With Dynamic Pixel-Wise Loss. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition . 7832–7841
Lele Chen, Ross K Maddox, Zhiyao Duan, and Chenliang Xu. 2019 · 2019
Earlier work this paper cites.
Accurate 3d face reconstruction with weakly-supervised learning: From single image to image set. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition workshops
Yu Deng, Jiaolong Yang, Sicheng Xu, Dong Chen, Yunde Jia, and Xin Tong. 2019 · 2019
Earlier work this paper cites.
You Said That?: Synthesising Talking Faces from Audio
Amir Jamaludin, Joon Son Chung, and Andrew Zisserman. 2019 · 2019
Earlier work this paper cites.
Differentiable surface splatting for point-based geometry processing
Wang Yifan, Felice Serena, Shihao Wu, Cengiz Öztireli, and Olga Sorkine-Hornung. 2019 · 2019
Earlier work this paper cites.
CurricularFace: Adaptive Curriculum Learning Loss for Deep Face Recognition. In CVPR
Yuge Huang, Yuhan Wang, Ying Tai, Xiaoming Liu, Pengcheng Shen, Shaoxin Li, Jilin Li, and Feiyue Huang. 2020 · 2020
Earlier work this paper cites.
NeRF: Representing Scenes as Neural Radiance Fields for View Synthesis. In ECCV
Ben Mildenhall, Pratul P. Srinivasan, Matthew Tancik, Jonathan T. Barron, Ravi Ramamoorthi, and Ren Ng. 2020 · 2020
Earlier work this paper cites.
A Lip Sync Expert Is All You Need for Speech to Lip Generation In the Wild. In Proceedings of the 28th ACM International Conference on Multimedia . 484–492
KR Prajwal, Rudrabha Mukhopadhyay, Vinay P Namboodiri, and CV Jawahar. 2020 · 2020
Earlier work this paper cites.
Neural Voice Puppetry: Audio-Driven Facial Reenactment. In Computer Vision–ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part XVI 16 . Springer, 716–731
Justus Thies, Mohamed Elgharib, Ayush Tewari, Christian Theobalt, and Matthias Nießner. 2020 · 2020
Earlier work this paper cites.
MEAD: A Large-Scale Audio-Visual Dataset for Emotional Talking-Face Generation. In Computer Vision–ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part XXI . Springer, 700–717
Kaisiyuan Wang, Qianyi Wu, Linsen Song, Zhuoqian Yang, Wayne Wu, Chen Qian, Ran He, Yu Qiao, and Chen Change Loy. 2020 · 2020
Earlier work this paper cites.
Multimodal inputs driven talking face generation with spatial–temporal dependency
Lingyun Yu, Jun Yu, Mengyan Li, and Qiang Ling. 2020 · 2020
Earlier work this paper cites.
Makelttalk: speaker-aware talking-head animation
Yang Zhou, Xintong Han, Eli Shechtman, Jose Echevarria, Evangelos Kalogerakis, and Dingzeyu Li. 2020 · 2020
Earlier work this paper cites.
AD-NeRF: Audio Driven Neural Radiance Fields for Talking Head Synthesis. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 5784–5794
Yudong Guo, Keyu Chen, Sen Liang, Yong-Jin Liu, Hujun Bao, and Juyong Zhang. 2021 · 2021
Cited alongside, same era.
Live Speech Portraits: Real-Time Photorealistic Talking-Head Animation
Yuanxun Lu, Jinxiang Chai, and Xun Cao. 2021 · 2021
Cited alongside, same era.
Speech2Talking-Face: Inferring and Driving a Face with Synchronized Audio-Visual Representation.. In IJCAI , Vol. 2. 4
Yasheng Sun, Hang Zhou, Ziwei Liu, and Hideki Koike. 2021 · 2021
Cited alongside, same era.
Flow-Guided One-Shot Talking Face Generation With a High-Resolution Audio-Visual Dataset. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
Zhimeng Zhang, Lincheng Li, Yu Ding, and Changjie Fan. 2021 · 2021
Cited alongside, same era.
Pose-Controllable Talking Face Generation by Implicitly Modularized Audio-Visual Representation. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition . 4176–4186
Im avatar: Implicit morphable head avatars from videos. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 13545–13555
Yufeng Zheng, Victoria Fernández Abrevaya, Marcel C Bühler, Xu Chen, Michael J Black, and Otmar Hilliges. 2022 · 2022
Later among the works it cites.
HexPlane: A Fast Representation for Dynamic Scenes
Ang Cao and Justin Johnson. 2023 · 2023
Later among the works it cites.
Jiazhong Cen, Jiemin Fang, Chen Yang, Lingxi Xie, Xiaopeng Zhang, Wei Shen, and Qi Tian. 2023 · 2023
Later among the works it cites.
Monogaussianavatar: Monocular gaussian point-based head avatar
Yufan Chen, Lizhen Wang, Qijing Li, Hongjiang Xiao, Shengping Zhang, Hongxun Yao, and Yebin Liu. 2023 · 2023
Later among the works it cites.
Headgas: Real-time animatable head avatars via 3d gaussian splatting
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Hang Zhou, Yasheng Sun, Wayne Wu, Chen Change Loy, Xiaogang Wang, and Ziwei Liu. 2021 · 2021
Cited alongside, same era.
Rignerf: Fully controllable neural 3d portraits. In Proceedings of the IEEE/CVF conference on Computer Vision and Pattern Recognition . 20364–20373
ShahRukh Athar, Zexiang Xu, Kalyan Sunkavalli, Eli Shechtman, and Zhixin Shu. 2022 · 2022
Cited alongside, same era.
Efficient Geometry-Aware 3D Generative Adversarial Networks. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 16123–16133
Eric R Chan, Connor Z Lin, Matthew A Chan, Koki Nagano, Boxiao Pan, Shalini De Mello, Orazio Gallo, Leonidas J Guibas, Jonathan Tremblay, Sameh Khamis, et al · 2022
Cited alongside, same era.
Fast Dynamic Radiance Fields with Time-Aware Neural Voxels. In SIGGRAPH Asia 2022 Conference Papers
Jiemin Fang, Taoran Yi, Xinggang Wang, Lingxi Xie, Xiaopeng Zhang, Wenyu Liu, Matthias Nießner, and Qi Tian. 2022 · 2022
Cited alongside, same era.
Reconstructing personalized semantic facial nerf models from monocular video
Xuan Gao, Chenglai Zhong, Jun Xiang, Yang Hong, Yudong Guo, and Juyong Zhang. 2022 · 2022
Cited alongside, same era.
Neural head avatars from monocular rgb videos. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 18653–18664
Philip-William Grassal, Malte Prinzler, Titus Leistner, Carsten Rother, Matthias Nießner, and Justus Thies. 2022 · 2022
Cited alongside, same era.
Realistic one-shot mesh-based head avatars. In European Conference on Computer Vision . Springer, 345–362
Taras Khakhulin, Vanessa Sklyarova, Victor Lempitsky, and Egor Zakharov. 2022 · 2022
Cited alongside, same era.
Semantic-Aware Implicit Neural Audio-Driven Video Portrait Generation. In Computer Vision–ECCV 2022: 17th European Conference, Tel Aviv, Israel, October 23–27, 2022, Proceedings, Part XXXVII . Springer, 106–125
Xian Liu, Yinghao Xu, Qianyi Wu, Hang Zhou, Wayne Wu, and Bolei Zhou. 2022 · 2022
Cited alongside, same era.
Helisa Dhamo, Yinyu Nie, Arthur Moreau, Jifei Song, Richard Shaw, Yiren Zhou, and Eduardo Pérez-Pellitero. 2023 · 2023
Later among the works it cites.
Gaussianeditor: Editing 3d gaussians delicately with text instructions
Jiemin Fang, Junjie Wang, Xiaopeng Zhang, Lingxi Xie, and Qi Tian. 2023 · 2023
Later among the works it cites.
K-Planes: Explicit Radiance Fields in Space, Time, and Appearance
Sara Fridovich-Keil, Giacomo Meanti, Frederik Warburg, Benjamin Recht, and Angjoo Kanazawa. 2023 · 2023
Later among the works it cites.
3d gaussian splatting for real-time radiance field rendering
Bernhard Kerbl, Georgios Kopanas, Thomas Leimkühler, and George Drettakis. 2023 · 2023
Later among the works it cites.
Efficient Region-Aware Neural Radiance Fields for High-Fidelity Talking Portrait Synthesis
Jiahe Li, Jiawei Zhang, Xiao Bai, Jun Zhou, and Lin Gu. 2023 · 2023
Later among the works it cites.
Animatable 3D Gaussian: Fast and High-Quality Reconstruction of Multiple Human Avatars
Yang Liu, Xiang Huang, Minghan Qin, Qinwei Lin, and Haoqian Wang. 2023 · 2023
Later among the works it cites.
Dynamic 3D Gaussians: Tracking by Persistent Dynamic View Synthesis
Jonathon Luiten, Georgios Kopanas, Bastian Leibe, and Deva Ramanan. 2023 · 2023
Later among the works it cites.
Gaussianavatars: Photorealistic head avatars with rigged 3d gaussians
Shenhan Qian, Tobias Kirschstein, Liam Schoneveld, Davide Davoli, Simon Giebenhain, and Matthias Nießner. 2023 · 2023
Later among the works it cites.
4D Gaussian Splatting for Real-Time Dynamic Scene Rendering
Guanjun Wu, Taoran Yi, Jiemin Fang, Lingxi Xie, Xiaopeng Zhang, Wei Wei, Wenyu Liu, Qi Tian, and Xinggang Wang. 2023 · 2023
Later among the works it cites.
GeneFace++: Generalized and Stable Real-Time Audio-Driven 3D Talking Face Generation
Zhenhui Ye, Jinzheng He, Ziyue Jiang, Rongjie Huang, Jiawei Huang, Jinglin Liu, Yi Ren, Xiang Yin, Zejun Ma, and Zhou Zhao. 2023 · 2023
Later among the works it cites.
SadTalker: Learning Realistic 3D Motion Coefficients for Stylized Audio-Driven Single Image Talking Face Animation. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 8652–8661
Wenxuan Zhang, Xiaodong Cun, Xuan Wang, Yong Zhang, Xi Shen, Yu Guo, Ying Shan, and Fei Wang. 2023 · 2023
Later among the works it cites.
Mesh-based Gaussian Splatting for Real-time Large-scale Deformation
Lin Gao, Jie Yang, Bo-Tao Zhang, Jia-Mu Sun, Yu-Jie Yuan, Hongbo Fu, and Yu-Kun Lai. 2024 · 2024
Closest in time.
GauHuman: Articulated Gaussian Splatting from Monocular Human Videos. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)
Shoukang Hu and Ziwei Liu. 2024 · 2024
Closest in time.
Animatable Gaussians: Learning Pose-dependent Gaussian Maps for High-fidelity Human Avatar Modeling. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)
Zhe Li, Zerong Zheng, Lizhen Wang, and Yebin Liu. 2024 · 2024
Closest in time.
GaussianHead: High-fidelity Head Avatars with Learnable Gaussian Derivation
Jie Wang, Jiu-Cheng Xie, Xianyan Li, Feng Xu, Chi-Man Pun, and Hao Gao. 2024 · 2024
Closest in time.