Fetching the paper…
Reading the bibliography…
We propose X-Portrait, an innovative conditional diffusion model tailored for generating expressive and temporally coherent portrait animation.
Perceptual losses for real-time style transfer and super-resolution. In Computer Vision–ECCV 2016: 14th European Conference, Amsterdam, The Netherlands, October 11-14, 2016, Proceedings, Part II 14 . Springer, 694–711
Justin Johnson, Alexandre Alahi, and Li Fei-Fei. 2016 · 2016
Earlier work this paper cites.
Warp-guided gans for single-photo facial animation
Jiahao Geng, Tianjia Shao, Youyi Zheng, Yanlin Weng, and Kun Zhou. 2018 · 2018
Earlier work this paper cites.
OpenPose: Realtime Multi-Person 2D Pose Estimation using Part Affinity Fields
Z. Cao, G. Hidalgo Martinez, T. Simon, S. Wei, and Y. A. Sheikh. 2019 · 2019
Earlier work this paper cites.
Arcface: Additive angular margin loss for deep face recognition. In CVPR . 4690–4699
Jiankang Deng, Jia Guo, Niannan Xue, and Stefanos Zafeiriou. 2019 · 2019
Earlier work this paper cites.
Disentangled and Controllable Face Image Generation via 3D Imitative-Contrastive Learning. In IEEE Computer Vision and Pattern Recognition
Yu Deng, Jiaolong Yang, Dong Chen, Fang Wen, and Xin Tong. 2020 · 2020
Earlier work this paper cites.
Denoising diffusion probabilistic models
Jonathan Ho, Ajay Jain, and Pieter Abbeel. 2020 · 2020
Earlier work this paper cites.
KonIQ-10k: An Ecologically Valid Database for Deep Learning of Blind Image Quality Assessment
Vlad Hosu, Hanhe Lin, Tamas Sziranyi, and Dietmar Saupe. 2020 · 2020
Earlier work this paper cites.
Denoising diffusion implicit models
Jiaming Song, Chenlin Meng, and Stefano Ermon. 2020a · 2020
Earlier work this paper cites.
Score-based generative modeling through stochastic differential equations
Yang Song, Jascha Sohl-Dickstein, Diederik P Kingma, Abhishek Kumar, Stefano Ermon, and Ben Poole. 2020b · 2020
Earlier work this paper cites.
Blindly Assess Image Quality in the Wild Guided by a Self-Adaptive Hyper Network. In IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)
Shaolin Su, Qingsen Yan, Yu Zhu, Cheng Zhang, Xin Ge, Jinqiu Sun, and Yanning Zhang. 2020 · 2020
Earlier work this paper cites.
Learning an Animatable Detailed 3D Face Model from In-The-Wild Images
Yao Feng, Haiwen Feng, Michael J. Black, and Timo Bolkart. 2021 · 2021
Earlier work this paper cites.
AD-NeRF: Audio Driven Neural Radiance Fields for Talking Head Synthesis. In IEEE/CVF International Conference on Computer Vision (ICCV)
Yudong Guo, Keyu Chen, Sen Liang, Yongjin Liu, Hujun Bao, and Juyong Zhang. 2021 · 2021
Earlier work this paper cites.
High-Resolution Image Synthesis with Latent Diffusion Models
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer. 2021 · 2021
Earlier work this paper cites.
Motion Representations for Articulated Animation. In CVPR
Aliaksandr Siarohin, Oliver Woodford, Jian Ren, Menglei Chai, and Sergey Tulyakov. 2021 · 2021
Earlier work this paper cites.
One-Shot Free-View Neural Talking-Head Synthesis for Video Conferencing. In CVPR
Ting-Chun Wang, Arun Mallya, and Ming-Yu Liu. 2021 · 2021
Earlier work this paper cites.
Stable diffusion v1.5 model card
Stability AI. 2022 · 2022
Cited alongside, same era.
MegaPortraits: One-Shot Megapixel Neural Head Avatars. In Proceedings of the 30th ACM International Conference on Multimedia
Nikita Drobyshev, Jenya Chelishev, Taras Khakhulin, Aleksei Ivakhnenko, Victor Lempitsky, and Egor Zakharov. 2022 · 2022
Cited alongside, same era.
Depth-Aware Generative Adversarial Network for Talking Head Video Generation
Fa-Ting Hong, Longhao Zhang, Li Shen, and Dan Xu. 2022 · 2022
Cited alongside, same era.
Realistic One-shot Mesh-based Head Avatars. In European Conference of Computer vision (ECCV)
Taras Khakhulin, Vanessa Sklyarova, Victor Lempitsky, and Egor Zakharov. 2022 · 2022
Cited alongside, same era.
High-resolution image synthesis with latent diffusion models. In CVPR . 10684–10695
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer. 2022 · 2022
Cited alongside, same era.
SparseCtrl: Adding Sparse Controls to Text-to-Video Diffusion Models
Yuwei Guo, Ceyuan Yang, Anyi Rao, Maneesh Agrawala, Dahua Lin, and Bo Dai. 2023a · 2023
Later among the works it cites.
AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning
Yuwei Guo, Ceyuan Yang, Anyi Rao, Yaohui Wang, Yu Qiao, Dahua Lin, and Bo Dai. 2023b · 2023
Later among the works it cites.
Implicit Identity Representation Conditioned Memory Compensation Network for Talking Head video Generation. In ICCV
Fa-Ting Hong and Dan Xu. 2023 · 2023
Later among the works it cites.
Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation
Li Hu, Xin Gao, Peng Zhang, Ke Sun, Bang Zhang, and Liefeng Bo. 2023 · 2023
Later among the works it cites.
Consistent123: One Image to Highly Consistent 3D Asset Using Case-Aware Diffusion Priors
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Photorealistic text-to-image diffusion models with deep language understanding
Chitwan Saharia, William Chan, Saurabh Saxena, Lala Li, Jay Whang, Emily L Denton, Kamyar Ghasemipour, Raphael Gontijo Lopes, Burcu Karagol Ayan, Tim Salimans, et al · 2022
Cited alongside, same era.
Masked Lip-Sync Prediction by Audio-Visual Contextual Exploitation in Transformers. In SIGGRAPH Asia 2022 Conference Papers
Yasheng Sun, Hang Zhou, Kaisiyuan Wang, Qianyi Wu, Zhibin Hong, Jingtuo Liu, Errui Ding, Jingdong Wang, Ziwei Liu, and Koike Hideki. 2022 · 2022
Cited alongside, same era.
EDGE: Editable Dance Generation From Music
Jonathan Tseng, Rodrigo Castellon, and C. Karen Liu. 2022 · 2022
Cited alongside, same era.
Latent Image Animator: Learning to Animate Images via Latent Space Navigation. In International Conference on Learning Representations
Yaohui Wang, Di Yang, Francois Bremond, and Antitza Dantcheva. 2022 · 2022
Cited alongside, same era.
Thin-Plate Spline Motion Model for Image Animation
Jian Zhao and Hui Zhang. 2022 · 2022
Cited alongside, same era.
ARFaceAnchor.BlendShapeLocation
Apple. 2023 · 2023
Cited alongside, same era.
Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets
Andreas Blattmann, Tim Dockhorn, Sumith Kulal, Daniel Mendelevitch, Maciej Kilian, Dominik Lorenz, Yam Levi, Zion English, Vikram Voleti, Adam Letts, Varun Jampani, and Robin Rombach. 2023 · 2023
Cited alongside, same era.
Yukang Lin, Haonan Han, Chaoqun Gong, Zunnan Xu, Yachao Zhang, and Xiu Li. 2023 · 2023
Later among the works it cites.
Latent consistency models: Synthesizing high-resolution images with few-step inference
Simian Luo, Yiqin Tan, Longbo Huang, Jian Li, and Hang Zhao. 2023 · 2023
Later among the works it cites.
ReenactArtFace: Artistic Face Image Reenactment
Linzi Qu, Jiaxiang Shang, Xiaoguang Han, and Hongbo Fu. 2023 · 2023
Later among the works it cites.
Next3D: Generative Neural Texture Rasterization for 3D-Aware Head Avatars. In CVPR
Jingxiang Sun, Xuan Wang, Lizhen Wang, Xiaoyu Li, Yong Zhang, Hongwen Zhang, and Yebin Liu. 2023 · 2023
Later among the works it cites.
OmniAvatar: Geometry-Guided Controllable 3D Head Synthesis. In 2023 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)
H. Xu, G. Song, Z. Jiang, J. Zhang, Y. Shi, J. Liu, W. Ma, J. Feng, and L. Luo. 2023a · 2023
Later among the works it cites.
Face Animation with an Attribute-Guided Diffusion Model. In 2023 IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW)
Bohan Zeng, Xuhui Liu, Sicheng Gao, Boyu Liu, Hong Li, Jianzhuang Liu, and Baochang Zhang. 2023 · 2023
Later among the works it cites.
[major update] reference-only control · Mikubill/SD-webui-controlnet · discussion #1236
Lyumin Zhang. 2023 · 2023
Later among the works it cites.
MagicPose: Realistic Human Poses and Facial Expressions Retargeting with Identity-aware Diffusion
Di Chang, Yichun Shi, Quankai Gao, Jessica Fu, Hongyi Xu, Guoxian Song, Qing Yan, Yizhe Zhu, Xiao Yang, and Mohammad Soleymani. 2024 · 2024
Closest in time.
deviantart
DeviantArt. 2024 · 2024
Closest in time.
midjourney
Midjourney. 2024 · 2024
Closest in time.