Fetching the paper…
Reading the bibliography…
Recent advancements in human video synthesis have enabled the generation of high-quality videos through the application of stable diffusion models.
U-net: Convolutional networks for biomedical image segmentation
O. Ronneberger, P. Fischer, and T. Brox · 2015
Earlier work this paper cites.
Neural discrete representation learning
A. Van Den Oord, O. Vinyals, et al · 2017
Earlier work this paper cites.
Densepose: Dense human pose estimation in the wild
R. A. Güler, N. Neverova, and I. Kokkinos · 2018
Earlier work this paper cites.
Animating arbitrary objects via deep motion transfer
A. Siarohin, S. Lathuilière, S. Tulyakov, E. Ricci, and N. Sebe · 2019
Earlier work this paper cites.
First order motion model for image animation
A. Siarohin, S. Lathuilière, S. Tulyakov, E. Ricci, and N. Sebe · 2019
Earlier work this paper cites.
Dwnet: Dense warp-based network for pose-guided human video generation
P. Zablotskaia, A. Siarohin, B. Zhao, and L. Sigal · 2019
Earlier work this paper cites.
Boosting semantic human matting with coarse annotations
J. Liu, Y. Yao, W. Hou, M. Cui, X. Xie, C. Zhang, and X.-s. Hua · 2020
Earlier work this paper cites.
Learning high fidelity depths of dressed humans by watching social media dance videos
Y. Jafarian and H. S. Park · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
A. Radford, J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, et al · 2021
Earlier work this paper cites.
Motion representations for articulated animation
A. Siarohin, O. J. Woodford, J. Ren, M. Chai, and S. Tulyakov · 2021
Earlier work this paper cites.
High-resolution image synthesis with latent diffusion models
R. Rombach, A. Blattmann, D. Lorenz, P. Esser, and B. Ommer · 2022
Earlier work this paper cites.
Exploring dual-task correlation for pose guided person image generation
P. Zhang, L. Yang, J.-H. Lai, and X. Xie · 2022
Cited alongside, same era.
Thin-plate spline motion model for image animation
J. Zhao and H. Zhang · 2022
Cited alongside, same era.
Anydoor: Zero-shot object-level image customization
X. Chen, L. Huang, Y. Liu, Y. Shen, D. Zhao, and H. Zhao · 2023
Cited alongside, same era.
Animatediff: Animate your personalized text-to-image diffusion models without specific tuning
Y. Guo, C. Yang, A. Rao, Y. Wang, Y. Qiao, D. Lin, and B. Dai · 2023
Cited alongside, same era.
Animate anyone: Consistent and controllable image-to-video synthesis for character animation
L. Hu, X. Gao, P. Zhang, K. Sun, B. Zhang, and L. Bo · 2023
Cited alongside, same era.
Magicanimate: Temporally consistent human image animation using diffusion model
Z. Xu, J. Zhang, J. H. Liew, H. Yan, J.-W. Liu, C. Zhang, J. Feng, and M. Z. Shou · 2023
Later among the works it cites.
Effective whole-body pose estimation with two-stages distillation
Z. Yang, A. Zeng, C. Yuan, and Y. Li · 2023
Later among the works it cites.
Bidirectionally deformable motion modulation for video-based human pose transfer
W.-Y. Yu, L.-M. Po, R. C. Cheung, Y. Zhao, Y. Xue, and K. Li · 2023
Later among the works it cites.
Magic clothing: Controllable garment-driven image synthesis
W. Chen, T. Gu, Y. Xu, and C. Chen · 2024
Closest in time.
Unsupervised semantic correspondence using stable diffusion
E. Hedlin, G. Sharma, S. Mahajan, H. Isack, A. Kar, A. Tagliasacchi, and K. M. Yi · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cotracker: It is better to track together
N. Karaev, I. Rocco, B. Graham, N. Neverova, A. Vedaldi, and C. Rupprecht · 2023
Cited alongside, same era.
Dreampose: Fashion image-to-video synthesis via stable diffusion
J. Karras, A. Holynski, T.-C. Wang, and I. Kemelmacher-Shlizerman · 2023
Cited alongside, same era.
Mtvg: Multi-text video generation with text-to-video models
G. Oh, J. Jeong, S. Kim, W. Byeon, J. Kim, S. Kim, H. Kwon, and S. Kim · 2023
Cited alongside, same era.
Dinov2: Learning robust visual features without supervision
M. Oquab, T. Darcet, T. Moutakanni, H. Vo, M. Szafraniec, V. Khalidov, P. Fernandez, D. Haziza, F. Massa, A. El-Nouby, et al · 2023
Cited alongside, same era.
Stable diffusion v1-5 model, 2023
runwayml · 2023
Cited alongside, same era.
Disco: Disentangled control for realistic human dance generation
T. Wang, L. Li, K. Lin, Y. Zhai, C.-C. Lin, Z. Yang, H. Zhang, Z. Liu, and L. Wang · 2023
Cited alongside, same era.
Diffusion hyperfeatures: Searching through time and space for semantic correspondence
G. Luo, L. Dunlap, D. H. Park, A. Holynski, and T. Darrell · 2024
Closest in time.
L. Tian, Q. Wang, B. Zhang, and L. Bo · 2024
Closest in time.
Aniportrait: Audio-driven synthesis of photorealistic portrait animation
H. Wei, Z. Yang, and Z. Wang · 2024
Closest in time.
Ootdiffusion: Outfitting fusion based latent diffusion for controllable virtual try-on
Y. Xu, T. Gu, W. Chen, and C. Chen · 2024
Closest in time.
Champ: Controllable and consistent human image animation with 3d parametric guidance
S. Zhu, J. L. Chen, Z. Dai, Y. Xu, X. Cao, Y. Yao, H. Zhu, and S. Zhu · 2024
Closest in time.