Fetching the paper…
Reading the bibliography…
Existing works in single-image human reconstruction suffer from weak generalizability due to insufficient training data or 3D inconsistencies for a lack of comprehensive multi-view knowledge.
Detailed human shape and pose from images
Balan, A. O.; Sigal, L.; Black, M. J.; Davis, J. E.; and Haussecker, H. W. 2007 · 2007
Earlier work this paper cites.
Dynamic shape capture using multi-view photometric stereo
Vlasic, D.; Peers, P.; Baran, I.; Debevec, P.; Popović, J.; Rusinkiewicz, S.; and Matusik, W. 2009 · 2009
Earlier work this paper cites.
Denoising diffusion implicit models
Song, J.; Meng, C.; and Ermon, S. 2020 · 2010
Earlier work this paper cites.
SMPL: A Skinned Multi-Person Linear Model
Loper, M.; Mahmood, N.; Romero, J.; Pons-Moll, G.; and Black, M. J. 2015 · 2015
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation
Ronneberger, O.; Fischer, P.; and Brox, T. 2015 · 2015
Earlier work this paper cites.
Deep unsupervised learning using nonequilibrium thermodynamics
Sohl-Dickstein, J.; Weiss, E.; Maheswaranathan, N.; and Ganguli, S. 2015 · 2015
Earlier work this paper cites.
View synthesis by appearance flow
Zhou, T.; Tulsiani, S.; Sun, W.; Malik, J.; and Efros, A. A. 2016 · 2016
Earlier work this paper cites.
Pose guided person image generation
Ma, L.; Jia, X.; Sun, Q.; Schiele, B.; Tuytelaars, T.; and Van Gool, L. 2017 · 2017
Earlier work this paper cites.
A variational u-net for conditional appearance and shape generation
Esser, P.; Sutter, E.; and Ommer, B. 2018 · 2018
Earlier work this paper cites.
Deformable gans for pose-based human image generation
Siarohin, A.; Sangineto, E.; Lathuiliere, S.; and Sebe, N. 2018 · 2018
Earlier work this paper cites.
Multi-garment net: Learning to dress 3d people from images
Bhatnagar, B. L.; Tiwari, G.; Theobalt, C.; and Pons-Moll, G. 2019 · 2019
Earlier work this paper cites.
Dense intrinsic appearance flow for human pose transfer
Li, Y.; Huang, C.; and Loy, C. C. 2019 · 2019
Earlier work this paper cites.
Liquid warping gan: A unified framework for human motion imitation, appearance transfer and novel view synthesis
Liu, W.; Piao, Z.; Min, J.; Luo, W.; Ma, L.; and Gao, S. 2019 · 2019
Earlier work this paper cites.
Expressive body capture: 3d hands, face, and body from a single image
Pavlakos, G.; Choutas, V.; Ghorbani, N.; Bolkart, T.; Osman, A. A.; Tzionas, D.; and Black, M. J. 2019 · 2019
Earlier work this paper cites.
Pifu: Pixel-aligned implicit function for high-resolution clothed human digitization
Saito, S.; Huang, Z.; Natsume, R.; Morishima, S.; Kanazawa, A.; and Li, H. 2019 · 2019
Earlier work this paper cites.
Progressive pose attention transfer for person image generation
Zhu, Z.; Huang, T.; Shi, B.; Yu, M.; Wang, B.; and Bai, X. 2019 · 2019
Earlier work this paper cites.
Denoising diffusion probabilistic models
Ho, J.; Jain, A.; and Abbeel, P. 2020 · 2020
Earlier work this paper cites.
Controllable person image synthesis with attribute-decomposed gan
Men, Y.; Mao, Y.; Jiang, Y.; Ma, W.-Y.; and Lian, Z. 2020 · 2020
Earlier work this paper cites.
Deep image spatial transformation for person image generation
Ren, Y.; Yu, X.; Chen, J.; Li, T. H.; and Li, G. 2020 · 2020
Earlier work this paper cites.
Diffusion models beat gans on image synthesis
Dhariwal, P.; and Nichol, A. 2021 · 2021
Earlier work this paper cites.
Pathdreamer: A world model for indoor navigation
Koh, J. Y.; Lee, H.; Yang, Y.; Baldridge, J.; and Anderson, P. 2021 · 2021
Cited alongside, same era.
Infinite nature: Perpetual view generation of natural scenes from a single image
Liu, A.; Tucker, R.; Jampani, V.; Makadia, A.; Snavely, N.; and Kanazawa, A. 2021 · 2021
Cited alongside, same era.
Learning semantic person image generation by region-adaptive normalization
Lv, Z.; Li, X.; Li, X.; Li, F.; Lin, T.; He, D.; and Zuo, W. 2021 · 2021
Cited alongside, same era.
Glide: Towards photorealistic image generation and editing with text-guided diffusion models
Nichol, A.; Dhariwal, P.; Ramesh, A.; Shyam, P.; Mishkin, P.; McGrew, B.; Sutskever, I.; and Chen, M. 2021 · 2021
Cited alongside, same era.
Pamir: Parametric model-conditioned implicit representation for image-based human reconstruction
Zheng, Z.; Yu, T.; Liu, Y.; and Dai, Q. 2021 · 2021
Cited alongside, same era.
Learning locally editable virtual humans
Ho, H.-I.; Xue, L.; Song, J.; and Hilliges, O. 2023 · 2023
Later among the works it cites.
Dreampose: Fashion image-to-video synthesis via stable diffusion
Karras, J.; Holynski, A.; Wang, T.-C.; and Kemelmacher-Shlizerman, I. 2023 · 2023
Later among the works it cites.
Sdxl: Improving latent diffusion models for high-resolution image synthesis
Podell, D.; English, Z.; Lacey, K.; Blattmann, A.; Dockhorn, T.; Müller, J.; Penna, J.; and Rombach, R. 2023 · 2023
Later among the works it cites.
Mvdream: Multi-view diffusion for 3d generation
Shi, Y.; Wang, P.; Ye, J.; Long, M.; Li, K.; and Yang, X. 2023 · 2023
Later among the works it cites.
Dreamgaussian: Generative gaussian splatting for efficient 3d content creation
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Stylegan-human: A data-centric odyssey of human generation
Fu, J.; Li, S.; Jiang, Y.; Lin, K.-Y.; Qian, C.; Loy, C. C.; Wu, W.; and Liu, Z. 2022 · 2022
Cited alongside, same era.
Classifier-free diffusion guidance
Ho, J.; and Salimans, T. 2022 · 2022
Cited alongside, same era.
Tutel: Adaptive Mixture-of-Experts at Scale
Hwang, C.; Cui, W.; Xiong, Y.; Yang, Z.; Liu, Z.; Hu, H.; Wang, Z.; Salas, R.; Jose, J.; Ram, P.; Chau, J.; Cheng, P.; Yang, F.; Yang, M.; and Xiong, Y. 2022 · 2022
Cited alongside, same era.
Viewformer: Nerf-free neural rendering from few images using transformers
Kulhánek, J.; Derner, E.; Sattler, T.; and Babuška, R. 2022 · 2022
Cited alongside, same era.
Infinitenature-zero: Learning perpetual view generation of natural scenes from single images
Li, Z.; Wang, Q.; Snavely, N.; and Kanazawa, A. 2022 · 2022
Cited alongside, same era.
Dreamfusion: Text-to-3d using 2d diffusion
Poole, B.; Jain, A.; Barron, J. T.; and Mildenhall, B. 2022 · 2022
Cited alongside, same era.
Hierarchical text-conditional image generation with clip latents
Ramesh, A.; Dhariwal, P.; Nichol, A.; Chu, C.; and Chen, M. 2022 · 2022
Cited alongside, same era.
Tang, J.; Ren, J.; Zhou, H.; Liu, Z.; and Zeng, G. 2023 · 2023
Later among the works it cites.
Consistent View Synthesis with Pose-Guided Diffusion Models
Tseng, H.-Y.; Li, Q.; Kim, C.; Alsisan, S.; Huang, J.-B.; and Kopf, J. 2023 · 2023
Later among the works it cites.
Econ: Explicit clothed humans optimized via normal integration
Xiu, Y.; Yang, J.; Cao, X.; Tzionas, D.; and Black, M. J. 2023 · 2023
Later among the works it cites.
PyMAF-X: Towards Well-aligned Full-body Model Regression from Monocular Images
Zhang, H.; Tian, Y.; Zhang, Y.; Li, M.; An, L.; Sun, Z.; and Liu, Y. 2023 · 2023
Later among the works it cites.
Scaling rectified flow transformers for high-resolution image synthesis
Esser, P.; Kulal, S.; Blattmann, A.; Entezari, R.; Müller, J.; Saini, H.; Levi, Y.; Lorenz, D.; Sauer, A.; Boesel, F.; et al. 2024 · 2024
Closest in time.
Contex-human: Free-view rendering of human from a single image with texture-consistent synthesis
Gao, X.; Li, X.; Zhang, C.; Zhang, Q.; Cao, Y.; Shan, Y.; and Quan, L. 2024 · 2024
Closest in time.
Animate anyone: Consistent and controllable image-to-video synthesis for character animation
Hu, L. 2024 · 2024
Closest in time.
Tech: Text-guided reconstruction of lifelike clothed humans
Huang, Y.; Yi, H.; Xiu, Y.; Liao, T.; Tang, J.; Cai, D.; and Thies, J. 2024 · 2024
Closest in time.
Repurposing Diffusion-Based Image Generators for Monocular Depth Estimation
Ke, B.; Obukhov, A.; Huang, S.; Metzger, N.; Daudt, R. C.; and Schindler, K. 2024 · 2024
Closest in time.
Humangaussian: Text-driven 3d human generation with gaussian splatting
Liu, X.; Zhan, X.; Tang, J.; Shan, Y.; Zeng, G.; Lin, D.; Liu, X.; and Liu, Z. 2024 · 2024
Closest in time.
Wonder3d: Single image to 3d using cross-domain diffusion
Long, X.; Guo, Y.-C.; Lin, C.; Liu, Y.; Dou, Z.; Liu, L.; Ma, Y.; Zhang, S.-H.; Habermann, M.; Theobalt, C.; et al. 2024 · 2024
Closest in time.
Sv3d: Novel multi-view synthesis and 3d generation from a single image using latent video diffusion
Voleti, V.; Yao, C.-H.; Boss, M.; Letts, A.; Pankratz, D.; Tochilkin, D.; Laforte, C.; Rombach, R.; and Jampani, V. 2024 · 2024
Closest in time.
Magicanimate: Temporally consistent human image animation using diffusion model
Xu, Z.; Zhang, J.; Liew, J. H.; Yan, H.; Liu, J.-W.; Zhang, C.; Feng, J.; and Shou, M. Z. 2024 · 2024
Closest in time.
Humanref: Single image to 3d human generation via reference-guided diffusion
Zhang, J.; Li, X.; Zhang, Q.; Cao, Y.; Shan, Y.; and Liao, J. 2024 · 2024
Closest in time.
Champ: Controllable and consistent human image animation with 3d parametric guidance
Zhu, S.; Chen, J. L.; Dai, Z.; Xu, Y.; Cao, X.; Yao, Y.; Zhu, H.; and Zhu, S. 2024 · 2024
Closest in time.