Fetching the paper…
Reading the bibliography…
Creating expressive, diverse and high-quality 3D avatars from highly customized text descriptions and pose guidance is a challenging task, due to the intricacy of modeling and texturing in 3D that ensure details and various styles (realistic, fictional, etc).
NeRF: Representing Scenes as Neural Radiance Fields for View Synthesis
Mildenhall, B.; Srinivasan, P. P.; Tancik, M.; Barron, J. T.; Ramamoorthi, R.; and Ng, R. 2020 · 2003
Earlier work this paper cites.
Liu, L.; Gu, J.; Lin, K. Z.; Chua, T.-S.; and Theobalt, C. 2020 · 2007
Earlier work this paper cites.
SMPL: a skinned multi-person linear model
Loper, M.; Mahmood, N.; Romero, J.; Pons-Moll, G.; and Black, M. J. 2015 · 2015
Earlier work this paper cites.
Keep It SMPL: Automatic Estimation of 3D Human Pose and Shape from a Single Image
Bogo, F.; Kanazawa, A.; Lassner, C.; Gehler, P.; Romero, J.; and Black, M. J. 2016 · 2016
Earlier work this paper cites.
DeepFashion: Powering Robust Clothes Recognition and Retrieval with Rich Annotations
Liu, Z.; Luo, P.; Qiu, S.; Wang, X.; and Tang, X. 2016 · 2016
Earlier work this paper cites.
DensePose: Dense Human Pose Estimation in the Wild
Güler, R. A.; Neverova, N.; and Kokkinos, I. 2018 · 2018
Earlier work this paper cites.
Analyzing and Improving the Image Quality of StyleGAN
Karras, T.; Laine, S.; Aittala, M.; Hellsten, J.; Lehtinen, J.; and Aila, T. 2019 · 2020
Earlier work this paper cites.
Learning Transferable Visual Models From Natural Language Supervision
Radford, A.; Kim, J. W.; Hallacy, C.; Ramesh, A.; Goh, G.; Agarwal, S.; Sastry, G.; Askell, A.; Mishkin, P.; Clark, J.; Krueger, G.; and Sutskever, I. 2021 · 2021
Earlier work this paper cites.
Deep Marching Tetrahedra: a Hybrid Representation for High-Resolution 3D Shape Synthesis
Shen, T.; Gao, J.; Yin, K.; Liu, M.-Y.; and Fidler, S. 2021 · 2021
Earlier work this paper cites.
AvatarCLIP: Zero-Shot Text-Driven Generation and Animation of 3D Avatars
Hong, F.; Zhang, M.; Pan, L.; Cai, Z.; Yang, L.; and Liu, Z. 2022 · 2022
Earlier work this paper cites.
Zero-Shot Text-Guided Object Generation with Dream Fields
Jain, A.; Mildenhall, B.; Barron, J. T.; Abbeel, P.; and Poole, B. 2021 · 2022
Earlier work this paper cites.
InstantAvatar: Learning Avatars from Monocular Video in 60 Seconds
Jiang, T.; Chen, X.; Song, J.; and Hilliges, O. 2022 · 2022
Cited alongside, same era.
CLIP-Mesh: Generating textured meshes from text using pretrained image-text models
Khalid, N. M.; Xie, T.; Belilovsky, E.; and Popa, T. 2022 · 2022
Cited alongside, same era.
Magic3D: High-Resolution Text-to-3D Content Creation
Lin, C.-H.; Gao, J.; Tang, L.; Takikawa, T.; Zeng, X.; Huang, X.; Kreis, K.; Fidler, S.; Liu, M.-Y.; and Lin, T.-Y. 2022 · 2022
Cited alongside, same era.
Latent-NeRF for Shape-Guided Generation of 3D Shapes and Textures
Metzer, G.; Richardson, E.; Patashnik, O.; Giryes, R.; and Cohen-Or, D. 2022 · 2022
Cited alongside, same era.
DreamAvatar: Text-and-Shape Guided 3D Human Avatar Generation via Diffusion Models
Cao, Y.; Cao, Y.-P.; Han, K.; Shan, Y.; and Wong, K.-Y. K. 2023 · 2023
Closest in time.
HeadSculpt: Crafting 3D Head Avatars with Text
Han, X.; Cao, Y.; Han, K.; Zhu, X.; Deng, J.; Song, Y.-Z.; Xiang, T.; and Wong, K.-Y. K. 2023 · 2023
Closest in time.
DreamWaltz: Make a Scene with Complex 3D Animatable Avatars
Huang, Y.; Wang, J.; Zeng, A.; Cao, H.; Qi, X.; Shi, Y.; Zha, Z.; and Zhang, L. 2023 · 2023
Closest in time.
HumanRF: High-Fidelity Neural Radiance Fields for Humans in Motion
Isik, M.; Rünz, M.; Georgopoulos, M.; Khakhulin, T.; Starck, J.; de Agapito, L.; and Nießner, M. 2023 · 2023
Closest in time.
AvatarCraft: Transforming Text into Neural Human Avatars with Parameterized Shape and Pose Control
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Poole, B.; Jain, A.; Barron, J. T.; and Mildenhall, B. 2022 · 2022
Cited alongside, same era.
High-Resolution Image Synthesis with Latent Diffusion Models
Rombach, R.; Blattmann, A.; Lorenz, D.; Esser, P.; and Ommer, B. 2021 · 2022
Cited alongside, same era.
Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding
Saharia, C.; Chan, W.; Saxena, S.; Li, L.; Whang, J.; Denton, E. L.; Ghasemipour, S. K. S.; Ayan, B. K.; Mahdavi, S. S.; Lopes, R. G.; Salimans, T.; Ho, J.; Fleet, D. J.; and Norouzi, M. 2022 · 2022
Cited alongside, same era.
CLIP-Forge: Towards Zero-Shot Text-to-Shape Generation
Sanghi, A.; Chu, H.; Lambourne, J.; Wang, Y.; Cheng, C.-Y.; and Fumero, M. 2021 · 2022
Cited alongside, same era.
Direct Voxel Grid Optimization: Super-fast Convergence for Radiance Fields Reconstruction
Sun, C.; Sun, M.; and Chen, H.-T. 2021 · 2022
Cited alongside, same era.
Voxurf: Voxel-based Efficient and Accurate Neural Surface Reconstruction
Wu, T.; Wang, J.; Pan, X.; Xu, X.; Theobalt, C.; Liu, Z.; and Lin, D. 2022 · 2022
Cited alongside, same era.
ECON: Explicit Clothed humans Obtained from Normals
Xiu, Y.; Yang, J.; Cao, X.; Tzionas, D.; and Black, M. J. 2022 · 2022
Cited alongside, same era.
BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models
Li, J.; Li, D.; Savarese, S.; and Hoi, S. 2023a
Cited in the paper.
Jiang, R.; Wang, C.; Zhang, J.; Chai, M.; He, M.; Chen, D.; and Liao, J. 2023 · 2023
Closest in time.
DreamHuman: Animatable 3D Avatars from Text
Kolotouros, N.; Alldieck, T.; Zanfir, A.; Bazavan, E. G.; Fieraru, M.; and Sminchisescu, C. 2023 · 2023
Closest in time.
TEXTure: Text-Guided Texturing of 3D Shapes
Richardson, E.; Metzer, G.; Alaluf, Y.; Giryes, R.; and Cohen-Or, D. 2023 · 2023
Closest in time.
Adding Conditional Control to Text-to-Image Diffusion Models
Zhang, L.; and Agrawala, M. 2023 · 2023
Closest in time.
DreamFace: Progressive Generation of Animatable 3D Faces under Text Guidance
Zhang, L.; Qiu, Q.; Lin, H.; Zhang, Q.; Shi, C.; Yang, W.; Shi, Y.; Yang, S.; Xu, L.; and Yu, J. 2023 · 2023
Closest in time.
AvatarReX: Real-time Expressive Full-body Avatars
Zheng, Z.; Zhao, X.; Zhang, H.; Liu, B.; and Liu, Y. 2023 · 2023
Closest in time.
Cross-Domain and Disentangled Face Manipulation With 3D Guidance
Wang, C.; Chai, M.; He, M.; Chen, D.; and Liao, J. 2021 · 2066
Closest in time.