Fetching the paper…
Reading the bibliography…
Large-scale text-guided image diffusion models have shown astonishing results in text-to-image (T2I) generation.
Zippered polygon meshes from range images
Turk, G.; and Levoy, M. 1994 · 1994
Earlier work this paper cites.
Shapenet: An information-rich 3d model repository
Chang, A. X.; Funkhouser, T.; Guibas, L.; Hanrahan, P.; Huang, Q.; Li, Z.; Savarese, S.; Savva, M.; Song, S.; Su, H.; et al. 2015 · 2015
Earlier work this paper cites.
Gans trained by a two time-scale update rule converge to a local nash equilibrium
Heusel, M.; Ramsauer, H.; Unterthiner, T.; Nessler, B.; and Hochreiter, S. 2017 · 2017
Earlier work this paper cites.
Bińkowski, M.; Sutherland, D. J.; Arbel, M.; and Gretton, A. 2018 · 2018
Earlier work this paper cites.
Texture fields: Learning texture representations in function space
Oechsle, M.; Mescheder, L.; Niemeyer, M.; Strauss, T.; and Geiger, A. 2019 · 2019
Earlier work this paper cites.
Denoising diffusion probabilistic models
Ho, J.; Jain, A.; and Abbeel, P. 2020 · 2020
Earlier work this paper cites.
Modular Primitives for High-Performance Differentiable Rendering
Laine, S.; Hellsten, J.; Karras, T.; Seol, Y.; Lehtinen, J.; and Aila, T. 2020 · 2020
Earlier work this paper cites.
Glide: Towards photorealistic image generation and editing with text-guided diffusion models
Nichol, A.; Dhariwal, P.; Ramesh, A.; Shyam, P.; Mishkin, P.; McGrew, B.; Sutskever, I.; and Chen, M. 2021 · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
Radford, A.; Kim, J. W.; Hallacy, C.; Ramesh, A.; Goh, G.; Agarwal, S.; Sastry, G.; Askell, A.; Mishkin, P.; Clark, J.; et al. 2021 · 2021
Earlier work this paper cites.
Learning texture generators for 3d shape collections from internet photo sets
Yu, R.; Dong, Y.; Peers, P.; and Tong, X. 2021 · 2021
Earlier work this paper cites.
Auv-net: Learning aligned uv maps for texture transfer and synthesis
Chen, Z.; Yin, K.; and Fidler, S. 2022 · 2022
Earlier work this paper cites.
Prompt-to-prompt image editing with cross attention control
Hertz, A.; Mokady, R.; Tenenbaum, J.; Aberman, K.; Pritch, Y.; and Cohen-Or, D. 2022 · 2022
Earlier work this paper cites.
Point-e: A system for generating 3d point clouds from complex prompts
Nichol, A.; Jun, H.; Dhariwal, P.; Mishkin, P.; and Chen, M. 2022 · 2022
Earlier work this paper cites.
Dreamfusion: Text-to-3d using 2d diffusion
Poole, B.; Jain, A.; Barron, J. T.; and Mildenhall, B. 2022 · 2022
Earlier work this paper cites.
Hierarchical text-conditional image generation with clip latents
Ramesh, A.; Dhariwal, P.; Nichol, A.; Chu, C.; and Chen, M. 2022 · 2022
Earlier work this paper cites.
High-resolution image synthesis with latent diffusion models
Rombach, R.; Blattmann, A.; Lorenz, D.; Esser, P.; and Ommer, B. 2022 · 2022
Earlier work this paper cites.
Photorealistic text-to-image diffusion models with deep language understanding
Saharia, C.; Chan, W.; Saxena, S.; Li, L.; Whang, J.; Denton, E. L.; Ghasemipour, K.; Gontijo Lopes, R.; Karagol Ayan, B.; Salimans, T.; et al. 2022 · 2022
Cited alongside, same era.
Laion-5b: An open large-scale dataset for training next generation image-text models
Schuhmann, C.; Beaumont, R.; Vencu, R.; Gordon, C.; Wightman, R.; Cherti, M.; Coombes, T.; Katta, A.; Mullis, C.; Wortsman, M.; et al. 2022 · 2022
Cited alongside, same era.
Texturify: Generating textures on 3d shape surfaces
Siddiqui, Y.; Thies, J.; Ma, F.; Shan, Q.; Nießner, M.; and Dai, A. 2022 · 2022
Cited alongside, same era.
TexFusion: Synthesizing 3D Textures with Text-Guided Image Diffusion Models
Cao, T.; Kreis, K.; Fidler, S.; Sharp, N.; and Yin, K. 2023 · 2023
Cited alongside, same era.
Objaverse: A universe of annotated 3d objects
Deitke, M.; Schwenk, D.; Salvador, J.; Weihs, L.; Michel, O.; VanderBilt, E.; Schmidt, L.; Ehsani, K.; Kembhavi, A.; and Farhadi, A. 2023 · 2023
Cited alongside, same era.
Score jacobian chaining: Lifting pretrained 2d diffusion models for 3d generation
Wang, H.; Du, X.; Li, J.; Yeh, R. A.; and Shakhnarovich, G. 2023 · 2023
Later among the works it cites.
Dmv3d: Denoising multi-view diffusion using 3d large reconstruction model
Xu, Y.; Tan, H.; Luan, F.; Bi, S.; Wang, P.; Li, J.; Shi, Z.; Sunkavalli, K.; Wetzstein, G.; Xu, Z.; et al. 2023 · 2023
Later among the works it cites.
Rerender A Video: Zero-Shot Text-Guided Video-to-Video Translation
Yang, S.; Zhou, Y.; Liu, Z.; ; and Loy, C. C. 2023 · 2023
Later among the works it cites.
Youwang, K.; Oh, T.-H.; and Pons-Moll, G. 2023 · 2023
Later among the works it cites.
Adding conditional control to text-to-image diffusion models
Zhang, L.; Rao, A.; and Agrawala, M. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Hong, Y.; Zhang, K.; Gu, J.; Bi, S.; Zhou, Y.; Liu, D.; Liu, F.; Sunkavalli, K.; Bui, T.; and Tan, H. 2023 · 2023
Cited alongside, same era.
Textfield3d: Towards enhancing open-vocabulary 3d generation with noisy text fields
Huang, T.; Zeng, Y.; Dong, B.; Xu, H.; Xu, S.; Lau, R. W.; and Zuo, W. 2023 · 2023
Cited alongside, same era.
Shap-e: Generating conditional 3d implicit functions
Jun, H.; and Nichol, A. 2023 · 2023
Cited alongside, same era.
Text2video-zero: Text-to-image diffusion models are zero-shot video generators
Khachatryan, L.; Movsisyan, A.; Tadevosyan, V.; Henschel, R.; Wang, Z.; Navasardyan, S.; and Shi, H. 2023 · 2023
Cited alongside, same era.
Magic3d: High-resolution text-to-3d content creation
Lin, C.-H.; Gao, J.; Tang, L.; Takikawa, T.; Zeng, X.; Huang, X.; Kreis, K.; Fidler, S.; Liu, M.-Y.; and Lin, T.-Y. 2023 · 2023
Cited alongside, same era.
Wonder3d: Single image to 3d using cross-domain diffusion
Long, X.; Guo, Y.-C.; Lin, C.; Liu, Y.; Dou, Z.; Liu, L.; Ma, Y.; Zhang, S.-H.; Habermann, M.; Theobalt, C.; et al. 2023 · 2023
Cited alongside, same era.
Latent-nerf for shape-guided generation of 3d shapes and textures
Metzer, G.; Richardson, E.; Patashnik, O.; Giryes, R.; and Cohen-Or, D. 2023 · 2023
Cited alongside, same era.
civitai — The Home of Open-Source Generative AI
civitai. 2024 · 2024
Closest in time.
GenesisTex: Adapting Image Denoising Diffusion to Texture Space
Gao, C.; Jiang, B.; Li, X.; Zhang, Y.; and Yu, Q. 2024 · 2024
Closest in time.
Huggingface
Huggingface. 2024 · 2024
Closest in time.
SyncTweedies: A General Generative Framework Based on Synchronized Diffusions
Kim, J.; Koo, J.; Yeo, K.; and Sung, M. 2024 · 2024
Closest in time.
Pick-a-pic: An open dataset of user preferences for text-to-image generation
Kirstain, Y.; Polyak, A.; Singer, U.; Matiana, S.; Penna, J.; and Levy, O. 2024 · 2024
Closest in time.
One-2-3-45: Any single image to 3d mesh in 45 seconds without per-shape optimization
Liu, M.; Xu, C.; Jin, H.; Chen, L.; Varma T, M.; Xu, Z.; and Su, H. 2024 · 2024
Closest in time.
Meshy — 3D AI Generator
Meshy. 2024 · 2024
Closest in time.
T2i-adapter: Learning adapters to dig out more controllable ability for text-to-image diffusion models
Mou, C.; Wang, X.; Xie, L.; Wu, Y.; Zhang, J.; Qi, Z.; and Shan, Y. 2024 · 2024
Closest in time.
Prolificdreamer: High-fidelity and diverse text-to-3d generation with variational score distillation
Wang, Z.; Lu, C.; Wang, Y.; Bao, F.; Li, C.; Su, H.; and Zhu, J. 2024 · 2024
Closest in time.
FRESCO: Spatial-Temporal Correspondence for Zero-Shot Video Translation
Yang, S.; Zhou, Y.; Liu, Z.; and Loy, C. C. 2024 · 2024
Closest in time.