Fetching the paper…
Reading the bibliography…
Text-to-3D generation has attracted much attention from the computer vision community.
Brown T, Mann B, Ryder N, Subbiah M, Kaplan JD, Dhariwal P, Neelakantan A, Shyam P, Sastry G, Askell A, et al. (2020) Language models are few-shot learners. Advances in neural information processing systems 33:1877–1901
1901
Earlier work this paper cites.
Song Y, Sohl-Dickstein J, Kingma DP, Kumar A, Ermon S, Poole B (2021) Score-based generative modeling through stochastic differential equations. 2011.13456
2011
Earlier work this paper cites.
Kingma DP, Ba J (2014) Adam: A method for stochastic optimization. arXiv preprint arXiv:14126980
2014
Earlier work this paper cites.
Arjovsky M, Bottou L (2016) Towards principled methods for training generative adversarial networks. In: International Conference on Learning Representations
2016
Earlier work this paper cites.
Reed S, Akata Z, Yan X, Logeswaran L, Schiele B, Lee H (2016) Generative adversarial text to image synthesis. In: International conference on machine learning, PMLR, pp 1060–1069
2016
Earlier work this paper cites.
Vaswani A, Shazeer N, Parmar N, Uszkoreit J, Jones L, Gomez AN, Kaiser Ł, Polosukhin I (2017) Attention is all you need. Advances in neural information processing systems 30
2017
Earlier work this paper cites.
Xu T, Zhang P, Huang Q, Zhang H, Gan Z, Huang X, He X (2018) Attngan: Fine-grained text to image generation with attentional generative adversarial networks. In: Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
2018
Earlier work this paper cites.
Karras T, Laine S, Aila T (2019) A style-based generator architecture for generative adversarial networks. In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition, pp 4401–4410
2019
Earlier work this paper cites.
Qiao T, Zhang J, Xu D, Tao D (2019) Mirrorgan: Learning text-to-image generation by redescription. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)
2019
Earlier work this paper cites.
Tan H, Liu X, Li X, Zhang Y, Yin B (2019) Semantics-enhanced adversarial nets for text-to-image synthesis. In: Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV)
2019
Earlier work this paper cites.
Goodfellow I, Pouget-Abadie J, Mirza M, Xu B, Warde-Farley D, Ozair S, Courville A, Bengio Y (2020) Generative adversarial networks. Communications of the ACM 63(11):139–144
2020
Earlier work this paper cites.
Ho J, Jain A, Abbeel P (2020) Denoising diffusion probabilistic models. Advances in Neural Information Processing Systems 33:6840–6851
2020
Earlier work this paper cites.
Karras T, Laine S, Aittala M, Hellsten J, Lehtinen J, Aila T (2020) Analyzing and improving the image quality of StyleGAN. In: CVPR
2020
Earlier work this paper cites.
Mildenhall B, Srinivasan PP, Tancik M, Barron JT, Ramamoorthi R, Ng R (2020) Nerf: Representing scenes as neural radiance fields for view synthesis. In: ECCV
2020
Earlier work this paper cites.
Song J, Meng C, Ermon S (2020) Denoising diffusion implicit models. arXiv preprint arXiv:201002502
2020
Earlier work this paper cites.
Ding M, Yang Z, Hong W, Zheng W, Zhou C, Yin D, Lin J, Zou X, Shao Z, Yang H, et al. (2021) Cogview: Mastering text-to-image generation via transformers. Advances in Neural Information Processing Systems 34:19822–19835
2021
Earlier work this paper cites.
Esser P, Rombach R, Ommer B (2021) Taming transformers for high-resolution image synthesis. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pp 12873–12883
2021
Earlier work this paper cites.
Karras T, Aittala M, Laine S, Härkönen E, Hellsten J, Lehtinen J, Aila T (2021) Alias-free generative adversarial networks. In: NeurIPS
2021
Cited alongside, same era.
Nichol A, Dhariwal P, Ramesh A, Shyam P, Mishkin P, McGrew B, Sutskever I, Chen M (2021) Glide: Towards photorealistic image generation and editing with text-guided diffusion models. arXiv preprint arXiv:211210741
2021
Cited alongside, same era.
Radford A, Kim JW, Hallacy C, Ramesh A, Goh G, Agarwal S, Sastry G, Askell A, Mishkin P, Clark J, et al. (2021) Learning transferable visual models from natural language supervision. In: International Conference on Machine Learning, PMLR, pp 8748–8763
2021
Cited alongside, same era.
Ramesh A, Pavlov M, Goh G, Gray S, Voss C, Radford A, Chen M, Sutskever I (2021) Zero-shot text-to-image generation. In: International Conference on Machine Learning, PMLR, pp 8821–8831
2021
Cited alongside, same era.
Ramesh A, Dhariwal P, Nichol A, Chu C, Chen M (2022) Hierarchical text-conditional image generation with clip latents. arXiv preprint arXiv:220406125
2022
Later among the works it cites.
Rombach R, Blattmann A, Lorenz D, Esser P, Ommer B (2022) High-resolution image synthesis with latent diffusion models. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp 10684–10695
2022
Later among the works it cites.
Saharia C, Chan W, Saxena S, Li L, Whang J, Denton E, Ghasemipour SKS, Ayan BK, Mahdavi SS, Lopes RG, et al. (2022) Photorealistic text-to-image diffusion models with deep language understanding. arXiv preprint arXiv:220511487
2022
Later among the works it cites.
Sanghi A, Chu H, Lambourne JG, Wang Y, Cheng CY, Fumero M, Malekshan KR (2022) Clip-forge: Towards zero-shot text-to-shape generation. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pp 18603–18613
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ruan S, Zhang Y, Zhang K, Fan Y, Tang F, Liu Q, Chen E (2021) Dae-gan: Dynamic aspect-aware gan for text-to-image synthesis. In: Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV), pp 13960–13969
2021
Cited alongside, same era.
Schuhmann C, Vencu R, Beaumont R, Kaczmarczyk R, Mullis C, Katta A, Coombes T, Jitsev J, Komatsuzaki A (2021) Laion-400m: Open dataset of clip-filtered 400 million image-text pairs. arXiv preprint arXiv:211102114
2021
Cited alongside, same era.
Shen T, Gao J, Yin K, Liu MY, Fidler S (2021) Deep marching tetrahedra: a hybrid representation for high-resolution 3d shape synthesis. Advances in Neural Information Processing Systems 34:6087–6101
2021
Cited alongside, same era.
Chan ER, Lin CZ, Chan MA, Nagano K, Pan B, De Mello S, Gallo O, Guibas LJ, Tremblay J, Khamis S, et al. (2022) Efficient geometry-aware 3d generative adversarial networks. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp 16123–16133
2022
Cited alongside, same era.
Ho J, Salimans T (2022) Classifier-free diffusion guidance. arXiv preprint arXiv:220712598
2022
Cited alongside, same era.
Jain A, Mildenhall B, Barron JT, Abbeel P, Poole B (2022) Zero-shot text-guided object generation with dream fields. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pp 867–876
2022
Cited alongside, same era.
Khalid NM, Xie T, Belilovsky E, Tiberiu P (2022) Clip-mesh: Generating textured meshes from text using pretrained image-text models. SIGGRAPH Asia 2022 Conference Papers
2022
Cited alongside, same era.
Lee HH, Chang AX (2022) Understanding pure clip guidance for voxel grid nerf models. arXiv preprint arXiv:220915172
2022
Cited alongside, same era.
Schuhmann C, Beaumont R, Vencu R, Gordon C, Wightman R, Cherti M, Coombes T, Katta A, Mullis C, Wortsman M, et al. (2022) Laion-5b: An open large-scale dataset for training next generation image-text models. arXiv preprint arXiv:221008402
2022
Later among the works it cites.
Sun C, Sun M, Chen HT (2022) Direct voxel grid optimization: Super-fast convergence for radiance fields reconstruction. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pp 5459–5469
2022
Later among the works it cites.
2022
Later among the works it cites.
Armandpour M, Zheng H, Sadeghian A, Sadeghian A, Zhou M (2023) Re-imagine the negative prompt algorithm: Transform 2d diffusion into 3d, alleviate janus problem and beyond. arXiv preprint arXiv:230404968
2023
Closest in time.
Deitke M, Schwenk D, Salvador J, Weihs L, Michel O, VanderBilt E, Schmidt L, Ehsani K, Kembhavi A, Farhadi A (2023) Objaverse: A universe of annotated 3d objects. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp 13142–13153
2023
Closest in time.
Long X, Guo YC, Lin C, Liu Y, Dou Z, Liu L, Ma Y, Zhang SH, Habermann M, Theobalt C, et al. (2023) Wonder3d: Single image to 3d using cross-domain diffusion. arXiv preprint arXiv:231015008
2023
Closest in time.
Lorraine J, Xie K, Zeng X, Lin CH, Takikawa T, Sharp N, Lin TY, Liu MY, Fidler S, Lucas J (2023) Att3d: Amortized text-to-3d object synthesis. arXiv preprint arXiv:230607349
2023
Closest in time.
Metzer G, Richardson E, Patashnik O, Giryes R, Cohen-Or D (2023) Latent-nerf for shape-guided generation of 3d shapes and textures. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
2023
Closest in time.
Shi Y, Wang P, Ye J, Long M, Li K, Yang X (2023) Mvdream: Multi-view diffusion for 3d generation. arXiv preprint arXiv:230816512
2023
Closest in time.
2023
Closest in time.
Tsalicoglou C, Manhardt F, Tonioni A, Niemeyer M, Tombari F (2023) Textmesh: Generation of realistic 3d meshes from text prompts. arXiv preprint arXiv:230412439
2023
Closest in time.
Yi H, Zheng Z, Xu X, Chua Ts (2023) Progressive text-to-3d generation for automatic 3d prototyping. arXiv preprint arXiv:230914600
2023
Closest in time.