Fetching the paper…
Reading the bibliography…
This paper introduces WordArt Designer, a user-driven framework for artistic typography synthesis, relying on the Large Language Model (LLM).
Megatron-lm: Training multi-billion parameter language models using model parallelism
Mohammad Shoeybi, Mostofa Patwary, Raul Puri, et al. 2019 · 1909
Earlier work this paper cites.
FreeType 2
David Turner, Robert Wilhelm, and Werner Lemberg. 1996 · 1996
Earlier work this paper cites.
Video ecommerce: Towards online video advertising
Zhi-Qi Cheng, Yang Liu, Xiao Wu, and Xian-Sheng Hua. 2016 · 2016
Earlier work this paper cites.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Earlier work this paper cites.
Typography in destination advertising: An exploratory study and research perspectives
Jennifer Amar, Olivier Droulers, and Patrick Legohérel. 2017 · 2017
Earlier work this paper cites.
Images as a resource for supporting vocabulary learning: A multimodal analysis of thai efl tablet apps for primary school children
Sompatu Vungthong, Emilia Djonov, and Jane Torr. 2017 · 2017
Earlier work this paper cites.
Improving language understanding by generative pre-training
Alec Radford, Karthik Narasimhan, Tim Salimans, and Ilya Sutskever. 2018 · 2018
Earlier work this paper cites.
Personalized clothing recommendation combining user social circle and fashion style consistency
Guang-Lu Sun, Zhi-Qi Cheng, Xiao Wu, and Qiang Peng. 2018 · 2018
Earlier work this paper cites.
Multi-view image generation from a single-view
Bo Zhao, Xiao Wu, Zhi-Qi Cheng, et al. 2018 · 2018
Earlier work this paper cites.
BERT: pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Earlier work this paper cites.
Language models are few-shot learners
Tom B. Brown, Benjamin Mann, Nick Ryder, et al. 2020 · 2020
Earlier work this paper cites.
Denoising diffusion probabilistic models
Jonathan Ho, Ajay Jain, and Pieter Abbeel. 2020 · 2020
Earlier work this paper cites.
Generating person images with appearance-aware pose stylizer
Siyu Huang, Haoyi Xiong, Zhi-Qi Cheng, et al. 2020 · 2020
Earlier work this paper cites.
Differentiable vector graphics rasterization for editing and learning
Tzu-Mao Li, Michal Lukác, Michaël Gharbi, and Jonathan Ragan-Kelley. 2020 · 2020
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, et al. 2020 · 2020
Earlier work this paper cites.
Zero: memory optimizations toward training trillion parameter models
Samyam Rajbhandari, Jeff Rasley, Olatunji Ruwase, and Yuxiong He. 2020 · 2020
Cited alongside, same era.
Diffusion models beat gans on image synthesis
Prafulla Dhariwal and Alexander Quinn Nichol. 2021 · 2021
Cited alongside, same era.
Improved denoising diffusion probabilistic models
Alexander Quinn Nichol and Prafulla Dhariwal. 2021 · 2021
Cited alongside, same era.
Zero-shot text-to-image generation
Aditya Ramesh, Mikhail Pavlov, Gabriel Goh, et al. 2021 · 2021
Cited alongside, same era.
Score-based generative modeling through stochastic differential equations
Yang Song, Jascha Sohl-Dickstein, Diederik P. Kingma, et al. 2021 · 2021
Cited alongside, same era.
ediff-i: Text-to-image diffusion models with an ensemble of expert denoisers
An image is worth one word: Personalizing text-to-image generation using textual inversion
Rinon Gal, Yuval Alaluf, Yuval Atzmon, et al. 2023 · 2023
Closest in time.
Composer: Creative and controllable image synthesis with composable conditions
Lianghua Huang, Di Chen, Yu Liu, et al. 2023 · 2023
Closest in time.
Word-as-image for semantic typography
Shir Iluz, Yael Vinker, Amir Hertz, et al. 2023 · 2023
Closest in time.
Deepfloyd if
DeepFloyd Lab. 2023 · 2023
Closest in time.
Glyphdraw: Learning to draw chinese characters in image synthesis models coherently
Jian Ma, Mingjun Zhao, Chen Chen, et al. 2023 · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Yogesh Balaji, Seungjun Nah, Xun Huang, et al. 2022 · 2022
Cited alongside, same era.
Instructpix2pix: Learning to follow image editing instructions
Tim Brooks, Aleksander Holynski, and Alexei A. Efros. 2022 · 2022
Cited alongside, same era.
Clipdraw: Exploring text-to-drawing synthesis through language-image encoders
Kevin Frans, Lisa B. Soros, and Olaf Witkowski. 2022 · 2022
Cited alongside, same era.
Sdedit: Guided image synthesis and editing with stochastic differential equations
Chenlin Meng, Yutong He, Yang Song, et al. 2022 · 2022
Cited alongside, same era.
Hierarchical text-conditional image generation with CLIP latents
Aditya Ramesh, Prafulla Dhariwal, Alex Nichol, et al. 2022 · 2022
Cited alongside, same era.
High-resolution image synthesis with latent diffusion models
Robin Rombach, Andreas Blattmann, Dominik Lorenz, et al. 2022 · 2022
Cited alongside, same era.
Photorealistic text-to-image diffusion models with deep language understanding
Chitwan Saharia, William Chan, Saurabh Saxena, et al. 2022 · 2022
Cited alongside, same era.
Chong Mou, Xintao Wang, Liangbin Xie, et al. 2023 · 2023
Closest in time.
OpenAI. 2023 · 2023
Closest in time.
Dreamfusion: Text-to-3d using 2d diffusion
Ben Poole, Ajay Jain, Jonathan T. Barron, and Ben Mildenhall. 2023 · 2023
Closest in time.
Hugginggpt: Solving ai tasks with chatgpt and its friends in huggingface
Yongliang Shen, Kaitao Song, Xu Tan, et al. 2023 · 2023
Closest in time.
Ds-fusion: Artistic typography via discriminated and stylized diffusion
Maham Tanveer, Yizhi Wang, Ali Mahdavi-Amiri, and Hao Zhang. 2023 · 2023
Closest in time.
Visual chatgpt: Talking, drawing and editing with visual foundation models
Chenfei Wu, Shengming Yin, Weizhen Qi, et al. 2023 · 2023
Closest in time.
Glyphcontrol: Glyph conditional control for visual text generation
Yukang Yang, Dongnan Gui, Yuhui Yuan, et al. 2023 · 2023
Closest in time.
Adding conditional control to text-to-image diffusion models
Lvmin Zhang and Maneesh Agrawala. 2023 · 2023
Closest in time.
Controlvideo: Training-free controllable text-to-video generation
Yabo Zhang, Yuxiang Wei, Dongsheng Jiang, et al. 2023 · 2023
Closest in time.