Fetching the paper…
Reading the bibliography…
A hallmark of human intelligence is the ability to create complex artifacts through structured multi-step processes.
Paint by numbers: Abstract image representations
Haeberli, P · 1990
Earlier work this paper cites.
Processing images and video for an impressionist effect
Litwinowicz, P · 1997
Earlier work this paper cites.
A survey of stroke-based rendering
Hertzmann, A · 2003
Earlier work this paper cites.
Artist agent: A reinforcement learning approach to automatic stroke generation in oriental ink painting
Xie, N., Hachiya, H., and Sugiyama, M · 2013
Earlier work this paper cites.
A neural representation of sketch drawings
Ha, D. and Eck, D · 2017
Earlier work this paper cites.
Learning to sketch with deep q networks and demonstrated strokes
Zhou, T., Fang, C., Wang, Z., Yang, J., Kim, B., Chen, Z., Brandt, J., and Terzopoulos, D · 2018
Earlier work this paper cites.
Neural painters: A learned differentiable constraint for generating brushstroke paintings
Nakano, R · 2019
Earlier work this paper cites.
Denoising diffusion probabilistic models
Ho, J., Jain, A., and Abbeel, P · 2020
Earlier work this paper cites.
Multi-modal attention for speech emotion recognition
Pan, Z., Luo, Z., Yang, J., and Li, H · 2020
Earlier work this paper cites.
Denoising diffusion implicit models
Song, J., Meng, C., and Ermon, S · 2020
Earlier work this paper cites.
Rethinking style transfer: From pixels to parameterized brushstrokes
Kotovenko, D., Wright, M., Heimbrecht, A., and Ommer, B · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
Radford, A., Kim, J. W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., Krueger, G., and Sutskever, I · 2021
Earlier work this paper cites.
Clipdraw: Exploring text-to-drawing synthesis through language-image encoders
Frans, K., Soros, L., and Witkowski, O · 2022
Earlier work this paper cites.
Prompt-to-prompt image editing with cross attention control
Hertz, A., Mokady, R., Tenenbaum, J., Aberman, K., Pritch, Y., and Cohen-Or, D · 2022
Earlier work this paper cites.
LoRA: Low-rank adaptation of large language models
Hu, E. J., Shen, Y., Wallis, P., Allen-Zhu, Z., Li, Y., Wang, S., Wang, L., and Chen, W · 2022
Cited alongside, same era.
High-resolution image synthesis with latent diffusion models
Rombach, R., Blattmann, A., Lorenz, D., Esser, P., and Ommer, B · 2022
Cited alongside, same era.
Cliptexture: Text-driven texture synthesis
Song, Y · 2022
Cited alongside, same era.
Clipfont: Text guided vector wordart generation
Song, Y. and Zhang, Y · 2022
Cited alongside, same era.
Stable video diffusion: Scaling latent video diffusion models to large datasets
Blattmann, A., Dockhorn, T., Kulal, S., Mendelevitch, D., Kilian, M., Lorenz, D., Levi, Y., English, Z., Voleti, V., Letts, A., et al · 2023
Cited alongside, same era.
Instructpix2pix: Learning to follow image editing instructions
Adding conditional control to text-to-image diffusion models
Zhang, L. and Agrawala, M · 2023
Later among the works it cites.
Uni-controlnet: All-in-one control to text-to-image diffusion models
Zhao, S., Chen, D., Chen, Y.-C., Bao, J., Hao, S., Yuan, L., and Wong, K.-Y. K · 2023
Later among the works it cites.
Flux.1 ai, 2024
AI, F · 2024
Later among the works it cites.
Inverse painting: Reconstructing the painting process
Chen, B., Wang, Y., Curless, B., Kemelmacher-Shlizerman, I., and Seitz, S. M · 2024
Later among the works it cites.
Scaling rectified flow transformers for high-resolution image synthesis
Esser, P., Kulal, S., Blattmann, A., Entezari, R., Müller, J., Saini, H., Levi, Y., Lorenz, D., Sauer, A., Boesel, F., et al · 2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Brooks, T., Holynski, A., and Efros, A. A · 2023
Cited alongside, same era.
Animatediff: Animate your personalized text-to-image diffusion models without specific tuning
Guo, Y., Yang, C., Rao, A., Wang, Y., Qiao, Y., Lin, D., and Dai, B · 2023
Cited alongside, same era.
Multi-concept customization of text-to-image diffusion
Kumari, N., Zhang, B., Zhang, R., Shechtman, E., and Zhu, J.-Y · 2023
Cited alongside, same era.
Mou, C., Wang, X., Xie, L., Zhang, J., Qi, Z., Shan, Y., and Qie, X · 2023
Cited alongside, same era.
Scalable diffusion models with transformers
Peebles, W. and Xie, S · 2023
Cited alongside, same era.
Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation
Ruiz, N., Li, Y., Jampani, V., Pritch, Y., Rubinstein, M., and Aberman, K · 2023
Cited alongside, same era.
Clipvg: Text-guided image manipulation using differentiable vector graphics
Song, Y., Shao, X., Chen, K., Zhang, W., Jing, Z., and Li, M · 2023
Cited alongside, same era.
Roformer: Enhanced transformer with rotary position embedding
Su, J., Ahmed, M., Lu, Y., Pan, S., Bo, W., and Liu, Y · 2024
Later among the works it cites.
Ominicontrol: Minimal and universal control for diffusion transformer
Tan, Z., Liu, S., Yang, X., Xue, Q., and Wang, X · 2024
Later among the works it cites.
Paints-undo github page, 2024
Team, P.-U · 2024
Later among the works it cites.
Grid: Visual layout generation
Wan, C., Luo, X., Cai, Z., Song, Y., Zhao, Y., Bai, Y., He, Y., and Gong, Y · 2024
Later among the works it cites.
Fast personalized text to image synthesis with attention injection
Zhang, Y., Song, Y., Yu, J., Pan, H., and Jing, Z · 2024
Later among the works it cites.
Asymmetry in low-rank adapters of foundation models
Zhu, J., Greenewald, K., Nadjahi, K., Borde, H. S. d. O., Gabrielsson, R. B., Choshen, L., Ghassemi, M., Yurochkin, M., and Solomon, J · 2024
Later among the works it cites.
Civitai website
Civitai · 2025
Closest in time.
Ideogram ai, 2023
Ideogram · 2025
Closest in time.