Fetching the paper…
Reading the bibliography…
Recent progresses in large-scale text-to-image models have yielded remarkable accomplishments, finding various applications in art domain.
Multi-concept customization of text-to-image diffusion
Kumari, N.; Zhang, B.; Zhang, R.; Shechtman, E.; and Zhu, J.-Y. 2023 · 1941
Earlier work this paper cites.
Very deep convolutional networks for large-scale image recognition
Simonyan, K.; and Zisserman, A. 2014 · 2014
Earlier work this paper cites.
Image style transfer using convolutional neural networks
Gatys, L. A.; Ecker, A. S.; and Bethge, M. 2016 · 2016
Earlier work this paper cites.
Arbitrary style transfer in real-time with adaptive instance normalization
Huang, X.; and Belongie, S. 2017 · 2017
Earlier work this paper cites.
Improved ArtGAN for Conditional Synthesis of Natural Image and Artwork
Tan, W. R.; Chan, C. S.; Aguirre, H.; and Tanaka, K. 2019 · 2019
Earlier work this paper cites.
Artistic style transfer with internal-external learning and contrastive learning
Chen, H.; Wang, Z.; Zhang, H.; Zuo, Z.; Li, A.; Xing, W.; Lu, D.; et al. 2021 · 2021
Earlier work this paper cites.
Adaattn: Revisit attention mechanism in arbitrary neural style transfer
Liu, S.; Lin, T.; He, D.; Li, F.; Wang, M.; Li, X.; Sun, Z.; Li, Q.; and Ding, E. 2021 · 2021
Earlier work this paper cites.
Sdedit: Guided image synthesis and editing with stochastic differential equations
Meng, C.; He, Y.; Song, Y.; Song, J.; Wu, J.; Zhu, J.-Y.; and Ermon, S. 2021 · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
Radford, A.; Kim, J. W.; Hallacy, C.; Ramesh, A.; Goh, G.; Agarwal, S.; Sastry, G.; Askell, A.; Mishkin, P.; Clark, J.; et al. 2021 · 2021
Earlier work this paper cites.
ediffi: Text-to-image diffusion models with an ensemble of expert denoisers
Balaji, Y.; Nah, S.; Huang, X.; Vahdat, A.; Song, J.; Kreis, K.; Aittala, M.; Aila, T.; Laine, S.; Catanzaro, B.; et al. 2022 · 2022
Cited alongside, same era.
Perception prioritized training of diffusion models
Choi, J.; Lee, J.; Shin, C.; Kim, S.; Kim, H.; and Yoon, S. 2022 · 2022
Cited alongside, same era.
StyTr2: Image Style Transfer with Transformers
Deng, Y.; Tang, F.; Dong, W.; Ma, C.; Pan, X.; Wang, L.; and Xu, C. 2022 · 2022
Cited alongside, same era.
An image is worth one word: Personalizing text-to-image generation using textual inversion
Gal, R.; Alaluf, Y.; Atzmon, Y.; Patashnik, O.; Bermano, A. H.; Chechik, G.; and Cohen-Or, D. 2022 · 2022
Cited alongside, same era.
Classifier-free diffusion guidance
Ho, J.; and Salimans, T. 2022 · 2022
AesUST: towards aesthetic-enhanced universal style transfer
Wang, Z.; Zhang, Z.; Zhao, L.; Zuo, Z.; Li, A.; Xing, W.; and Lu, D. 2022 · 2022
Later among the works it cites.
Interactive Cartoonization with Controllable Perceptual Factors
Ahn, N.; Kwon, P.; Back, J.; Hong, K.; and Kim, S. 2023 · 2023
Closest in time.
AesPA-Net: Aesthetic Pattern-Aware Style Transfer Networks
Hong, K.; Jeon, S.; Lee, J.; Ahn, N.; Kim, K.; Lee, P.; Kim, D.; Uh, Y.; and Byun, H. 2023 · 2023
Closest in time.
Li, J.; Li, D.; Savarese, S.; and Hoi, S. 2023 · 2023
Closest in time.
Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation
Ruiz, N.; Li, Y.; Jampani, V.; Pritch, Y.; Rubinstein, M.; and Aberman, K. 2023 · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
UPainting: Unified Text-to-Image Diffusion Generation with Cross-modal Guidance
Li, W.; Xu, X.; Xiao, X.; Liu, J.; Yang, H.; Li, G.; Wang, Z.; Feng, Z.; She, Q.; Lyu, Y.; et al. 2022 · 2022
Cited alongside, same era.
Hierarchical text-conditional image generation with clip latents
Ramesh, A.; Dhariwal, P.; Nichol, A.; Chu, C.; and Chen, M. 2022 · 2022
Cited alongside, same era.
High-resolution image synthesis with latent diffusion models
Rombach, R.; Blattmann, A.; Lorenz, D.; Esser, P.; and Ommer, B. 2022 · 2022
Cited alongside, same era.
Photorealistic text-to-image diffusion models with deep language understanding
Saharia, C.; Chan, W.; Saxena, S.; Li, L.; Whang, J.; Denton, E. L.; Ghasemipour, K.; Gontijo Lopes, R.; Karagol Ayan, B.; Salimans, T.; et al. 2022 · 2022
Cited alongside, same era.
Sohn, K.; Ruiz, N.; Lee, K.; Chin, D. C.; Blok, I.; Chang, H.; Barber, J.; Jiang, L.; Entis, G.; Li, Y.; et al. 2023 · 2023
Closest in time.
P + P+ : Extended Textual Conditioning in Text-to-Image Generation
Voynov, A.; Chu, Q.; Cohen-Or, D.; and Aberman, K. 2023 · 2023
Closest in time.
Adding conditional control to text-to-image diffusion models
Zhang, L.; and Agrawala, M. 2023 · 2023
Closest in time.
Inversion-based style transfer with diffusion models
Zhang, Y.; Huang, N.; Tang, F.; Huang, H.; Ma, C.; Dong, W.; and Xu, C. 2023 · 2023
Closest in time.