Fetching the paper…
Reading the bibliography…
Although remarkable progress has been made in image style transfer, style is just one of the components of artistic paintings.
O. Ronneberger, P. Fischer, and T. Brox, “U-Net: Convolutional networks for biomedical image segmentation,” in Medical Image Computing and Computer-Assisted Intervention (MICCAI) . Springer, 2015, pp. 234–241
2015
Earlier work this paper cites.
X. Huang and S. Belongie, “Arbitrary style transfer in real-time with adaptive instance normalization,” in IEEE International Conference on Computer Vision (ICCV) , 2017, pp. 1501–1510
2017
Earlier work this paper cites.
J. Liao, Y. Yao, L. Yuan, G. Hua, and S. B. Kang, “Visual attribute transfer through deep image analogy,” ACM Transactions on Graphics , vol. 36, no. 4, pp. 120:1–120:15, 2017
2017
Earlier work this paper cites.
Y. Li, C. Fang, J. Yang, Z. Wang, X. Lu, and M.-H. Yang, “Universal style transfer via feature transforms,” in Advances Neural Information Processing Systems (NeurIPS) , 2017, pp. 386–396
2017
Earlier work this paper cites.
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, Ł. u. Kaiser, and I. Polosukhin, “Attention is all you need,” in Advances in Neural Information Processing Systems (NIPS) , vol. 30. Curran Associates, Inc., 2017
2017
Earlier work this paper cites.
S.-S. Lin, C. C. Morace, C.-H. Lin, L.-F. Hsu, and T.-Y. Lee, “Generation of escher arts with dual perception,” IEEE Transactions on Visualization and Computer Graphics , vol. 24, no. 2, pp. 1103–1113, 2018
2018
Earlier work this paper cites.
H. Talebi and P. Milanfar, “NIMA: Neural image assessment,” IEEE Transactions on Image Processing , vol. 27, no. 8, pp. 3998–4011, 2018
2018
Earlier work this paper cites.
D. Y. Park and K. H. Lee, “Arbitrary style transfer with style-attentional networks,” in IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2019, pp. 5880–5888
2019
Earlier work this paper cites.
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova, “BERT: Pre-training of deep bidirectional transformers for language understanding,” Annual Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (NAACL-HLT) , pp. 4171–4186, 2019
2019
Earlier work this paper cites.
A. Razavi, A. Van den Oord, and O. Vinyals, “Generating diverse high-fidelity images with VQ-VAE-2,” in Advances in Neural Information Processing Systems (NeurIPS) , vol. 32, 2019
2019
Earlier work this paper cites.
T. Karras, S. Laine, and T. Aila, “A style-based generator architecture for generative adversarial networks,” in IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2019, pp. 4401–4410
2019
Earlier work this paper cites.
H. Pflüger, D. Thom, A. Schütz, D. Bohde, and T. Ertl, “VeCHArt: Visually enhanced comparison of historic art using an automated line-based synchronization technique,” IEEE Transactions on Visualization and Computer Graphics , vol. 26, no. 10, pp. 3063–3076, 2020
2020
Earlier work this paper cites.
Y. Jing, Y. Yang, Z. Feng, J. Ye, Y. Yu, and M. Song, “Neural style transfer: A review,” IEEE Transactions on Visualization and Computer Graphics , vol. 26, no. 11, pp. 3365–3385, 2020
2020
Earlier work this paper cites.
Y. Deng, F. Tang, W. Dong, H. Huang, C. Ma, and C. Xu, “Arbitrary video style transfer via multi-channel correlation,” in AAAI Conference on Artificial Intelligence (AAAI) , 2021, pp. 1210–1217
2021
Earlier work this paper cites.
J. An, S. Huang, Y. Song, D. Dou, W. Liu, and J. Luo, “ArtFlow: Unbiased image style transfer via reversible neural flows,” in IEEE/CVF Conferences on Computer Vision and Pattern Recognition (CVPR) , 2021, pp. 862–871
2021
Earlier work this paper cites.
A. Ramesh, M. Pavlov, G. Goh, S. Gray, C. Voss, A. Radford, M. Chen, and I. Sutskever, “Zero-shot text-to-image generation,” in International Conference on Machine Learning (ICML) . PMLR, 2021, pp. 8821–8831
2021
Earlier work this paper cites.
J. Song, C. Meng, and S. Ermon, “Denoising diffusion implicit models,” in International Conference on Learning Representations (ICLR) , 2021
2021
Earlier work this paper cites.
P. Dhariwal and A. Nichol, “Diffusion models beat GANs on image synthesis,” in Advances in Neural Information Processing Systems (NeurIPS) , vol. 34, 2021, pp. 8780–8794
2021
Cited alongside, same era.
2021
Cited alongside, same era.
R. Rombach, A. Blattmann, D. Lorenz, P. Esser, and B. Ommer, “High-resolution image synthesis with latent diffusion models,” in IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2022, pp. 10 684–10 695
2022
Cited alongside, same era.
2022
Cited alongside, same era.
M. Cao, X. Wang, Z. Qi, Y. Shan, X. Qie, and Y. Zheng, “MasaCtrl: Tuning-free mutual self-attention control for consistent image synthesis and editing,” in IEEE/CVF International Conference on Computer Vision (ICCV) , 2023, pp. 22 560–22 570
2023
Later among the works it cites.
L. Zhang, A. Rao, and M. Agrawala, “Adding conditional control to text-to-image diffusion models,” in IEEE/CVF International Conference on Computer Vision (ICCV) , 2023, pp. 3836–3847
2023
Later among the works it cites.
Y. Zhang, W. Dong, F. Tang, N. Huang, H. Huang, C. Ma, T.-Y. Lee, O. Deussen, and C. Xu, “ProSpect: Prompt spectrum for attribute-aware personalization of diffusion models,” ACM Transactions on Graphics , vol. 42, no. 6, pp. 244:1–244:14, 2023
2023
Later among the works it cites.
R. Mokady, A. Hertz, K. Aberman, Y. Pritch, and D. Cohen-Or, “Null-text inversion for editing real images using guided diffusion models,” in IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2023, pp. 6038–6047
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A. Nichol, P. Dhariwal, A. Ramesh, P. Shyam, P. Mishkin, B. McGrew, I. Sutskever, and M. Chen, “GLIDE: Towards photorealistic image generation and editing with text-guided diffusion models,” in International Conference on Machine Learning (ICML) , 2022
2022
Cited alongside, same era.
C.-K. Yeh, Z. Liu, I.-H. Lin, E. Zhang, and T.-Y. Lee, “WYSIWYG design of hypnotic line art,” IEEE Transactions on Visualization and Computer Graphics , vol. 28, no. 6, pp. 2517–2529, 2022
2022
Cited alongside, same era.
Y. Zhang, F. Tang, W. Dong, H. Huang, C. Ma, T.-Y. Lee, and C. Xu, “Domain enhanced arbitrary image style transfer via contrastive learning,” in ACM SIGGRAPH 2022 Conference Proceedings , 2022, pp. 12:1–12:8
2022
Cited alongside, same era.
Y. Deng, F. Tang, W. Dong, C. Ma, X. Pan, L. Wang, and C. Xu, “StyTr 2 : Image style transfer with transformers,” in IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2022, pp. 11 326–11 336
2022
Cited alongside, same era.
N. Huang, F. Tang, W. Dong, and C. Xu, “Draw your art dream: Diverse digital art synthesis with multimodal guided diffusion,” in ACM International Conference on Multimedia , 2022, p. 1085–1094
2022
Cited alongside, same era.
N. Ruiz, Y. Li, V. Jampani, Y. Pritch, M. Rubinstein, and K. Aberman, “DreamBooth: Fine tuning text-to-image diffusion models for subject-driven generation,” in IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2023, pp. 22 500–22 510
2023
Cited alongside, same era.
Y. Tewel, R. Gal, G. Chechik, and Y. Atzmon, “Key-locked rank one editing for text-to-image personalization,” in ACM SIGGRAPH 2023 Conference Proceedings , ser. SIGGRAPH ’23. New York, NY, USA: Association for Computing Machinery, 2023, pp. 12:1–12:11
2023
Cited alongside, same era.
R. Gal, M. Arar, Y. Atzmon, A. H. Bermano, G. Chechik, and D. Cohen-Or, “Encoder-based domain tuning for fast personalization of text-to-image models,” ACM Transactions on Graphics , vol. 42, no. 4, pp. 150:1–150:13, 2023
2023
Cited alongside, same era.
2023
Later among the works it cites.
2023
Later among the works it cites.
X. Xu, Z. Wang, G. Zhang, K. Wang, and H. Shi, “Versatile diffusion: Text, images and variations all in one diffusion model,” in IEEE/CVF International Conference on Computer Vision (ICCV) , 2023, pp. 7754–7765
2023
Later among the works it cites.
T. Brooks, A. Holynski, and A. A. Efros, “InstructPix2Pix: Learning to follow image editing instructions,” in IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2023, pp. 18 392–18 402
2023
Later among the works it cites.
Y. Cao, X. Meng, P. Y. Mok, T.-Y. Lee, X. Liu, and P. Li, “AnimeDiffusion: Anime diffusion colorization,” IEEE Transactions on Visualization and Computer Graphics , vol. 30, no. 10, p. 6956–6969, 2024
2024
Closest in time.
H. Ma, J. Yang, and H. Huang, “Taming diffusion model for exemplar-based image translation,” Computational Visual Media , vol. 10, p. 1031–1043, 2024
2024
Closest in time.
S. H. Lee, S. Kim, W. Byeon, G. Oh, S. In, H. Park, S. H. Yoon, S.-H. Hong, J. Kim, and S. Kim, “Audio-guided implicit neural representation for local image stylization,” Computational Visual Media , vol. 10, no. 6, pp. 1185–1204, 2024
2024
Closest in time.
Y. Deng, X. He, F. Tang, and W. Dong, “Z*: Zero-shot style transfer via attention reweighting,” in IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2024, pp. 6934–6944
2024
Closest in time.
A. Hertz, A. Voynov, S. Fruchter, and D. Cohen-Or, “Style aligned image generation via shared attention,” in IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , 2024, pp. 4775–4785
2024
Closest in time.
Pinkney, Justin, “Image mixer,” 2024, last accessed on 2024-01-10. [Online]. Available: https://github.com/justinpinkney/stable-diffusion#image-mixer
2024
Closest in time.
X. Pan, L. Dong, S. Huang, Z. Peng, W. Chen, and F. Wei, “Kosmos-G: Generating images in context with multimodal large language models,” in International Conference on Learning Representations (ICLR) , 2024
2024
Closest in time.
Beaumont, Romain and Schuhmann, Christoph, “Aesthetic predictor,” 2024, last accessed on 2024-01-15. [Online]. Available: https://github.com/LAION-AI/aesthetic-predictor
2024
Closest in time.
N. Huang, Y. Zhang, F. Tang, C. Ma, H. Huang, W. Dong, and C. Xu, “DiffStyler: Controllable dual diffusion for text-driven image stylization,” IEEE Transactions on Neural Networks and Learning Systems , vol. 36, no. 2, pp. 3370–3383, 2025
2025
Closest in time.