Fetching the paper…
Reading the bibliography…
We consider the problem of constraining diffusion model outputs with a user-supplied reference image.
Multi-concept customization of text-to-image diffusion
Kumari, N.; Zhang, B.; Zhang, R.; Shechtman, E.; and Zhu, J.-Y. 2023 · 1941
Earlier work this paper cites.
Inverting the generator of a generative adversarial network
Creswell, A.; and Bharath, A. A. 2018 · 1974
Earlier work this paper cites.
Denoising diffusion implicit models
Song, J.; Meng, C.; and Ermon, S. 2020 · 2010
Earlier work this paper cites.
Color thief
Dhakar, L. 2015 · 2015
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation
Ronneberger, O.; Fischer, P.; and Brox, T. 2015 · 2015
Earlier work this paper cites.
Deep unsupervised learning using nonequilibrium thermodynamics
Sohl-Dickstein, J.; Weiss, E.; Maheswaranathan, N.; and Ganguli, S. 2015 · 2015
Earlier work this paper cites.
Precise recovery of latent vectors from generative adversarial networks
Lipton, Z. C.; and Tripathi, S. 2017 · 2017
Earlier work this paper cites.
Image2stylegan: How to embed images into the stylegan latent space?
Abdal, R.; Qin, Y.; and Wonka, P. 2019 · 2019
Earlier work this paper cites.
Image2stylegan++: How to edit the embedded images?
Abdal, R.; Qin, Y.; and Wonka, P. 2020 · 2020
Earlier work this paper cites.
Denoising diffusion probabilistic models
Ho, J.; Jain, A.; and Abbeel, P. 2020 · 2020
Earlier work this paper cites.
Glide: Towards photorealistic image generation and editing with text-guided diffusion models
Nichol, A.; Dhariwal, P.; Ramesh, A.; Shyam, P.; Mishkin, P.; McGrew, B.; Sutskever, I.; and Chen, M. 2021 · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
Radford, A.; Kim, J. W.; Hallacy, C.; Ramesh, A.; Goh, G.; Agarwal, S.; Sastry, G.; Askell, A.; Mishkin, P.; Clark, J.; et al. 2021 · 2021
Cited alongside, same era.
Hyperstyle: Stylegan inversion with hypernetworks for real image editing
Alaluf, Y.; Tov, O.; Mokady, R.; Gal, R.; and Bermano, A. 2022 · 2022
Cited alongside, same era.
State-of-the-Art in the Architecture, Methods and Applications of StyleGAN
Bermano, A. H.; Gal, R.; Alaluf, Y.; Mokady, R.; Nitzan, Y.; Tov, O.; Patashnik, O.; and Cohen-Or, D. 2022 · 2022
Cited alongside, same era.
Training-free structured diffusion guidance for compositional text-to-image synthesis
Feng, W.; He, X.; Fu, T.-J.; Jampani, V.; Akula, A.; Narayana, P.; Basu, S.; Wang, X. E.; and Wang, W. Y. 2022 · 2022
Cited alongside, same era.
An image is worth one word: Personalizing text-to-image generation using textual inversion
Gal, R.; Alaluf, Y.; Atzmon, Y.; Patashnik, O.; Bermano, A. H.; Chechik, G.; and Cohen-Or, D. 2022 · 2022
Photorealistic text-to-image diffusion models with deep language understanding
Saharia, C.; Chan, W.; Saxena, S.; Li, L.; Whang, J.; Denton, E.; Ghasemipour, S. K. S.; Ayan, B. K.; Mahdavi, S. S.; Lopes, R. G.; et al. 2022 · 2022
Later among the works it cites.
Gan inversion: A survey
Xia, W.; Zhang, Y.; Yang, Y.; Xue, J.-H.; Zhou, B.; and Yang, M.-H. 2022 · 2022
Later among the works it cites.
A-STAR: Test-time Attention Segregation and Retention for Text-to-image Synthesis
Agarwal, A.; Karanam, S.; Joseph, K.; Saxena, A.; Goswami, K.; and Srinivasan, B. V. 2023 · 2023
Closest in time.
Attend-and-Excite: Attention-Based Semantic Guidance for Text-to-Image Diffusion Models
Chefer, H.; Alaluf, Y.; Vinker, Y.; Wolf, L.; and Cohen-Or, D. 2023 · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Compositional visual generation with composable diffusion models
Liu, N.; Li, S.; Du, Y.; Torralba, A.; and Tenenbaum, J. B. 2022 · 2022
Cited alongside, same era.
Mystyle: A personalized generative prior
Nitzan, Y.; Aberman, K.; He, Q.; Liba, O.; Yarom, M.; Gandelsman, Y.; Mosseri, I.; Pritch, Y.; and Cohen-Or, D. 2022 · 2022
Cited alongside, same era.
Hierarchical text-conditional image generation with clip latents
Ramesh, A.; Dhariwal, P.; Nichol, A.; Chu, C.; and Chen, M. 2022 · 2022
Cited alongside, same era.
Pivotal tuning for latent-based editing of real images
Roich, D.; Mokady, R.; Bermano, A. H.; and Cohen-Or, D. 2022 · 2022
Cited alongside, same era.
High-resolution image synthesis with latent diffusion models
Rombach, R.; Blattmann, A.; Lorenz, D.; Esser, P.; and Ommer, B. 2022 · 2022
Cited alongside, same era.
Mou, C.; Wang, X.; Xie, L.; Zhang, J.; Qi, Z.; Shan, Y.; and Qie, X. 2023 · 2023
Closest in time.
Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation
Ruiz, N.; Li, Y.; Jampani, V.; Pritch, Y.; Rubinstein, M.; and Aberman, K. 2023 · 2023
Closest in time.
Key-locked rank one editing for text-to-image personalization
Tewel, Y.; Gal, R.; Chechik, G.; and Atzmon, Y. 2023 · 2023
Closest in time.
P + P+ : Extended Textual Conditioning in Text-to-Image Generation
Voynov, A.; Chu, Q.; Cohen-Or, D.; and Aberman, K. 2023 · 2023
Closest in time.
Adding conditional control to text-to-image diffusion models
Zhang, L.; and Agrawala, M. 2023 · 2023
Closest in time.
ProSpect: Expanded Conditioning for the Personalization of Attribute-aware Image Generation
Zhang, Y.; Dong, W.; Tang, F.; Huang, N.; Huang, H.; Ma, C.; Lee, T.-Y.; Deussen, O.; and Xu, C. 2023 · 2023
Closest in time.