Fetching the paper…
Reading the bibliography…
Personalized image generation has made significant strides in adapting content to novel concepts.
Multi-concept customization of text-to-image diffusion
Kumari, N.; Zhang, B.; Zhang, R.; Shechtman, E.; and Zhu, J.-Y. 2023 · 1941
Earlier work this paper cites.
Labeled faces in the wild: A database forstudying face recognition in unconstrained environments
Huang, G. B.; Mattar, M.; Berg, T.; and Learned-Miller, E. 2008 · 2008
Earlier work this paper cites.
Adam: A method for stochastic optimization
Kingma, D. P.; and Ba, J. 2015 · 2015
Earlier work this paper cites.
Ba, J. L.; Kiros, J. R.; and Hinton, G. E. 2016 · 2016
Earlier work this paper cites.
Joint face detection and alignment using multitask cascaded convolutional networks
Zhang, K.; Zhang, Z.; Li, Z.; and Qiao, Y. 2016 · 2016
Earlier work this paper cites.
Progressive growing of gans for improved quality, stability, and variation
Karras, T.; Aila, T.; Laine, S.; and Lehtinen, J. 2018 · 2018
Earlier work this paper cites.
Cosface: Large margin cosine loss for deep face recognition
Wang, H.; Wang, Y.; Zhou, Z.; Ji, X.; Gong, D.; Zhou, J.; Li, Z.; and Liu, W. 2018 · 2018
Earlier work this paper cites.
Arcface: Additive angular margin loss for deep face recognition
Deng, J.; Guo, J.; Xue, N.; and Zafeiriou, S. 2019 · 2019
Earlier work this paper cites.
A style-based generator architecture for generative adversarial networks
Karras, T.; Laine, S.; and Aila, T. 2019 · 2019
Earlier work this paper cites.
Countering language drift via visual grounding
Lee, J.; Cho, K.; and Kiela, D. 2019 · 2019
Earlier work this paper cites.
Generative modeling by estimating gradients of the data distribution
Song, Y.; and Ermon, S. 2019 · 2019
Earlier work this paper cites.
Generative adversarial networks
Goodfellow, I.; Pouget-Abadie, J.; Mirza, M.; Xu, B.; Warde-Farley, D.; Ozair, S.; Courville, A.; and Bengio, Y. 2020 · 2020
Earlier work this paper cites.
Analyzing and improving the image quality of stylegan
Karras, T.; Laine, S.; Aittala, M.; Hellsten, J.; Lehtinen, J.; and Aila, T. 2020 · 2020
Earlier work this paper cites.
Countering language drift with seeded iterated learning
Lu, Y.; Singhal, S.; Strub, F.; Courville, A.; and Pietquin, O. 2020 · 2020
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Raffel, C.; Shazeer, N.; Roberts, A.; Lee, K.; Narang, S.; Matena, M.; Zhou, Y.; Li, W.; and Liu, P. J. 2020 · 2020
Earlier work this paper cites.
Taming transformers for high-resolution image synthesis
Esser, P.; Rombach, R.; and Ommer, B. 2021 · 2021
Earlier work this paper cites.
CLIPScore: A Reference-free Evaluation Metric for Image Captioning
Hessel, J.; Holtzman, A.; Forbes, M.; Bras, R. L.; and Choi, Y. 2021 · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision
Radford, A.; Kim, J. W.; Hallacy, C.; Ramesh, A.; Goh, G.; Agarwal, S.; Sastry, G.; Askell, A.; Mishkin, P.; Clark, J.; et al. 2021 · 2021
Cited alongside, same era.
Zero-shot text-to-image generation
Ramesh, A.; Pavlov, M.; Goh, G.; Gray, S.; Voss, C.; Radford, A.; Chen, M.; and Sutskever, I. 2021 · 2021
Cited alongside, same era.
Denoising diffusion implicit models
Song, J.; Meng, C.; and Ermon, S. 2021 · 2021
Cited alongside, same era.
Score-based generative modeling through stochastic differential equations
Song, Y.; Sohl-Dickstein, J.; Kingma, D. P.; Kumar, A.; Ermon, S.; and Poole, B. 2021 · 2021
Cited alongside, same era.
From continuity to editability: Inverting gans with consecutive images
Pivotal tuning for latent-based editing of real images
Roich, D.; Mokady, R.; Bermano, A. H.; and Cohen-Or, D. 2022 · 2022
Later among the works it cites.
High-resolution image synthesis with latent diffusion models
Rombach, R.; Blattmann, A.; Lorenz, D.; Esser, P.; and Ommer, B. 2022 · 2022
Later among the works it cites.
Photorealistic text-to-image diffusion models with deep language understanding
Saharia, C.; Chan, W.; Saxena, S.; Li, L.; Whang, J.; Denton, E. L.; Ghasemipour, K.; Gontijo Lopes, R.; Karagol Ayan, B.; Salimans, T.; et al. 2022 · 2022
Later among the works it cites.
Editing out-of-domain gan inversion via differential activations
Song, H.; Du, Y.; Xiang, T.; Dong, J.; Qin, J.; and He, S. 2022 · 2022
Later among the works it cites.
Improved vector quantized diffusion models
Tang, Z.; Gu, S.; Bao, J.; Chen, D.; and Wen, F. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Xu, Y.; Du, Y.; Xiao, W.; Xu, X.; and He, S. 2021 · 2021
Cited alongside, same era.
Discovering interpretable latent space directions of gans beyond binary attributes
Yang, H.; Chai, L.; Wen, Q.; Zhao, S.; Sun, Z.; and He, S. 2021 · 2021
Cited alongside, same era.
Hyperstyle: Stylegan inversion with hypernetworks for real image editing
Alaluf, Y.; Tov, O.; Mokady, R.; Gal, R.; and Bermano, A. 2022 · 2022
Cited alongside, same era.
ediffi: Text-to-image diffusion models with an ensemble of expert denoisers
Balaji, Y.; Nah, S.; Huang, X.; Vahdat, A.; Song, J.; Kreis, K.; Aittala, M.; Aila, T.; Laine, S.; Catanzaro, B.; et al. 2022 · 2022
Cited alongside, same era.
Stable Diffusion
CompVis. 2022 · 2022
Cited alongside, same era.
Vector quantized diffusion model for text-to-image synthesis
Gu, S.; Chen, D.; Bao, J.; Wen, F.; Zhang, B.; Chen, D.; Yuan, L.; and Guo, B. 2022 · 2022
Cited alongside, same era.
Classifier-free diffusion guidance
Ho, J.; and Salimans, T. 2022 · 2022
Cited alongside, same era.
Diffusers: State-of-the-art diffusion models
von Platen, P.; Patil, S.; Lozhkov, A.; Cuenca, P.; Lambert, N.; Rasul, K.; Davaadorj, M.; and Wolf, T. 2022 · 2022
Later among the works it cites.
Scaling Autoregressive Models for Content-Rich Text-to-Image Generation
Yu, J.; Xu, Y.; Koh, J. Y.; Luong, T.; Baid, G.; Wang, Z.; Vasudevan, V.; Ku, A.; Yang, Y.; Ayan, B. K.; Hutchinson, B.; Han, W.; Parekh, Z.; Li, X.; Zhang, H.; Baldridge, J.; and Wu, Y. 2022 · 2022
Later among the works it cites.
A neural space-time representation for text-to-image personalization
Alaluf, Y.; Richardson, E.; Metzer, G.; and Cohen-Or, D. 2023 · 2023
Later among the works it cites.
Blended latent diffusion
Avrahami, O.; Fried, O.; and Lischinski, D. 2023 · 2023
Later among the works it cites.
Key-Locked-Rank-One-Editing-for-Text-to-Image-Personalization
ChenDarYen. 2023 · 2023
Later among the works it cites.
Dreamlike Photoreal
Dreamlike.art. 2023 · 2023
Later among the works it cites.
Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation
Ruiz, N.; Li, Y.; Jampani, V.; Pritch, Y.; Rubinstein, M.; and Aberman, K. 2023 · 2023
Later among the works it cites.
Key-locked rank one editing for text-to-image personalization
Tewel, Y.; Gal, R.; Chechik, G.; and Atzmon, Y. 2023 · 2023
Later among the works it cites.
Ip-adapter: Text compatible image prompt adapter for text-to-image diffusion models
Ye, H.; Zhang, J.; Liu, S.; Han, X.; and Yang, W. 2023 · 2023
Later among the works it cites.
Photomaker: Customizing realistic human photos via stacked id embedding
Li, Z.; Cao, M.; Wang, X.; Qi, Z.; Cheng, M.-M.; and Shan, Y. 2024 · 2024
Closest in time.