Fetching the paper…
Reading the bibliography…
Recent advancements in generative models have significantly facilitated the development of personalized content creation.
Brown T, Mann B, Ryder N, Subbiah M, Kaplan JD, Dhariwal P, Neelakantan A, Shyam P, Sastry G, Askell A, et al.. Language models are few-shot learners. Advances in neural information processing systems , 2020, 33: 1877–1901
1901
Earlier work this paper cites.
Kumari N, Zhang B, Zhang R, Shechtman E, Zhu JY. Multi-concept customization of text-to-image diffusion. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, 1931–1941
1941
Earlier work this paper cites.
Creswell A, Bharath AA. Inverting the generator of a generative adversarial network. IEEE transactions on neural networks and learning systems , 2018, 30(7): 1967–1974
1974
Earlier work this paper cites.
2015
Earlier work this paper cites.
2015
Earlier work this paper cites.
Zhu JY, Krähenbühl P, Shechtman E, Efros AA. Generative visual manipulation on the natural image manifold. In Computer Vision–ECCV 2016: 14th European Conference, Amsterdam, The Netherlands, October 11-14, 2016, Proceedings, Part V 14 , 2016, 597–613
2016
Earlier work this paper cites.
Szegedy C, Vanhoucke V, Ioffe S, Shlens J, Wojna Z. Rethinking the inception architecture for computer vision. In Proceedings of the IEEE conference on computer vision and pattern recognition , 2016, 2818–2826
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
Zhang H, Xu T, Li H, Zhang S, Wang X, Huang X, Metaxas DN. Stackgan: Text to photo-realistic image synthesis with stacked generative adversarial networks. In Proceedings of the IEEE international conference on computer vision , 2017, 5907–5915
2017
Earlier work this paper cites.
Mao X, Li Q, Xie H, Lau RY, Wang Z, Paul Smolley S. Least squares generative adversarial networks. In Proceedings of the IEEE international conference on computer vision , 2017, 2794–2802
2017
Earlier work this paper cites.
Arjovsky M, Chintala S, Bottou L. Wasserstein generative adversarial networks. In International conference on machine learning , 2017, 214–223
2017
Earlier work this paper cites.
Gulrajani I, Ahmed F, Arjovsky M, Dumoulin V, Courville AC. Improved training of wasserstein gans. Advances in neural information processing systems , 2017, 30
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
Heusel M, Ramsauer H, Unterthiner T, Nessler B, Hochreiter S. Gans trained by a two time-scale update rule converge to a local nash equilibrium. Advances in neural information processing systems , 2017, 30
2017
Earlier work this paper cites.
Shah V, Hegde C. Solving linear inverse problems using gan priors: An algorithm with provable guarantees. In 2018 IEEE international conference on acoustics, speech and signal processing (ICASSP) , 2018, 4609–4613
2018
Earlier work this paper cites.
Ma F, Ayaz U, Karaman S. Invertibility of convolutional generative networks from partial measurements. Advances in Neural Information Processing Systems , 2018, 31
2018
Earlier work this paper cites.
Shen Y, Yang C, Tang X, Zhou B. Interfacegan: Interpreting the disentangled face representation learned by gans. IEEE transactions on pattern analysis and machine intelligence , 2020, 44(4): 2004–2018
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
Radford A. Improving language understanding by generative pre-training, 2018
2018
Earlier work this paper cites.
Abdal R, Qin Y, Wonka P. Image2stylegan: How to embed images into the stylegan latent space? In Proceedings of the IEEE/CVF international conference on computer vision , 2019, 4432–4441
2019
Earlier work this paper cites.
Karras T, Laine S, Aila T. A style-based generator architecture for generative adversarial networks. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2019, 4401–4410
2019
Earlier work this paper cites.
Radford A, Wu J, Child R, Luan D, Amodei D, Sutskever I, et al.. Language models are unsupervised multitask learners. OpenAI blog , 2019, 1(8): 9
2019
Earlier work this paper cites.
Liu M, Ding Y, Xia M, Liu X, Ding E, Zuo W, Wen S. Stgan: A unified selective transfer network for arbitrary image attribute editing. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2019, 3673–3682
2019
Earlier work this paper cites.
Bau D, Zhu JY, Wulff J, Peebles W, Strobelt H, Zhou B, Torralba A. Inverting layers of a large generator. In ICLR workshop , 2019, 4
2019
Earlier work this paper cites.
Deng J, Guo J, Xue N, Zafeiriou S. Arcface: Additive angular margin loss for deep face recognition. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2019, 4690–4699
2019
Earlier work this paper cites.
2019
Earlier work this paper cites.
Goodfellow I, Pouget-Abadie J, Mirza M, Xu B, Warde-Farley D, Ozair S, Courville A, Bengio Y. Generative adversarial networks. Communications of the ACM , 2020, 63(11): 139–144
2020
Earlier work this paper cites.
Ho J, Jain A, Abbeel P. Denoising diffusion probabilistic models. Advances in neural information processing systems , 2020, 33: 6840–6851
2020
Earlier work this paper cites.
Abdal R, Qin Y, Wonka P. Image2stylegan++: How to edit the embedded images? In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2020, 8296–8305
2020
Earlier work this paper cites.
Viazovetskyi Y, Ivashkin V, Kashin E. Stylegan2 distillation for feed-forward image manipulation. In Computer Vision–ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part XXII 16 , 2020, 170–186
2020
Earlier work this paper cites.
Peebles W, Peebles J, Zhu JY, Efros A, Torralba A. The hessian penalty: A weak prior for unsupervised disentanglement. In Computer Vision–ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part VI 16 , 2020, 581–597
2020
Earlier work this paper cites.
Voynov A, Babenko A. Unsupervised discovery of interpretable directions in the gan latent space. In International conference on machine learning , 2020, 9786–9796
2020
Earlier work this paper cites.
Härkönen E, Hertzmann A, Lehtinen J, Paris S. Ganspace: Discovering interpretable gan controls. Advances in neural information processing systems , 2020, 33: 9841–9850
2020
Earlier work this paper cites.
Song J, Meng C, Ermon S. Denoising diffusion implicit models. arXiv preprint arXiv:2010.02502 , 2020
2020
Earlier work this paper cites.
Karras T, Laine S, Aittala M, Hellsten J, Lehtinen J, Aila T. Analyzing and improving the image quality of stylegan. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2020, 8110–8119
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
He K, Fan H, Wu Y, Xie S, Girshick R. Momentum contrast for unsupervised visual representation learning. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2020, 9729–9738
2020
Earlier work this paper cites.
Zhu J, Shen Y, Zhao D, Zhou B. In-domain gan inversion for real image editing. In European conference on computer vision , 2020, 592–608
2020
Earlier work this paper cites.
Huang Y, Wang Y, Tai Y, Liu X, Shen P, Li S, Li J, Huang F. Curricularface: adaptive curriculum learning loss for deep face recognition. In proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2020, 5901–5910
2020
Earlier work this paper cites.
Richardson E, Alaluf Y, Patashnik O, Nitzan Y, Azar Y, Shapiro S, Cohen-Or D. Encoding in style: a stylegan encoder for image-to-image translation. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2021, 2287–2296
2021
Earlier work this paper cites.
Tov O, Alaluf Y, Nitzan Y, Patashnik O, Cohen-Or D. Designing an encoder for stylegan image manipulation. ACM Transactions on Graphics (TOG) , 2021, 40(4): 1–14
2021
Earlier work this paper cites.
Ramesh A, Pavlov M, Goh G, Gray S, Voss C, Radford A, Chen M, Sutskever I. Zero-shot text-to-image generation. In International conference on machine learning , 2021, 8821–8831
2021
Earlier work this paper cites.
Ding M, Yang Z, Hong W, Zheng W, Zhou C, Yin D, Lin J, Zou X, Shao Z, Yang H, et al.. Cogview: Mastering text-to-image generation via transformers. Advances in neural information processing systems , 2021, 34: 19822–19835
2021
Earlier work this paper cites.
Wu Z, Lischinski D, Shechtman E. Stylespace analysis: Disentangled controls for stylegan image generation. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2021, 12863–12872
2021
Earlier work this paper cites.
Alaluf Y, Patashnik O, Cohen-Or D. Restyle: A residual-based stylegan encoder via iterative refinement. In Proceedings of the IEEE/CVF international conference on computer vision , 2021, 6711–6720
2021
Earlier work this paper cites.
Shen Y, Zhou B. Closed-form factorization of latent semantics in gans. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2021, 1532–1540
2021
Earlier work this paper cites.
Wang HP, Yu N, Fritz M. Hijack-gan: Unintended-use of pretrained, black-box gans. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2021, 7872–7881
2021
Earlier work this paper cites.
Abdal R, Zhu P, Mitra NJ, Wonka P. Styleflow: Attribute-conditioned exploration of stylegan-generated images using conditional continuous normalizing flows. ACM Transactions on Graphics (ToG) , 2021, 40(3): 1–21
2021
Earlier work this paper cites.
Tzelepis C, Tzimiropoulos G, Patras I. Warpedganspace: Finding non-linear rbf paths in gan latent space. In Proceedings of the IEEE/CVF international conference on computer vision , 2021, 6393–6402
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
Karras T, Aittala M, Laine S, Härkönen E, Hellsten J, Lehtinen J, Aila T. Alias-free generative adversarial networks. Advances in neural information processing systems , 2021, 34: 852–863
2021
Earlier work this paper cites.
Radford A, Kim JW, Hallacy C, Ramesh A, Goh G, Agarwal S, Sastry G, Askell A, Mishkin P, Clark J, et al.. Learning transferable visual models from natural language supervision. In International conference on machine learning , 2021, 8748–8763
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
Kang K, Kim S, Cho S. Gan inversion for out-of-range images with geometric transformations. In Proceedings of the IEEE/CVF international conference on computer vision , 2021, 13941–13949
2021
Earlier work this paper cites.
Chai L, Zhu JY, Shechtman E, Isola P, Zhang R. Ensembling with deep generative views. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2021, 14997–15007
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
Rombach R, Blattmann A, Lorenz D, Esser P, Ommer B. High-resolution image synthesis with latent diffusion models. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2022, 10684–10695
2022
Earlier work this paper cites.
Saharia C, Chan W, Saxena S, Li L, Whang J, Denton EL, Ghasemipour K, Gontijo Lopes R, Karagol Ayan B, Salimans T, et al.. Photorealistic text-to-image diffusion models with deep language understanding. Advances in neural information processing systems , 2022, 35: 36479–36494
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
Xia W, Zhang Y, Yang Y, Xue JH, Zhou B, Yang MH. Gan inversion: A survey. IEEE transactions on pattern analysis and machine intelligence , 2022, 45(3): 3121–3138
2022
Earlier work this paper cites.
Wang T, Zhang Y, Fan Y, Wang J, Chen Q. High-fidelity gan inversion for image attribute editing. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2022, 11379–11388
2022
Earlier work this paper cites.
Parmar G, Li Y, Lu J, Zhang R, Zhu JY, Singh KK. Spatially-adaptive multilayer selection for gan inversion and editing. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2022, 11399–11409
2022
Earlier work this paper cites.
Roich D, Mokady R, Bermano AH, Cohen-Or D. Pivotal tuning for latent-based editing of real images. ACM Transactions on graphics (TOG) , 2022, 42(1): 1–13
2022
Earlier work this paper cites.
Alaluf Y, Tov O, Mokady R, Gal R, Bermano A. Hyperstyle: Stylegan inversion with hypernetworks for real image editing. In Proceedings of the IEEE/CVF conference on computer Vision and pattern recognition , 2022, 18511–18521
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
Dinh TM, Tran AT, Nguyen R, Hua BS. Hyperinverter: Improving stylegan inversion via hypernetwork. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2022, 11389–11398
2022
Earlier work this paper cites.
Wei T, Chen D, Zhou W, Liao J, Zhang W, Yuan L, Hua G, Yu N. E2Style: Improve the efficiency and effectiveness of StyleGAN inversion. IEEE Transactions on Image Processing , 2022, 31: 3267–3280
2022
Earlier work this paper cites.
Hu X, Huang Q, Shi Z, Li S, Gao C, Sun L, Li Q. Style transformer for image inversion and editing. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, 11337–11346
2022
Earlier work this paper cites.
Kocasari U, Dirik A, Tiftikci M, Yanardag P. Stylemc: Multi-channel based fast text-guided image generation and manipulation. In Proceedings of the IEEE/CVF Winter Conference on Applications of Computer vision , 2022, 895–904
2022
Earlier work this paper cites.
Xu Z, Lin T, Tang H, Li F, He D, Sebe N, Timofte R, Van Gool L, Ding E. Predict, prevent, and evaluate: Disentangled text-driven image manipulation empowered by pre-trained vision-language model. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, 18229–18238
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
Zhu Y, Liu H, Song Y, Yuan Z, Han X, Yuan C, Chen Q, Wang J. One model to edit them all: Free-form text-driven image manipulation with semantic modulations. Advances in Neural Information Processing Systems , 2022, 35: 25146–25159
2022
Earlier work this paper cites.
Wei T, Chen D, Zhou W, Liao J, Tan Z, Yuan L, Zhang W, Yu N. Hairclip: Design your hair by text and reference image. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2022, 18072–18081
2022
Earlier work this paper cites.
Gal R, Patashnik O, Maron H, Bermano AH, Chechik G, Cohen-Or D. Stylegan-nada: Clip-guided domain adaptation of image generators. ACM Transactions on Graphics (TOG) , 2022, 41(4): 1–13
2022
Earlier work this paper cites.
Zhang Y, Yao M, Wei Y, Ji Z, Bai J, Zuo W, et al.. Towards diverse and faithful one-shot adaption of generative adversarial networks. Advances in Neural Information Processing Systems , 2022, 35: 37297–37308
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
Schuhmann C, Beaumont R, Vencu R, Gordon C, Wightman R, Cherti M, Coombes T, Katta A, Mullis C, Wortsman M, et al.. Laion-5b: An open large-scale dataset for training next generation image-text models. Advances in Neural Information Processing Systems , 2022, 35: 25278–25294
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
Ding M, Zheng W, Hong W, Tang J. Cogview2: Faster and better text-to-image generation via hierarchical transformers. Advances in Neural Information Processing Systems , 2022, 35: 16890–16902
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
Ho J, Salimans T. Classifier-free diffusion guidance. arXiv preprint arXiv:2207.12598 , 2022
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
2022
Earlier work this paper cites.
Sun Q, Yu Q, Cui Y, Zhang F, Zhang X, Wang Y, Gao H, Liu J, Huang T, Wang X. Emu: Generative pretraining in multimodality. In The Twelfth International Conference on Learning Representations , 2023
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
Zhang L, Rao A, Agrawala M. Adding conditional control to text-to-image diffusion models. In Proceedings of the IEEE/CVF International Conference on Computer Vision , 2023, 3836–3847
2023
Earlier work this paper cites.
Cao M, Wang X, Qi Z, Shan Y, Qie X, Zheng Y. Masactrl: Tuning-free mutual self-attention control for consistent image synthesis and editing. In Proceedings of the IEEE/CVF International Conference on Computer Vision , 2023, 22560–22570
2023
Earlier work this paper cites.
Ruiz N, Li Y, Jampani V, Pritch Y, Rubinstein M, Aberman K. Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , 2023, 22500–22510
2023
Earlier work this paper cites.
Wei Y, Zhang Y, Ji Z, Bai J, Zhang L, Zuo W. Elite: Encoding visual concepts into textual embeddings for customized text-to-image generation. In Proceedings of the IEEE/CVF International Conference on Computer Vision , 2023, 15943–15953
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
Liu M, Wei Y, Wu X, Zuo W, Zhang L. Survey on leveraging pre-trained generative adversarial networks for image editing and restoration. Science China Information Sciences , 2023, 66(5): 151101
2023
Earlier work this paper cites.
Gal R, Arar M, Atzmon Y, Bermano AH, Chechik G, Cohen-Or D. Encoder-based domain tuning for fast personalization of text-to-image models. ACM Transactions on Graphics (TOG) , 2023, 42(4): 1–13
2023
Earlier work this paper cites.
Arar M, Gal R, Atzmon Y, Chechik G, Cohen-Or D, Shamir A, H Bermano A. Domain-agnostic tuning-encoder for fast personalization of text-to-image models. In SIGGRAPH Asia 2023 Conference Papers , 2023, 1–10
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
Zhuang C, Gao P, Smolic A. StylePrompter: All Styles Need Is Attention. In Proceedings of the 31st ACM International Conference on Multimedia , 2023, 2487–2497
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
Pan X, Tewari A, Leimkühler T, Liu L, Meka A, Theobalt C. Drag your gan: Interactive point-based manipulation on the generative image manifold. In ACM SIGGRAPH 2023 Conference Proceedings , 2023, 1–11
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
Mokady R, Hertz A, Aberman K, Pritch Y, Cohen-Or D. Null-text inversion for editing real images using guided diffusion models. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2023, 6038–6047
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
Alaluf Y, Richardson E, Metzer G, Cohen-Or D. A neural space-time representation for text-to-image personalization. ACM Transactions on Graphics (TOG) , 2023, 42(6): 1–10
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
Tewel Y, Gal R, Chechik G, Atzmon Y. Key-locked rank one editing for text-to-image personalization. In ACM SIGGRAPH 2023 Conference Proceedings , 2023, 1–11
2023
Earlier work this paper cites.
2023
Earlier work this paper cites.
Epstein D, Jabri A, Poole B, Efros A, Holynski A. Diffusion self-guidance for controllable image generation. Advances in Neural Information Processing Systems , 2023, 36: 16222–16239
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Qiu Z, Liu W, Feng H, Xue Y, Feng Y, Liu Z, Zhang D, Weller A, Schölkopf B. Controlling text-to-image diffusion by orthogonal finetuning. Advances in Neural Information Processing Systems , 2023, 36: 79320–79362
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
Avrahami O, Aberman K, Fried O, Cohen-Or D, Lischinski D. Break-a-scene: Extracting multiple concepts from a single image. In SIGGRAPH Asia 2023 Conference Papers , 2023, 1–12
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Esser P, Kulal S, Blattmann A, Entezari R, Müller J, Saini H, Levi Y, Lorenz D, Sauer A, Boesel F, et al.. Scaling rectified flow transformers for high-resolution image synthesis. In Forty-first International Conference on Machine Learning , 2024
2024
Later among the works it cites.
Team K. Kolors: Effective Training of Diffusion Model for Photorealistic Text-to-Image Synthesis. arXiv preprint , 2024
2024
Later among the works it cites.
Labs BF. FLUX. https://github.com/black-forest-labs/flux , 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Zhang H, Wu C, Cao G, Wang H, Cao W. HyperEditor: Achieving Both Authenticity and Cross-Domain Capability in Image Editing via Hypernetworks. In Proceedings of the AAAI Conference on Artificial Intelligence , 2024, 7051–7059
2024
Later among the works it cites.
Yildirim AB, Pehlivan H, Dundar A. Warping the residuals for image editing with stylegan. International Journal of Computer Vision , 2024: 1–16
2024
Later among the works it cites.
Nguyen T, Ojha U, Li Y, Liu H, Lee YJ. Edit One for All: Interactive Batch Image Editing. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, 8271–8280
2024
Later among the works it cites.
Huberman-Spiegelglas I, Kulikov V, Michaeli T. An edit friendly ddpm noise space: Inversion and manipulations. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, 12469–12478
2024
Later among the works it cites.
Xu J, Liu X, Wu Y, Tong Y, Li Q, Ding M, Tang J, Dong Y. Imagereward: Learning and evaluating human preferences for text-to-image generation. Advances in Neural Information Processing Systems , 2024, 36
2024
Later among the works it cites.
Chen X, Huang L, Liu Y, Shen Y, Zhao D, Zhao H. Anydoor: Zero-shot object-level image customization. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, 6593–6602
2024
Later among the works it cites.
Chen W, Hu H, Li Y, Ruiz N, Jia X, Chang MW, Cohen WW. Subject-driven text-to-image generation via apprenticeship learning. Advances in Neural Information Processing Systems , 2024, 36
2024
Later among the works it cites.
2024
Later among the works it cites.
Ruiz N, Li Y, Jampani V, Wei W, Hou T, Pritch Y, Wadhwa N, Rubinstein M, Aberman K. Hyperdreambooth: Hypernetworks for fast personalization of text-to-image models. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, 6527–6536
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Ham C, Fisher M, Hays J, Kolkin N, Liu Y, Zhang R, Hinz T. Personalized Residuals for Concept-Driven Text-to-Image Generation. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, 8186–8195
2024
Later among the works it cites.
Zhang X, Wei XY, Wu J, Zhang T, Zhang Z, Lei Z, Li Q. Compositional inversion for stable diffusion models. In Proceedings of the AAAI Conference on Artificial Intelligence , 2024, 7350–7358
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Le DH, Pham T, Lee S, Clark C, Kembhavi A, Mandt S, Krishna R, Lu J. One Diffusion to Generate Them All, 2024
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Hyung J, Shin J, Choo J. Magicapture: High-resolution multi-concept portrait customization. In Proceedings of the AAAI Conference on Artificial Intelligence , 2024, 2445–2453
2024
Later among the works it cites.
2024
Later among the works it cites.
Cui S, Guo J, An X, Deng J, Zhao Y, Wei X, Feng Z. IDAdapter: Learning Mixed Features for Tuning-Free Personalization of Text-to-Image Models. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, 950–959
2024
Later among the works it cites.
Wang Y, Zhang W, Zheng J, Jin C. High-fidelity Person-centric Subject-to-Image Synthesis. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, 7675–7684
2024
Later among the works it cites.
2024
Later among the works it cites.
Liu R, Ma B, Zhang W, Hu Z, Fan C, Lv T, Ding Y, Cheng X. Towards a simultaneous and granular identity-expression control in personalized face generation. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, 2114–2123
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Xu Y, Zhai B, Zhang C, Li M, Li Y, Du S. Diff-PC: Identity-preserving and 3D-aware controllable diffusion for zero-shot portrait customization. Information Fusion , 2024: 102869
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Hertz A, Voynov A, Fruchter S, Cohen-Or D. Style aligned image generation via shared attention. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, 4775–4785
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Gu Y, Wang X, Wu JZ, Shi Y, Chen Y, Fan Z, Xiao W, Zhao R, Chang S, Wu W, et al.. Mix-of-show: Decentralized low-rank adaptation for multi-concept customization of diffusion models. Advances in Neural Information Processing Systems , 2024, 36
2024
Later among the works it cites.
Po R, Yang G, Aberman K, Wetzstein G. Orthogonal adaptation for modular customization of diffusion models. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, 7964–7973
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Zhang Y, Yang M, Zhou Q, Wang Z. Attention Calibration for Disentangled Text-to-Image Personalization. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, 4764–4774
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Pan X, Qin P, Li Y, Xue H, Chen W. Synthesizing coherent story with auto-regressive latent diffusion models. In Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision , 2024, 2920–2930
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Huang M, Mao Z, Liu M, He Q, Zhang Y. RealCustom: Narrowing Real Text Word for Real-Time Open-Domain Text-to-Image Customization. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, 7476–7485
2024
Later among the works it cites.
Wang KC, Ostashev D, Fang Y, Tulyakov S, Aberman K. Moa: Mixture-of-attention for subject-context disentanglement in personalized image generation. In SIGGRAPH Asia 2024 Conference Papers , 2024, 1–12
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Zhang Y, Xing Z, Zeng Y, Fang Y, Chen K. Pia: Your personalized image animator via plug-and-play modules in text-to-image models. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, 7747–7756
2024
Later among the works it cites.
Li X, Jia X, Wang Q, Diao H, Ge M, Li P, He Y, Lu H. Motrans: Customized motion transfer with text-driven video diffusion models. In Proceedings of the 32nd ACM International Conference on Multimedia , 2024, 3421–3430
2024
Later among the works it cites.
2024
Later among the works it cites.
2024
Later among the works it cites.
Jiang Y, Wu T, Yang S, Si C, Lin D, Qiao Y, Loy CC, Liu Z. Videobooth: Diffusion-based video generation with image prompts. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, 6689–6700
2024
Later among the works it cites.
Song K, Zhu Y, Liu B, Yan Q, Elgammal A, Yang X. Moma: Multimodal llm adapter for fast personalized image generation. In European Conference on Computer Vision , 2025, 117–132
2025
Closest in time.
Guan S, Ge Y, Tai Y, Yang J, Li W, You M. HybridBooth: Hybrid Prompt Inversion for Efficient Subject-Driven Generation. In European Conference on Computer Vision , 2025, 403–419
2025
Closest in time.
2025
Closest in time.
Ram S, Neiman T, Feng Q, Stuart A, Tran S, Chilimbi T. DreamBlend: Advancing personalized fine-tuning of text-to-image diffusion models, 2025
2025
Closest in time.
Parihar R, Sachidanand V, Mani S, Karmali T, Venkatesh Babu R. Precisecontrol: Enhancing text-to-image diffusion models with fine-grained attribute control. In European Conference on Computer Vision , 2025, 469–487
2025
Closest in time.
Wu Y, Li Z, Zheng H, Wang C, Li B. Infinite-ID: Identity-preserved Personalization via ID-semantics Decoupling Paradigm. In European Conference on Computer Vision , 2025, 279–296
2025
Closest in time.
Wei Y, Ji Z, Bai J, Zhang H, Zhang L, Zuo W. MasterWeaver: Taming Editability and Face Identity for Personalized Text-to-Image Generation. In European Conference on Computer Vision , 2025, 252–271
2025
Closest in time.
2025
Closest in time.
Shah V, Ruiz N, Cole F, Lu E, Lazebnik S, Li Y, Jampani V. Ziplora: Any subject in any style by effectively merging loras. In European Conference on Computer Vision , 2025, 422–438
2025
Closest in time.
Kong Z, Zhang Y, Yang T, Wang T, Zhang K, Wu B, Chen G, Liu W, Luo W. Omg: Occlusion-friendly personalized multi-concept generation in diffusion models. In European Conference on Computer Vision , 2025, 253–270
2025
Closest in time.
2025
Closest in time.
Jin J, Shen Y, Fu Z, Yang J. Customized Generation Reimagined: Fidelity and Editability Harmonized. In European Conference on Computer Vision , 2025, 410–426
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
Frenkel Y, Vinker Y, Shamir A, Cohen-Or D. Implicit style-content separation using b-lora. In European Conference on Computer Vision , 2025, 181–198
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
2025
Closest in time.
Li N, Liu Q, Singh KK, Wang Y, Zhang J, Plummer BA, Lin Z. UniHuman: A Unified Model For Editing Human Images in the Wild. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , 2024, 2039–2048
2048
Closest in time.
Patashnik O, Wu Z, Shechtman E, Cohen-Or D, Lischinski D. Styleclip: Text-driven manipulation of stylegan imagery. In Proceedings of the IEEE/CVF international conference on computer vision , 2021, 2085–2094
2094
Closest in time.