Fetching the paper…
Reading the bibliography…
Consistent editing of real images is a challenging task, as it requires performing non-rigid edits (e.g., changing postures) to the main objects in the input image without changing their identity or attributes.
Stackgan++: Realistic image synthesis with stacked generative adversarial networks
Zhang, H.; Xu, T.; Li, H.; Zhang, S.; Wang, X.; Huang, X.; and Metaxas, D. N. 2018 · 1962
Earlier work this paper cites.
Denoising diffusion implicit models
Song, J.; Meng, C.; and Ermon, S. 2020 · 2010
Earlier work this paper cites.
Microsoft coco: Common objects in context
Lin, T.-Y.; Maire, M.; Belongie, S.; Hays, J.; Perona, P.; Ramanan, D.; Dollár, P.; and Zitnick, C. L. 2014 · 2014
Earlier work this paper cites.
Deep residual learning for image recognition
He, K.; Zhang, X.; Ren, S.; and Sun, J. 2016 · 2016
Earlier work this paper cites.
Generative adversarial text to image synthesis
Reed, S.; Akata, Z.; Yan, X.; Logeswaran, L.; Schiele, B.; and Lee, H. 2016 · 2016
Earlier work this paper cites.
Attention is all you need
Vaswani, A.; Shazeer, N.; Parmar, N.; Uszkoreit, J.; Jones, L.; Gomez, A. N.; Kaiser, Ł.; and Polosukhin, I. 2017 · 2017
Earlier work this paper cites.
Large scale GAN training for high fidelity natural image synthesis
Brock, A.; Donahue, J.; and Simonyan, K. 2018 · 2018
Earlier work this paper cites.
Text-adaptive generative adversarial networks: manipulating images with natural language
Nam, S.; Kim, Y.; and Kim, S. J. 2018 · 2018
Earlier work this paper cites.
Attngan: Fine-grained text to image generation with attentional generative adversarial networks
Xu, T.; Zhang, P.; Huang, Q.; Zhang, H.; Gan, Z.; Huang, X.; and He, X. 2018 · 2018
Earlier work this paper cites.
Generative modeling by estimating gradients of the data distribution
Song, Y.; and Ermon, S. 2019 · 2019
Earlier work this paper cites.
Denoising diffusion probabilistic models
Ho, J.; Jain, A.; and Abbeel, P. 2020 · 2020
Earlier work this paper cites.
Manigan: Text-guided image manipulation
Li, B.; Qi, X.; Lukasiewicz, T.; and Torr, P. H. 2020 · 2020
Earlier work this paper cites.
Diffusion models beat gans on image synthesis
Dhariwal, P.; and Nichol, A. 2021 · 2021
Earlier work this paper cites.
Cogview: Mastering text-to-image generation via transformers
Ding, M.; Yang, Z.; Hong, W.; Zheng, W.; Zhou, C.; Yin, D.; Lin, J.; Zou, X.; Shao, Z.; Yang, H.; et al. 2021 · 2021
Earlier work this paper cites.
Sdedit: Guided image synthesis and editing with stochastic differential equations
Meng, C.; He, Y.; Song, Y.; Song, J.; Wu, J.; Zhu, J.-Y.; and Ermon, S. 2021 · 2021
Earlier work this paper cites.
Glide: Towards photorealistic image generation and editing with text-guided diffusion models
Nichol, A.; Dhariwal, P.; Ramesh, A.; Shyam, P.; Mishkin, P.; McGrew, B.; Sutskever, I.; and Chen, M. 2021 · 2021
Cited alongside, same era.
Improved denoising diffusion probabilistic models
Nichol, A. Q.; and Dhariwal, P. 2021 · 2021
Cited alongside, same era.
Zero-shot text-to-image generation
Ramesh, A.; Pavlov, M.; Goh, G.; Gray, S.; Voss, C.; Radford, A.; Chen, M.; and Sutskever, I. 2021 · 2021
Cited alongside, same era.
Tedigan: Text-guided diverse face image generation and manipulation
Xia, W.; Yang, Y.; Xue, J.-H.; and Wu, B. 2021 · 2021
Cited alongside, same era.
Cross-modal contrastive learning for text-to-image generation
Zhang, H.; Koh, J. Y.; Baldridge, J.; Lee, H.; and Yang, Y. 2021 · 2021
Cited alongside, same era.
Hierarchical text-conditional image generation with clip latents
Ramesh, A.; Dhariwal, P.; Nichol, A.; Chu, C.; and Chen, M. 2022 · 2022
Later among the works it cites.
High-resolution image synthesis with latent diffusion models
Rombach, R.; Blattmann, A.; Lorenz, D.; Esser, P.; and Ommer, B. 2022 · 2022
Later among the works it cites.
Df-gan: A simple and effective baseline for text-to-image synthesis
Tao, M.; Tang, H.; Wu, F.; Jing, X.-Y.; Bao, B.-K.; and Xu, C. 2022 · 2022
Later among the works it cites.
Plug-and-Play Diffusion Features for Text-Driven Image-to-Image Translation
Tumanyan, N.; Geyer, M.; Bagon, S.; and Dekel, T. 2022 · 2022
Later among the works it cites.
Scaling autoregressive models for content-rich text-to-image generation
Yu, J.; Xu, Y.; Koh, J. Y.; Luong, T.; Baid, G.; Wang, Z.; Vasudevan, V.; Ku, A.; Yang, Y.; Ayan, B. K.; et al. 2022 · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Avrahami, O.; Fried, O.; and Lischinski, D. 2022 · 2022
Cited alongside, same era.
InstructPix2Pix: Learning to Follow Image Editing Instructions
Brooks, T.; Holynski, A.; and Efros, A. A. 2022 · 2022
Cited alongside, same era.
Vqgan-clip: Open domain image generation and editing with natural language guidance
Crowson, K.; Biderman, S.; Kornis, D.; Stander, D.; Hallahan, E.; Castricato, L.; and Raff, E. 2022 · 2022
Cited alongside, same era.
CogView2: Faster and Better Text-to-Image Generation via Hierarchical Transformers
Ding, M.; Zheng, W.; Hong, W.; and Tang, J. 2022 · 2022
Cited alongside, same era.
Vector quantized diffusion model for text-to-image synthesis
Gu, S.; Chen, D.; Bao, J.; Wen, F.; Zhang, B.; Chen, D.; Yuan, L.; and Guo, B. 2022 · 2022
Cited alongside, same era.
Prompt-to-prompt image editing with cross attention control
Hertz, A.; Mokady, R.; Tenenbaum, J.; Aberman, K.; Pritch, Y.; and Cohen-Or, D. 2022 · 2022
Cited alongside, same era.
Classifier-free diffusion guidance
Ho, J.; and Salimans, T. 2022 · 2022
Cited alongside, same era.
Later among the works it cites.
Tigan: Text-based interactive image generation and manipulation
Zhou, Y.; Zhang, R.; Gu, J.; Tensmeyer, C.; Yu, T.; Chen, C.; Xu, J.; and Sun, T. 2022 · 2022
Later among the works it cites.
MasaCtrl: Tuning-Free Mutual Self-Attention Control for Consistent Image Synthesis and Editing
Cao, M.; Wang, X.; Qi, Z.; Shan, Y.; Qie, X.; and Zheng, Y. 2023 · 2023
Closest in time.
Prompt Tuning Inversion for Text-Driven Image Editing Using Diffusion Models
Dong, W.; Xue, S.; Duan, X.; and Han, S. 2023 · 2023
Closest in time.
Improving Tuning-Free Real Image Editing with Proximal Guidance
Han, L.; Wen, S.; Chen, Q.; Zhang, Z.; Song, K.; Ren, M.; Gao, R.; Chen, Y.; 0003, D. L.; Zhangli, Q.; et al. 2023 · 2023
Closest in time.
Mou, C.; Wang, X.; Xie, L.; Zhang, J.; Qi, Z.; Shan, Y.; and Qie, X. 2023 · 2023
Closest in time.
Drag Your GAN: Interactive Point-based Manipulation on the Generative Image Manifold
Pan, X.; Tewari, A.; Leimkuhler, T.; Liu, L.; Meka, A.; and Theobalt, C. 2023 · 2023
Closest in time.
Zero-shot Image-to-Image Translation
Parmar, G.; Singh, K. K.; Zhang, R.; Li, Y.; Lu, J.; and Zhu, J.-Y. 2023 · 2023
Closest in time.
Adding Conditional Control to Text-to-Image Diffusion Models
Zhang, L.; and Agrawala, M. 2023 · 2023
Closest in time.
Styleclip: Text-driven manipulation of stylegan imagery
Patashnik, O.; Wu, Z.; Shechtman, E.; Cohen-Or, D.; and Lischinski, D. 2021 · 2094
Closest in time.