Fetching the paper…
Reading the bibliography…
Diffusion models have opened the path to a wide range of text-based image editing frameworks.
Improved StyleGAN Embedding: Where are the Good Latents?
Peihao Zhu, Rameen Abdal, Yipeng Qin, and Peter Wonka. 2020 · 2012
Earlier work this paper cites.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio. 2014 · 2014
Earlier work this paper cites.
Deep unsupervised learning using nonequilibrium thermodynamics. In International conference on machine learning . PMLR, 2256–2265
Jascha Sohl-Dickstein, Eric Weiss, Niru Maheswaranathan, and Surya Ganguli. 2015 · 2015
Earlier work this paper cites.
Generative visual manipulation on the natural image manifold. In European conference on computer vision . Springer, 597–613
Jun-Yan Zhu, Philipp Krähenbühl, Eli Shechtman, and Alexei A Efros. 2016 · 2016
Earlier work this paper cites.
The Unreasonable Effectiveness of Deep Features as a Perceptual Metric. In CVPR
Richard Zhang, Phillip Isola, Alexei A Efros, Eli Shechtman, and Oliver Wang. 2018 · 2018
Earlier work this paper cites.
Image2stylegan: How to embed images into the stylegan latent space?. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 4432–4441
Rameen Abdal, Yipeng Qin, and Peter Wonka. 2019 · 2019
Earlier work this paper cites.
Image2stylegan++: How to edit the embedded images?. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition . 8296–8305
Rameen Abdal, Yipeng Qin, and Peter Wonka. 2020 · 2020
Earlier work this paper cites.
Denoising diffusion probabilistic models
Jonathan Ho, Ajay Jain, and Pieter Abbeel. 2020 · 2020
Earlier work this paper cites.
Encoding in Style: a StyleGAN Encoder for Image-to-Image Translation
Elad Richardson, Yuval Alaluf, Or Patashnik, Yotam Nitzan, Yaniv Azar, Stav Shapiro, and Daniel Cohen-Or. 2020 · 2020
Earlier work this paper cites.
Denoising Diffusion Implicit Models. In International Conference on Learning Representations
Jiaming Song, Chenlin Meng, and Stefano Ermon. 2020 · 2020
Earlier work this paper cites.
ReStyle: A Residual-Based StyleGAN Encoder via Iterative Refinement
Yuval Alaluf, Or Patashnik, and Daniel Cohen-Or. 2021a · 2021
Earlier work this paper cites.
ILVR: Conditioning Method for Denoising Diffusion Probabilistic Models. In 2021 IEEE/CVF International Conference on Computer Vision (ICCV) . IEEE Computer Society, 14347–14356
Jooyoung Choi, Sungwon Kim, Yonghyun Jeong, Youngjune Gwon, and Sungroh Yoon. 2021 · 2021
Earlier work this paper cites.
Diffusion models beat gans on image synthesis
Prafulla Dhariwal and Alexander Nichol. 2021 · 2021
Earlier work this paper cites.
Stylegan-nada: Clip-guided domain adaptation of image generators
Rinon Gal, Or Patashnik, Haggai Maron, Gal Chechik, and Daniel Cohen-Or. 2021 · 2021
Earlier work this paper cites.
Classifier-Free Diffusion Guidance. In NeurIPS 2021 Workshop on Deep Generative Models and Downstream Applications
Jonathan Ho and Tim Salimans. 2021 · 2021
Earlier work this paper cites.
Glide: Towards photorealistic image generation and editing with text-guided diffusion models
Alex Nichol, Prafulla Dhariwal, Aditya Ramesh, Pranav Shyam, Pamela Mishkin, Bob McGrew, Ilya Sutskever, and Mark Chen. 2021 · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Earlier work this paper cites.
Designing an Encoder for StyleGAN Image Manipulation
Omer Tov, Yuval Alaluf, Yotam Nitzan, Or Patashnik, and Daniel Cohen-Or. 2021 · 2021
Earlier work this paper cites.
Hyperinverter: Improving stylegan inversion via hypernetwork. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 11389–11398
Tan M Dinh, Anh Tuan Tran, Rang Nguyen, and Binh-Son Hua. 2022 · 2022
Earlier work this paper cites.
Ziyi Dong, Pengxu Wei, and Liang Lin. 2022 · 2022
Earlier work this paper cites.
An image is worth one word: Personalizing text-to-image generation using textual inversion
Rinon Gal, Yuval Alaluf, Yuval Atzmon, Or Patashnik, Amit H Bermano, Gal Chechik, and Daniel Cohen-Or. 2022 · 2022
Earlier work this paper cites.
Prompt-to-prompt image editing with cross attention control
Amir Hertz, Ron Mokady, Jay Tenenbaum, Kfir Aberman, Yael Pritch, and Daniel Cohen-Or. 2022 · 2022
Earlier work this paper cites.
Elucidating the design space of diffusion-based generative models
Tero Karras, Miika Aittala, Timo Aila, and Samuli Laine. 2022 · 2022
Earlier work this paper cites.
Pseudo Numerical Methods for Diffusion Models on Manifolds. In International Conference on Learning Representations
Luping Liu, Yi Ren, Zhijie Lin, and Zhou Zhao. 2022 · 2022
Cited alongside, same era.
DPM-Solver: A Fast ODE Solver for Diffusion Probabilistic Model Sampling in Around 10 Steps. In Advances in Neural Information Processing Systems , Alice H. Oh, Alekh Agarwal, Danielle Belgrave, and Kyunghyun Cho (Eds.)
Cheng Lu, Yuhao Zhou, Fan Bao, Jianfei Chen, Chongxuan Li, and Jun Zhu. 2022 · 2022
Cited alongside, same era.
SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations. In International Conference on Learning Representations
Chenlin Meng, Yutong He, Yang Song, Jiaming Song, Jiajun Wu, Jun-Yan Zhu, and Stefano Ermon. 2022 · 2022
Cited alongside, same era.
Spatially-Adaptive Multilayer Selection for GAN Inversion and Editing. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 11399–11409
Gaurav Parmar, Yijun Li, Jingwan Lu, Richard Zhang, Jun-Yan Zhu, and Krishna Kumar Singh. 2022 · 2022
Cited alongside, same era.
Effective Real Image Editing with Accelerated Iterative Diffusion Inversion. In Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) . 15912–15921
Zhihong Pan, Riccardo Gherardi, Xiufeng Xie, and Stephen Huang. 2023 · 2023
Later among the works it cites.
Zero-shot Image-to-Image Translation
Gaurav Parmar, Krishna Kumar Singh, Richard Zhang, Yijun Li, Jingwan Lu, and Jun-Yan Zhu. 2023 · 2023
Later among the works it cites.
Localizing object-level shape variations with text-to-image diffusion models. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 23051–23061
Or Patashnik, Daniel Garibi, Idan Azuri, Hadar Averbuch-Elor, and Daniel Cohen-Or. 2023 · 2023
Later among the works it cites.
DreamFusion: Text-to-3D using 2D Diffusion. In The Eleventh International Conference on Learning Representations
Ben Poole, Ajay Jain, Jonathan T. Barron, and Ben Mildenhall. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Aditya Ramesh, Prafulla Dhariwal, Alex Nichol, Casey Chu, and Mark Chen. 2022 · 2022
Cited alongside, same era.
High-Resolution Image Synthesis with Latent Diffusion Models. In IEEE/CVF Conference on Computer Vision and Pattern Recognition, CVPR 2022, New Orleans, LA, USA, June 18-24, 2022 . IEEE, 10674–10685
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Björn Ommer. 2022 · 2022
Cited alongside, same era.
Progressive Distillation for Fast Sampling of Diffusion Models. In International Conference on Learning Representations
Tim Salimans and Jonathan Ho. 2022 · 2022
Cited alongside, same era.
Plug-and-Play Diffusion Features for Text-Driven Image-to-Image Translation
Narek Tumanyan, Michal Geyer, Shai Bagon, and Tali Dekel. 2022 · 2022
Cited alongside, same era.
Cross-image attention for zero-shot appearance transfer
Yuval Alaluf, Daniel Garibi, Or Patashnik, Hadar Averbuch-Elor, and Daniel Cohen-Or. 2023a · 2023
Cited alongside, same era.
Blended Latent Diffusion
Omri Avrahami, Ohad Fried, and Dani Lischinski. 2023 · 2023
Cited alongside, same era.
Ledits++: Limitless image editing using text-to-image models
Manuel Brack, Felix Friedrich, Katharina Kornmeier, Linoy Tsaban, Patrick Schramowski, Kristian Kersting, and Apolinário Passos. 2023 · 2023
Cited alongside, same era.
InstructPix2Pix: Learning to Follow Image Editing Instructions. In CVPR
Tim Brooks, Aleksander Holynski, and Alexei A. Efros. 2023 · 2023
Cited alongside, same era.
Axel Sauer, Dominik Lorenz, Andreas Blattmann, and Robin Rombach. 2023 · 2023
Later among the works it cites.
Consistency models. In Proceedings of the 40th International Conference on Machine Learning . 32211–32252
Yang Song, Prafulla Dhariwal, Mark Chen, and Ilya Sutskever. 2023 · 2023
Later among the works it cites.
Ledits: Real image editing with ddpm inversion and semantic guidance
Linoy Tsaban and Apolinário Passos. 2023 · 2023
Later among the works it cites.
Unitune: Text-driven image editing by fine tuning a diffusion model on a single image
Dani Valevski, Matan Kalman, Eyal Molad, Eyal Segalis, Yossi Matias, and Yaniv Leviathan. 2023 · 2023
Later among the works it cites.
p+: Extended textual conditioning in text-to-image generation
Andrey Voynov, Qinghao Chu, Daniel Cohen-Or, and Kfir Aberman. 2023 · 2023
Later among the works it cites.
A Latent Space of Stochastic Diffusion Models for Zero-Shot Image Editing and Guidance. In ICCV
Chen Henry Wu and Fernando De la Torre. 2023 · 2023
Later among the works it cites.
CCM: Adding Conditional Controls to Text-to-Image Consistency Models
Jie Xiao, Kai Zhu, Han Zhang, Zhiheng Liu, Yujun Shen, Yu Liu, Xueyang Fu, and Zheng-Jun Zha. 2023 · 2023
Later among the works it cites.
ProSpect: Prompt Spectrum for Attribute-Aware Personalization of Diffusion Models
Yuxin Zhang, Weiming Dong, Fan Tang, Nisha Huang, Haibin Huang, Chongyang Ma, Tong-Yee Lee, Oliver Deussen, and Changsheng Xu. 2023 · 2023
Later among the works it cites.
LCM-Lookahead for Encoder-based Text-to-Image Personalization
Rinon Gal, Or Lichter, Elad Richardson, Or Patashnik, Amit H. Bermano, Gal Chechik, and Daniel Cohen-Or. 2024 · 2024
Closest in time.
ReNoise: Real Image Inversion Through Iterative Noising
Daniel Garibi, Or Patashnik, Andrey Voynov, Hadar Averbuch-Elor, and Daniel Cohen-Or. 2024 · 2024
Closest in time.
PuLID: Pure and Lightning ID Customization via Contrastive Alignment
Zinan Guo, Yanze Wu, Zhuowei Chen, Lang Chen, and Qian He. 2024 · 2024
Closest in time.
Consistency Trajectory Models: Learning Probability Flow ODE Trajectory of Diffusion. In The Twelfth International Conference on Learning Representations
Dongjun Kim, Chieh-Hsin Lai, Wei-Hsiang Liao, Naoki Murata, Yuhta Takida, Toshimitsu Uesaka, Yutong He, Yuki Mitsufuji, and Stefano Ermon. 2024 · 2024
Closest in time.
Posterior distillation sampling. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 13352–13361
Juil Koo, Chanho Park, and Minhyuk Sung. 2024 · 2024
Closest in time.
SDXL-Lightning: Progressive Adversarial Diffusion Distillation
Shanchuan Lin, Anran Wang, and Xiao Yang. 2024 · 2024
Closest in time.
One-Step Image Translation with Text-to-Image Models
Gaurav Parmar, Taesung Park, Srinivasa Narasimhan, and Jun-Yan Zhu. 2024 · 2024
Closest in time.
SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis. In The Twelfth International Conference on Learning Representations
Dustin Podell, Zion English, Kyle Lacey, Andreas Blattmann, Tim Dockhorn, Jonas Müller, Joe Penna, and Robin Rombach. 2024 · 2024
Closest in time.
Training-Free Consistent Text-to-Image Generation
Yoad Tewel, Omri Kaduri, Rinon Gal, Yoni Kasten, Lior Wolf, Gal Chechik, and Yuval Atzmon. 2024 · 2024
Closest in time.
One-step Diffusion with Distribution Matching Distillation
Tianwei Yin, Michaël Gharbi, Richard Zhang, Eli Shechtman, Frédo Durand, William T Freeman, and Taesung Park. 2024 · 2024
Closest in time.