Fetching the paper…
Reading the bibliography…
Image manipulation under the guidance of textual descriptions has recently received a broad range of attention.
Denoising diffusion probabilistic models
Jonathan Ho, Ajay Jain, and Pieter Abbeel. 2020 · 2006
Earlier work this paper cites.
Adversarial score matching and improved sampling for image generation. In International Conference on Learning Representations . ICLR
Alexia Jolicoeur-Martineau, Rémi Piché-Taillefer, Ioannis Mitliagkas, and Remi Tachet des Combes. 2021 · 2009
Earlier work this paper cites.
Denoising Diffusion Implicit Models. In International Conference on Learning Representations . ICLR
Jiaming Song, Chenlin Meng, and Stefano Ermon. 2021 · 2010
Earlier work this paper cites.
Auto-encoding variational bayes
Diederik P Kingma and Max Welling. 2013 · 2013
Earlier work this paper cites.
Microsoft COCO: Common Objects in Context. In Computer Vision – ECCV 2014 . Springer International Publishing, 740–755
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C. Lawrence Zitnick. 2014 · 2014
Earlier work this paper cites.
Generating images from captions with attention
Elman Mansimov, Emilio Parisotto, Jimmy Lei Ba, and Ruslan Salakhutdinov. 2015 · 2015
Earlier work this paper cites.
Deep unsupervised learning using nonequilibrium thermodynamics. In International Conference on Machine Learning . PMLR, ICML, 2256–2265
Jascha Sohl-Dickstein, Eric Weiss, Niru Maheswaranathan, and Surya Ganguli. 2015 · 2015
Earlier work this paper cites.
Semantic Image Synthesis via Adversarial Learning. In 2017 IEEE International Conference on Computer Vision (ICCV) . IEEE
Hao Dong, Simiao Yu, Chao Wu, and Yike Guo. 2017 · 2017
Earlier work this paper cites.
Arbitrary Style Transfer in Real-Time with Adaptive Instance Normalization. In 2017 IEEE International Conference on Computer Vision (ICCV) . IEEE
Xun Huang and Serge Belongie. 2017 · 2017
Earlier work this paper cites.
Invited Talk: U-Net Convolutional Networks for Biomedical Image Segmentation. In Informatik aktuell . Springer Berlin Heidelberg, 3–3
Olaf Ronneberger. 2017 · 2017
Earlier work this paper cites.
Large scale GAN training for high fidelity natural image synthesis
Andrew Brock, Jeff Donahue, and Karen Simonyan. 2018 · 2018
Earlier work this paper cites.
Feature-wise transformations
Vincent Dumoulin, Ethan Perez, Nathan Schucher, Florian Strub, Harm Vries, Aaron Courville, and Yoshua Bengio. 2018 · 2018
Earlier work this paper cites.
Text-adaptive generative adversarial networks: manipulating images with natural language
Seonghyeon Nam, Yunji Kim, and Seon Joo Kim. 2018 · 2018
Earlier work this paper cites.
The unreasonable effectiveness of deep features as a perceptual metric. In IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . 586–595
Richard Zhang, Phillip Isola, Alexei A Efros, Eli Shechtman, and Oliver Wang. 2018 · 2018
Earlier work this paper cites.
Instagan: Instance-aware image-to-image translation. In International Conference on Learning Representations . ICLR
Sangwoo Mo, Minsu Cho, and Jinwoo Shin. 2019 · 2019
Earlier work this paper cites.
DoveNet: Deep Image Harmonization via Domain Verification. In 2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE
Wenyan Cong, Jianfu Zhang, Li Niu, Liu Liu, Zhixin Ling, Weiyuan Li, and Liqing Zhang. 2020 · 2020
Earlier work this paper cites.
Analyzing and Improving the Image Quality of StyleGAN. In 2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE
Tero Karras, Samuli Laine, Miika Aittala, Janne Hellsten, Jaakko Lehtinen, and Timo Aila. 2020 · 2020
Cited alongside, same era.
Simplified Fréchet Distance for Generative Adversarial Nets
Chung-Il Kim, Meejoung Kim, Seungwon Jung, and Eenjun Hwang. 2020 · 2020
Cited alongside, same era.
ManiGAN: Text-Guided Image Manipulation. In 2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE
Bowen Li, Xiaojuan Qi, Thomas Lukasiewicz, and Philip H.S. Torr. 2020 · 2020
Cited alongside, same era.
Open-Edit: Open-Domain Image Manipulation with Open-Vocabulary Instructions. In Computer Vision – ECCV 2020 . Springer International Publishing, 89–106
Xihui Liu, Zhe Lin, Jianming Zhang, Handong Zhao, Quan Tran, Xiaogang Wang, and Hongsheng Li. 2020 · 2020
Cited alongside, same era.
Blended Diffusion for Text-driven Editing of Natural Images. In 2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE
Omri Avrahami, Dani Lischinski, and Ohad Fried. 2022b · 2022
Later among the works it cites.
Text2live: Text-driven layered image and video editing. In European Conference on Computer Vision . Springer, 707–723
Omer Bar-Tal, Dolev Ofri-Amar, Rafail Fridman, Yoni Kasten, and Tali Dekel. 2022 · 2022
Later among the works it cites.
FlexIT: Towards Flexible Semantic Image Translation. In 2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE
Guillaume Couairon, Asya Grechka, Jakob Verbeek, Holger Schwenk, and Matthieu Cord. 2022 · 2022
Later among the works it cites.
CLIP guided diffusion HQ 256x256
Katherine Crowson. 2022 · 2022
Later among the works it cites.
StyleGAN-NADA: CLIP-guided domain adaptation of image generators
Rinon Gal, Or Patashnik, Haggai Maron, Amit H. Bermano, Gal Chechik, and Daniel Cohen-Or. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Dmitry Baranchuk, Ivan Rubachev, Andrey Voynov, Valentin Khrulkov, and Artem Babenko. 2021 · 2021
Cited alongside, same era.
David Bau, Alex Andonian, Audrey Cui, YeonHwan Park, Ali Jahanian, Aude Oliva, and Antonio Torralba. 2021 · 2021
Cited alongside, same era.
ILVR: Conditioning Method for Denoising Diffusion Probabilistic Models. In 2021 IEEE/CVF International Conference on Computer Vision (ICCV) . IEEE
Jooyoung Choi, Sungwon Kim, Yonghyun Jeong, Youngjune Gwon, and Sungroh Yoon. 2021 · 2021
Cited alongside, same era.
Diffusion Models Beat GANs on Image Synthesis. In Advances Neural Information Processing Systems . NeurIPS, 8780–8794
Prafulla Dhariwal and Alex Nichol. 2021 · 2021
Cited alongside, same era.
Taming Transformers for High-Resolution Image Synthesis. In 2021 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE
Patrick Esser, Robin Rombach, and Bjorn Ommer. 2021 · 2021
Cited alongside, same era.
Classifier-Free Diffusion Guidance. In NeurIPS 2021 Workshop on Deep Generative Models and Downstream Applications
Jonathan Ho and Tim Salimans. 2021 · 2021
Cited alongside, same era.
Language-Guided Global Image Editing via Cross-Modal Cyclic Mechanism. In 2021 IEEE/CVF International Conference on Computer Vision (ICCV) . IEEE
Wentao Jiang, Ning Xu, Jiayun Wang, Chen Gao, Jing Shi, Zhe Lin, and Si Liu. 2021 · 2021
Cited alongside, same era.
Glide: Towards photorealistic image generation and editing with text-guided diffusion models
Alex Nichol, Prafulla Dhariwal, Aditya Ramesh, Pranav Shyam, Pamela Mishkin, Bob McGrew, Ilya Sutskever, and Mark Chen. 2021 · 2021
Cited alongside, same era.
Draw Your Art Dream: Diverse Digital Art Synthesis with Multimodal Guided Diffusion. In Proceedings of the 30th ACM International Conference on Multimedia . ACM
Nisha Huang, Fan Tang, Weiming Dong, and Changsheng Xu. 2022 · 2022
Later among the works it cites.
Denoising diffusion restoration models
Bahjat Kawar, Michael Elad, Stefano Ermon, and Jiaming Song. 2022 · 2022
Later among the works it cites.
DiffusionCLIP: Text-Guided Diffusion Models for Robust Image Manipulation. In 2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE
Gwanghyun Kim, Taesung Kwon, and Jong Chul Ye. 2022 · 2022
Later among the works it cites.
CLIPstyler: Image Style Transfer with a Single Text Condition. In 2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE
Gihyun Kwon and Jong Chul Ye. 2022 · 2022
Later among the works it cites.
Image Segmentation Using Text and Image Prompts. In 2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE
Timo Luddecke and Alexander Ecker. 2022 · 2022
Later among the works it cites.
Hierarchical text-conditional image generation with clip latents
Aditya Ramesh, Prafulla Dhariwal, Alex Nichol, Casey Chu, and Mark Chen. 2022 · 2022
Later among the works it cites.
Latent diffusion LAION 400M model text to image inpainting
Robin Rombach. 2022 · 2022
Later among the works it cites.
High-Resolution Image Synthesis with Latent Diffusion Models. In 2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE
Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Bjorn Ommer. 2022 · 2022
Later among the works it cites.
Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding
Chitwan Saharia, William Chan, Saurabh Saxena, Lala Li, Jay Whang, Emily Denton, Seyed Kamyar Seyed Ghasemipour, Burcu Karagol Ayan, S Sara Mahdavi, Rapha Gontijo Lopes, et al · 2022
Later among the works it cites.
ManiTrans: Entity-Level Text-Guided Image Manipulation via Token-wise Semantic Alignment and Generation. In 2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . IEEE
Jianan Wang, Guansong Lu, Hang Xu, Zhenguo Li, Chunjing Xu, and Yanwei Fu. 2022 · 2022
Later among the works it cites.
More Control for Free! Image Synthesis with Semantic Diffusion Guidance
Xihui Liu, Dong Huk Park, Samaneh Azadi, Gong Zhang, Arman Chopikyan, Yuxiao Hu, Humphrey Shi, Anna Rohrbach, and Trevor Darrell. 2023 · 2023
Closest in time.