Fetching the paper…
Reading the bibliography…
Diffusion-based image translation guided by semantic texts or a single target image has enabled flexible style transfer which is not limited to the specific domains.
Learning hybrid image templates (hit) by information projection
Zhangzhang Si and Song-Chun Zhu · 2011
Earlier work this paper cites.
Score-based generative modeling through stochastic differential equations
Yang Song, Jascha Sohl-Dickstein, Diederik P Kingma, Abhishek Kumar, Stefano Ermon, and Ben Poole · 2011
Earlier work this paper cites.
Image style transfer using convolutional neural networks
Leon A Gatys, Alexander S Ecker, and Matthias Bethge · 2016
Earlier work this paper cites.
Gans trained by a two time-scale update rule converge to a local nash equilibrium
Martin Heusel, Hubert Ramsauer, Thomas Unterthiner, Bernhard Nessler, and Sepp Hochreiter · 2017
Earlier work this paper cites.
Arbitrary style transfer in real-time with adaptive instance normalization
Xun Huang and Serge Belongie · 2017
Earlier work this paper cites.
Image-to-image translation with conditional adversarial networks
Phillip Isola, Jun-Yan Zhu, Tinghui Zhou, and Alexei A Efros · 2017
Earlier work this paper cites.
Scene parsing through ade20k dataset
Bolei Zhou, Hang Zhao, Xavier Puig, Sanja Fidler, Adela Barriuso, and Antonio Torralba · 2017
Earlier work this paper cites.
Unpaired image-to-image translation using cycle-consistent adversarial networks
Jun-Yan Zhu, Taesung Park, Phillip Isola, and Alexei A. Efros · 2017
Earlier work this paper cites.
Combogan: Unrestrained scalability for image domain translation
Asha Anoosheh, Eirikur Agustsson, Radu Timofte, and Luc Van Gool · 2018
Earlier work this paper cites.
Cartoongan: Generative adversarial networks for photo cartoonization
Yang Chen, Yu-Kun Lai, and Yong-Jin Liu · 2018
Earlier work this paper cites.
Stargan: Unified generative adversarial networks for multi-domain image-to-image translation
Yunjey Choi, Minje Choi, Munyoung Kim, Jung-Woo Ha, Sunghun Kim, and Jaegul Choo · 2018
Earlier work this paper cites.
Arcface: Additive angular margin loss for deep face recognition
Jiankang Deng, Jia Guo, Niannan Xue, and Stefanos Zafeiriou · 2019
Earlier work this paper cites.
Style transfer by relaxed optimal transport and self-similarity
Nicholas Kolkin, Jason Salavon, and Gregory Shakhnarovich · 2019
Earlier work this paper cites.
Collagan: Collaborative gan for missing image data imputation
Dongwook Lee, Junyoung Kim, Won-Jin Moon, and Jong Chul Ye · 2019
Earlier work this paper cites.
Arbitrary style transfer with style-attentional networks
Dae Young Park and Kwang Hee Lee · 2019
Earlier work this paper cites.
Photorealistic style transfer via wavelet transforms
Jaejun Yoo, Youngjung Uh, Sanghyuk Chun, Byeongkyu Kang, and Jung-Woo Ha · 2019
Earlier work this paper cites.
An image is worth 16x16 words: Transformers for image recognition at scale
Alexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn, Xiaohua Zhai, Thomas Unterthiner, Mostafa Dehghani, Matthias Minderer, Georg Heigold, Sylvain Gelly, et al · 2020
Earlier work this paper cites.
Denoising diffusion probabilistic models
Jonathan Ho, Ajay Jain, and Pieter Abbeel · 2020
Cited alongside, same era.
Analyzing and improving the image quality of stylegan
Tero Karras, Samuli Laine, Miika Aittala, Janne Hellsten, Jaakko Lehtinen, and Timo Aila · 2020
Cited alongside, same era.
Simplified fréchet distance for generative adversarial nets
Chung-Il Kim, Meejoung Kim, Seungwon Jung, and Eenjun Hwang · 2020
Cited alongside, same era.
Tuigan: Learning versatile image-to-image translation with two unpaired images
Jianxin Lin, Yingxue Pang, Yingce Xia, Zhibo Chen, and Jiebo Luo · 2020
Cited alongside, same era.
Contrastive learning for unpaired image-to-image translation
Taesung Park, Alexei A. Efros, Richard Zhang, and Jun-Yan Zhu · 2020
Cited alongside, same era.
Emerging properties in self-supervised vision transformers
Mathilde Caron, Hugo Touvron, Ishan Misra, Hervé Jégou, Julien Mairal, Piotr Bojanowski, and Armand Joulin · 2021
Styleclip: Text-driven manipulation of stylegan imagery
Or Patashnik, Zongze Wu, Eli Shechtman, Daniel Cohen-Or, and Dani Lischinski · 2021
Later among the works it cites.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, et al · 2021
Later among the works it cites.
Unit-ddpm: Unpaired image translation with denoising diffusion probabilistic models
Hiroshi Sasaki, Chris G Willcocks, and Toby P Breckon · 2021
Later among the works it cites.
Peihao Zhu, Rameen Abdal, John Femiani, and Peter Wonka · 2021
Later among the works it cites.
Blended diffusion for text-driven editing of natural images
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Jojogan: One shot face stylization
Min Jin Chong and David Forsyth · 2021
Cited alongside, same era.
Diffusion models beat gans on image synthesis
Prafulla Dhariwal and Alexander Nichol · 2021
Cited alongside, same era.
Taming transformers for high-resolution image synthesis
Patrick Esser, Robin Rombach, and Bjorn Ommer · 2021
Cited alongside, same era.
Language-driven image style transfer
Tsu-Jui Fu, Xin Eric Wang, and William Yang Wang · 2021
Cited alongside, same era.
Stylegan-nada: Clip-guided domain adaptation of image generators
Rinon Gal, Or Patashnik, Haggai Maron, Gal Chechik, and Daniel Cohen-Or · 2021
Cited alongside, same era.
Plenopticam v1.0: A light-field imaging framework
Christopher Hahne and Amar Aggoun · 2021
Cited alongside, same era.
Omri Avrahami, Dani Lischinski, and Ohad Fried · 2022
Closest in time.
Flexit: Towards flexible semantic image translation
Guillaume Couairon, Asya Grechka, Jakob Verbeek, Holger Schwenk, and Matthieu Cord · 2022
Closest in time.
Clip-guided diffusion
Katherine Crowson · 2022
Closest in time.
Vqgan-clip: Open domain image generation and editing with natural language guidance
Katherine Crowson, Stella Biderman, Daniel Kornis, Dashiell Stander, Eric Hallahan, Louis Castricato, and Edward Raff · 2022
Closest in time.
Drop the gan: In defense of patches nearest neighbors as single image generative models
Niv Granot, Ben Feinstein, Assaf Shocher, Shai Bagon, and Michal Irani · 2022
Closest in time.
Exploring patch-wise semantic relation for contrastive learning in image-to-image translation tasks
Chanyong Jung, Gihyun Kwon, and Jong Chul Ye · 2022
Closest in time.
Diffusionclip: Text-guided diffusion models for robust image manipulation
Gwanghyun Kim, Taesung Kwon, and Jong Chul Ye · 2022
Closest in time.
One-shot adaptation of gan in just one clip
Gihyun Kwon and Jong Chul Ye · 2022
Closest in time.
Image segmentation using text and image prompts
Timo Lüddecke and Alexander Ecker · 2022
Closest in time.
Hierarchical text-conditional image generation with clip latents
Aditya Ramesh, Prafulla Dhariwal, Alex Nichol, Casey Chu, and Mark Chen · 2022
Closest in time.
Palette: Image-to-image diffusion models
Chitwan Saharia, William Chan, Huiwen Chang, Chris Lee, Jonathan Ho, Tim Salimans, David Fleet, and Mohammad Norouzi · 2022
Closest in time.
Splicing vit features for semantic appearance transfer
Narek Tumanyan, Omer Bar-Tal, Shai Bagon, and Tali Dekel · 2022
Closest in time.
Hairclip: Design your hair by text and reference image
Tianyi Wei, Dongdong Chen, Wenbo Zhou, Jing Liao, Zhentao Tan, Lu Yuan, Weiming Zhang, and Nenghai Yu · 2022
Closest in time.