Fetching the paper…
Reading the bibliography…
To achieve disentangled image manipulation, previous works depend heavily on manual annotation.
Deep learning face attributes in the wild
Ziwei Liu, Ping Luo, Xiaogang Wang, and Xiaoou Tang · 2015
Earlier work this paper cites.
Infogan: Interpretable representation learning by information maximizing generative adversarial nets
Xi Chen, Yan Duan, Rein Houthooft, John Schulman, Ilya Sutskever, and Pieter Abbeel · 2016
Earlier work this paper cites.
Semantic image synthesis via adversarial learning
Hao Dong, Simiao Yu, Chao Wu, and Yike Guo · 2017
Earlier work this paper cites.
Progressive growing of gans for improved quality, stability, and variation
Tero Karras, Timo Aila, Samuli Laine, and Jaakko Lehtinen · 2017
Earlier work this paper cites.
Isolating sources of disentanglement in variational autoencoders
Ricky T. Q. Chen, Xuechen Li, Roger B Grosse, and David K Duvenaud · 2018
Earlier work this paper cites.
Disentangling by factorising
Hyunjik Kim and Andriy Mnih · 2018
Earlier work this paper cites.
Text-adaptive generative adversarial networks: Manipulating images with natural language
Seonghyeon Nam, Yunji Kim, and Seon Joo Kim · 2018
Earlier work this paper cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Earlier work this paper cites.
A style-based generator architecture for generative adversarial networks
Tero Karras, Samuli Laine, and Timo Aila · 2019
Earlier work this paper cites.
Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks
Jiasen Lu, Dhruv Batra, Devi Parikh, and Stefan Lee · 2019
Earlier work this paper cites.
LXMERT: Learning cross-modality encoder representations from transformers
Hao Tan and Mohit Bansal · 2019
Earlier work this paper cites.
Demystifying inter-class disentanglement
Aviv Gabbay and Yedid Hoshen · 2020
Earlier work this paper cites.
Ganspace: Discovering interpretable gan controls
Erik Härkönen, Aaron Hertzmann, Jaakko Lehtinen, and Sylvain Paris · 2020
Earlier work this paper cites.
On the ”steerability” of generative adversarial networks
Ali Jahanian, Lucy Chai, and Phillip Isola · 2020
Earlier work this paper cites.
Analyzing and improving the image quality of stylegan
Tero Karras, Samuli Laine, Miika Aittala, Janne Hellsten, Jaakko Lehtinen, and Timo Aila · 2020
Earlier work this paper cites.
Manigan: Text-guided image manipulation
Bowen Li, Xiaojuan Qi, Thomas Lukasiewicz, and Philip HS Torr · 2020
Earlier work this paper cites.
Lightweight generative adversarial networks for text-guided image manipulation
Bowen Li, Xiaojuan Qi, Philip Torr, and Thomas Lukasiewicz · 2020
Earlier work this paper cites.
Oscar: Object-semantics aligned pre-training for vision-language tasks
Xiujun Li, Xi Yin, Chunyuan Li, Xiaowei Hu, Pengchuan Zhang, Lei Zhang, Lijuan Wang, Houdong Hu, Li Dong, Furu Wei, Yejin Choi, and Jianfeng Gao · 2020
Cited alongside, same era.
Describe what to change: A text-guided unsupervised image-to-image translation approach
Yahui Liu, Marco De Nadai, Deng Cai, Huayang Li, Xavier Alameda-Pineda, Nicu Sebe, and Bruno Lepri · 2020
Cited alongside, same era.
Semi-supervised stylegan for disentanglement learning
Weili Nie, Tero Karras, Animesh Garg, Shoubhik Debnath, Anjul Patney, Ankit Patel, and Animashree Anandkumar · 2020
Cited alongside, same era.
Face identity disentanglement via latent space mapping
Yotam Nitzan, Amit Bermano, Yangyan Li, and Daniel Cohen-Or · 2020
Cited alongside, same era.
Interpreting the latent space of gans for semantic face editing
Yujun Shen, Jinjin Gu, Xiaoou Tang, and Bolei Zhou · 2020
Cited alongside, same era.
Blendgan: Implicitly gan blending for arbitrary stylized face generation
Mingcong Liu, Qiang Li, Zekui Qin, Guoxin Zhang, Pengfei Wan, and Wen Zheng · 2021
Closest in time.
Smoothing the disentangled latent style space for unsupervised image-to-image translation
Yahui Liu, Enver Sangineto, Yajing Chen, Linchao Bao, Haoxian Zhang, Nicu Sebe, Bruno Lepri, Wei Wang, and Marco De Nadai · 2021
Closest in time.
Official implementation of StyleCLIP
Or Patashnik and Zongze Wu · 2021
Closest in time.
Styleclip: Text-driven manipulation of stylegan imagery
Or Patashnik, Zongze Wu, Eli Shechtman, Daniel Cohen-Or, and Dani Lischinski · 2021
Closest in time.
Learning transferable visual models from natural language supervision
Alec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh, Gabriel Goh, Sandhini Agarwal, Girish Sastry, Amanda Askell, Pamela Mishkin, Jack Clark, Gretchen Krueger, and Ilya Sutskever · 2021
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Weijie Su, Xizhou Zhu, Yue Cao, Bin Li, Lewei Lu, Furu Wei, and Jifeng Dai · 2020
Cited alongside, same era.
Stylerig: Rigging stylegan for 3d control over portrait images
Ayush Tewari, Mohamed Elgharib, Gaurav Bharaj, Florian Bernard, Hans-Peter Seidel, Patrick Pérez, Michael Zollhofer, and Christian Theobalt · 2020
Cited alongside, same era.
Unsupervised discovery of interpretable directions in the gan latent space
Andrey Voynov and Artem Babenko · 2020
Cited alongside, same era.
In-domain gan inversion for real image editing
Jiapeng Zhu, Yujun Shen, Deli Zhao, and Bolei Zhou · 2020
Cited alongside, same era.
Styleflow: Attribute-conditioned exploration of stylegan-generated images using conditional continuous normalizing flows
Rameen Abdal, Peihao Zhu, Niloy J Mitra, and Peter Wonka · 2021
Cited alongside, same era.
Only a matter of style: Age transformation using a style-based regression model
Yuval Alaluf, Or Patashnik, and Daniel Cohen-Or · 2021
Cited alongside, same era.
Clipdraw: Exploring text-to-drawing synthesis through language-image encoders
Kevin Frans, Lisa B. Soros, and Olaf Witkowski · 2021
Cited alongside, same era.
Aditya Ramesh, Mikhail Pavlov, Gabriel Goh, Scott Gray, Chelsea Voss, Alec Radford, Mark Chen, and Ilya Sutskever · 2021
Closest in time.
Encoding in style: a stylegan encoder for image-to-image translation
Elad Richardson, Yuval Alaluf, Or Patashnik, Yotam Nitzan, Yaniv Azar, Stav Shapiro, and Daniel Cohen-Or · 2021
Closest in time.
Pivotal tuning for latent-based editing of real images
Daniel Roich, Ron Mokady, Amit H Bermano, and Daniel Cohen-Or · 2021
Closest in time.
Closed-form factorization of latent semantics in gans
Yujun Shen and Bolei Zhou · 2021
Closest in time.
Designing an encoder for stylegan image manipulation
Omer Tov, Yuval Alaluf, Yotam Nitzan, Or Patashnik, and Daniel Cohen-Or · 2021
Closest in time.
A geometric analysis of deep generative image models and its applications
Binxu Wang and Carlos R Ponce · 2021
Closest in time.
Stylespace analysis: Disentangled controls for stylegan image generation
Zongze Wu, Dani Lischinski, and Eli Shechtman · 2021
Closest in time.
Tedigan: Text-guided diverse face image generation and manipulation
Weihao Xia, Yujiu Yang, Jing-Hao Xue, and Baoyuan Wu · 2021
Closest in time.
Towards open-world text-guided face image generation and manipulation
Weihao Xia, Yujiu Yang, Jing-Hao Xue, and Baoyuan Wu · 2021
Closest in time.
A latent transformer for disentangled face editing in images and videos
Xu Yao, Alasdair Newson, Yann Gousseau, and Pierre Hellier · 2021
Closest in time.
Vinvl: Making visual representations matter in vision-language models
Pengchuan Zhang, Xiujun Li, Xiaowei Hu, Jianwei Yang, Lei Zhang, Lijuan Wang, Yejin Choi, and Jianfeng Gao · 2021
Closest in time.