Fetching the paper…
Reading the bibliography…
Generating and editing images from open domain text prompts is a challenging task that heretofore has required expensive and specially trained models.
“The Cost of Training NLP Models: A Concise Overview”, 2020
Or Sharir, Barak Peleg and Yoav Shoham · 2004
Earlier work this paper cites.
“Imagenet: A large-scale hierarchical image database”
Jia Deng et al · 2009
Earlier work this paper cites.
“Adam: A Method for Stochastic Optimization”, 2014
Diederik Kingma and Jimmy Ba · 2014
Earlier work this paper cites.
Karen Simonyan, Andrea Vedaldi and Andrew Zisserman · 2014
Earlier work this paper cites.
“DeepDream - a code example for visualizing Neural Networks”, Google AI Blog, 2015
Alexander Mordvintsev, Christopher Olah and Mike Tyka · 2015
Earlier work this paper cites.
“Understanding Neural Networks Through Deep Visualization”, 2015
Jason Yosinski et al · 2015
Earlier work this paper cites.
“Semantic Image Synthesis via Adversarial Learning”
Hao Dong, Simiao Yu, Chao Wu and Yike Guo · 2017
Earlier work this paper cites.
“Neural Discrete Representation Learning”
Aäron van Oord, Oriol Vinyals and Koray Kavukcuoglu · 2017
Earlier work this paper cites.
“Grad-CAM: Visual Explanations from Deep Networks via Gradient-based Localization”
Ramprasaath Selvaraju et al · 2017
Earlier work this paper cites.
“Text-Adaptive Generative Adversarial Networks: Manipulating Images with Natural Language”
Seonghyeon Nam, Yunji Kim and Seon Kim · 2018
Earlier work this paper cites.
“Parameter-efficient transfer learning for NLP”
Neil Houlsby et al · 2019
Earlier work this paper cites.
“Rewriting a deep generative model”
David Bau et al · 2020
Earlier work this paper cites.
“Manigan: Text-guided image manipulation”
Bowen Li, Xiaojuan Qi, Thomas Lukasiewicz and Philip Torr · 2020
Earlier work this paper cites.
“Open-edit: Open-domain image manipulation with open-vocabulary instructions”
Xihui Liu et al · 2020
Earlier work this paper cites.
“SESAME: semantic editing of scenes by adding, manipulating or erasing objects”
Evangelos Ntavelis et al · 2020
Earlier work this paper cites.
“Kornia: an Open Source Differentiable Computer Vision Library for PyTorch”
Edgar Riba et al · 2020
Earlier work this paper cites.
“Semantic Pyramid for Image Generation”
Assaf Shocher et al · 2020
Earlier work this paper cites.
“Alien Dreams: An Emerging Art Scene”, Machine Learning @ Berkeley, 2020
Charlie Snell · 2020
Earlier work this paper cites.
“Telling Creative Stories Using Generative Visual Aids”, 2021
Safinah Ali and Devi Parikh · 2021
Cited alongside, same era.
“Blended Diffusion for Text-driven Editing of Natural Images”, 2021
Omri Avrahami, Dani Lischinski and Ohad Fried · 2021
Cited alongside, same era.
“diffvg+CLIP: Generating Painting Trajectories from Text”
Gerry Chen, Alice Dumay and Mengyi Tang · 2021
Cited alongside, same era.
“Editing Factual Knowledge in Language Models”, 2021
Nicola De, Wilker Aziz and Ivan Titov · 2021
Cited alongside, same era.
“MAGMA – Multimodal Augmentation of Generative Models through Adapter-based Finetuning”, 2021
Constantin Eichenberg et al · 2021
“Fast Model Editing at Scale”, 2021
Eric Mitchell et al · 2021
Later among the works it cites.
“GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models”, 2021
Alex Nichol et al · 2021
Later among the works it cites.
“Learning Transferable Visual Models From Natural Language Supervision”
Alec Radford et al · 2021
Later among the works it cites.
“Zero-Shot Text-to-Image Generation”
Aditya Ramesh et al · 2021
Later among the works it cites.
“The Dawn of the Human-Machine Era: A forecast of new and emerging language technologies.”, 2021
Dave Sayers et al · 2021
Later among the works it cites.
“Insiders and Outsiders in Research on Machine Learning and Society”
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
“Taming Transformers for High-Resolution Image Synthesis”
Patrick Esser, Robin Rombach and Björn Ommer · 2021
Cited alongside, same era.
“WenLan 2.0: Make AI Imagine via a Multimodal Foundation Model”, 2021
Nanyi Fei et al · 2021
Cited alongside, same era.
“CLIPDraw: Exploring Text-to-Drawing Synthesis through Language-Image Encoders”, 2021
Kevin Frans, LB Soros and Olaf Witkowski · 2021
Cited alongside, same era.
“AffectGAN: Affect-Based Generative Art Driven by Semantics”
Theodoros Galanos, Antonios Liapis and Georgios Yannakakis · 2021
Cited alongside, same era.
“Vector Quantized Diffusion Model for Text-to-Image Synthesis”, 2021
Shuyang Gu et al · 2021
Cited alongside, same era.
“MUSE: Textual Attributes Guided Portrait Painting Generation”
Xiaodan Hu et al · 2021
Cited alongside, same era.
“minDALL-E on Conceptual Captions”, 2021
Saehoon Kim et al · 2021
Cited alongside, same era.
Yu Tao and Kush Varshney · 2021
Later among the works it cites.
“Modern Evolution Strategies for Creativity: Fitting Concrete Images and Abstract Concepts”, 2021
Yingtao Tian and David Ha · 2021
Later among the works it cites.
“Multimodal Few-Shot Learning with Frozen Language Models”
Maria Tsimpoukelli et al · 2021
Later among the works it cites.
“Mapping the latent spaces of culture”, 2021
Ted Underwood · 2021
Later among the works it cites.
“Wav2CLIP: Learning Robust Audio Representations From CLIP”, 2021
Ho-Hsiang Wu, Prem Seetharaman, Kundan Kumar and Juan Bello · 2021
Later among the works it cites.
“Words to Matter: De novo
Zhenze Yang and Markus. Buehler · 2021
Later among the works it cites.
“The values encoded in machine learning research”
Abeba Birhane et al · 2022
Closest in time.
“GPT-NeoX-20B: An Open-Source Autoregressive Language Model”
Sid Black et al · 2022
Closest in time.
“FlexIT: Towards Flexible Semantic Image Translation”, 2022
Guillaume Couairon et al · 2022
Closest in time.
“Music2Video: Automatic Generation of Music Video with fusion of audio and text”, 2022
Joel Jang, Sumin Shin and Yoonjeon Kim · 2022
Closest in time.
“Studying up machine learning data: Why talk about bias when we mean power?”
Milagros Miceli, Julian Posada and Tianling Yang · 2022
Closest in time.
“CLIP-GEN: Language-Free Training of a Text-to-Image Generator with CLIP”, 2022
Zihao Wang et al · 2022
Closest in time.
“StyleCLIP: Text-Driven Manipulation of StyleGAN Imagery”
Or Patashnik et al · 2094
Closest in time.