Fetching the paper…
Reading the bibliography…
While diffusion-based text-to-image (T2I) models provide a simple and powerful way to generate images, guiding this generation remains a challenge.
Personal causation: The international affective determinants of behavior
Richard De Charms. 1970 · 1970
Earlier work this paper cites.
What do children learn when they paint?
Elliot W Eisner. 1978 · 1978
Earlier work this paper cites.
Direct Manipulation vs. Interface Agents
Ben Shneiderman and Pattie Maes. 1997 · 1997
Earlier work this paper cites.
Design Principles for Tools to Support Creative Thinking
Mitchel Resnick, Brad Myers, Kumiyo Nakakoji, Ben Shneiderman, Randy Pausch, Ted Selker, and Mike Eisenberg. 2005 · 2005
Earlier work this paper cites.
GeDi: Generative Discriminator Guided Sequence Generation
Ben Krause, Akhilesh Deepak Gotmare, Bryan McCann, Nitish Shirish Keskar, Shafiq Joty, Richard Socher, and Nazneen Fatema Rajani. 2020 · 2009
Earlier work this paper cites.
ICanDraw: Using Sketch Recognition and Corrective Feedback to Assist a User in Drawing Human Faces. In Proceedings of the SIGCHI Conference on Human Factors in Computing Systems (Atlanta, Georgia, USA) (CHI ’10) . Association for Computing Machinery, New York, NY, USA, 897–906
Daniel Dixon, Manoj Prasad, and Tracy Hammond. 2010 · 2010
Earlier work this paper cites.
The Art Teacher’s Book of Lists
H.D. Hume. 2010 · 2010
Earlier work this paper cites.
Nonlinear Revision Control for Images
Hsiang-Ting Chen, Li-Yi Wei, and Chun-Fa Chang. 2011 · 2011
Earlier work this paper cites.
Sketch-Sketch Revolution: An Engaging Tutorial System for Guided Sketching and Application Learning. In Proceedings of the 24th Annual ACM Symposium on User Interface Software and Technology (Santa Barbara, California, USA) (UIST ’11) . Association for Computing Machinery, New York, NY, USA, 373–382
Jennifer Fernquist, Tovi Grossman, and George Fitzmaurice. 2011 · 2011
Earlier work this paper cites.
The Drawing Assistant: Automated Drawing Guidance and Feedback from Photographs. In Proceedings of the 26th Annual ACM Symposium on User Interface Software and Technology (St. Andrews, Scotland, United Kingdom) (UIST ’13) . Association for Computing Machinery, New York, NY, USA, 183–192
Emmanuel Iarussi, Adrien Bousseau, and Theophanis Tsandilas. 2013 · 2013
Earlier work this paper cites.
Real-Time Drawing Assistance through Crowdsourcing
Alex Limpaecher, Nicolas Feltman, Adrien Treuille, and Michael Cohen. 2013 · 2013
Earlier work this paper cites.
Painting with Bob: Assisted Creativity for Novices. In Proceedings of the 27th Annual ACM Symposium on User Interface Software and Technology (Honolulu, Hawaii, USA) (UIST ’14) . Association for Computing Machinery, New York, NY, USA, 419–428
Luca Benedetti, Holger Winnemöller, Massimiliano Corsini, and Roberto Scopigno. 2014 · 2014
Earlier work this paper cites.
Quantifying the Creativity Support of Digital Tools through the Creativity Support Index
Erin Cherry and Celine Latulipe. 2014 · 2014
Earlier work this paper cites.
Generative Adversarial Nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio. 2014 · 2014
Earlier work this paper cites.
EZ-Sketching: Three-Level Optimization for Error-Tolerant Image Tracing
Qingkun Su, Wing Ho Andy Li, Jue Wang, and Hongbo Fu. 2014 · 2014
Earlier work this paper cites.
PortraitSketch: Face Sketching Assistance for Novices. In Proceedings of the 27th Annual ACM Symposium on User Interface Software and Technology (Honolulu, Hawaii, USA) (UIST ’14) . Association for Computing Machinery, New York, NY, USA, 407–417
Jun Xie, Aaron Hertzmann, Wilmot Li, and Holger Winnemöller. 2014 · 2014
Earlier work this paper cites.
A Neural Algorithm of Artistic Style
Leon A. Gatys, Alexander S. Ecker, and Matthias Bethge. 2015 · 2015
Earlier work this paper cites.
Empirically Studying Participatory Sense-Making in Abstract Drawing with a Co-Creative Cognitive Agent. In Proceedings of the 21st International Conference on Intelligent User Interfaces (Sonoma, California, USA) (IUI ’16) . Association for Computing Machinery, New York, NY, USA, 196–207
Nicholas Davis, Chih-PIn Hsiao, Kunwar Yashraj Singh, Lisa Li, and Brian Magerko. 2016 · 2016
Earlier work this paper cites.
A learned representation for artistic style
Vincent Dumoulin, Jonathon Shlens, and Manjunath Kudlur. 2016 · 2016
Earlier work this paper cites.
Alec Radford, Luke Metz, and Soumith Chintala. 2016 · 2016
Earlier work this paper cites.
Playful Palette: An Interactive Parametric Color Mixer for Artists
Maria Shugrina, Jingwan Lu, and Stephen Diverdi. 2017 · 2017
Earlier work this paper cites.
StarGAN: Unified Generative Adversarial Networks for Multi-Domain Image-to-Image Translation. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR)
Yunjey Choi, Minje Choi, Munyoung Kim, Jung-Woo Ha, Sunghun Kim, and Jaegul Choo. 2018 · 2018
Earlier work this paper cites.
Extending Manual Drawing Practices with Artist-Centric Programming Tools. In Proceedings of the 2018 CHI Conference on Human Factors in Computing Systems (Montreal QC, Canada) (CHI ’18) . Association for Computing Machinery, New York, NY, USA, 1–13
Jennifer Jacobs, Joel Brandt, Radomír Mech, and Mitchel Resnick. 2018 · 2018
Earlier work this paper cites.
I Lead, You Help but Only with Enough Details: Understanding User Experience of Co-Creation with Artificial Intelligence. In Proceedings of the 2018 CHI Conference on Human Factors in Computing Systems (Montreal QC, Canada) (CHI ’18) . Association for Computing Machinery, New York, NY, USA, 1–13
Changhoon Oh, Jungwoo Song, Jinhan Choi, Seonghyeon Kim, Sungwoo Lee, and Bongwon Suh. 2018 · 2018
Earlier work this paper cites.
Avatar-Net: Multi-scale Zero-Shot Style Transfer by Feature Decoration
Lu Sheng, Ziyi Lin, Jing Shao, and Xiaogang Wang. 2018 · 2018
Earlier work this paper cites.
A Style-Based Generator Architecture for Generative Adversarial Networks. In 2019 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . 4396–4405
Tero Karras, Samuli Laine, and Timo Aila. 2019 · 2019
Earlier work this paper cites.
Arbitrary Style Transfer With Style-Attentional Networks. In 2019 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . 5873–5881
D. Y. Park and K. H. Lee. 2019 · 2019
Earlier work this paper cites.
Painting with CATS: Camera-Aided Texture Synthesis. In Proceedings of the 2019 CHI Conference on Human Factors in Computing Systems (Glasgow, Scotland Uk) (CHI ’19) . Association for Computing Machinery, New York, NY, USA, 1–9
Ticha Sethapakdi and James McCann. 2019 · 2019
Earlier work this paper cites.
Color Builder: A Direct Manipulation Interface for Versatile Color Theme Authoring. In Proceedings of the 2019 CHI Conference on Human Factors in Computing Systems (Glasgow, Scotland Uk) (CHI ’19) . Association for Computing Machinery, New York, NY, USA, 1–12
Maria Shugrina, Wenjia Zhang, Fanny Chevalier, Sanja Fidler, and Karan Singh. 2019 · 2019
Earlier work this paper cites.
DrawMyPhoto: Assisting Novices in Drawing from Photographs. In Proceedings of the 2019 on Creativity and Cognition (San Diego, CA, USA) (C&C ’19) . Association for Computing Machinery, New York, NY, USA, 198–209
Blake Williford, Abhay Doke, Michel Pahud, Ken Hinckley, and Tracy Hammond. 2019 · 2019
Earlier work this paper cites.
Language Models are Few-Shot Learners. In Advances in Neural Information Processing Systems , H. Larochelle, M. Ranzato, R. Hadsell, M. F. Balcan, and H. Lin (Eds.), Vol. 33. Curran Associates, Inc., 1877–1901
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel Ziegler, Jeffrey Wu, Clemens Winter, Chris Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020 · 2020
Cited alongside, same era.
Human-in-the-Loop Differential Subspace Search in High-Dimensional Latent Space
Chia-Hsing Chiu, Yuki Koyama, Yu-Chi Lai, Takeo Igarashi, and Yonghao Yue. 2020 · 2020
Cited alongside, same era.
Denoising Diffusion Probabilistic Models. In Advances in Neural Information Processing Systems , H. Larochelle, M. Ranzato, R. Hadsell, M.F. Balcan, and H. Lin (Eds.), Vol. 33. Curran Associates, Inc., 6840–6851
Jonathan Ho, Ajay Jain, and Pieter Abbeel. 2020 · 2020
Cited alongside, same era.
Sequential Gallery for Interactive Visual Design Optimization
Yuki Koyama, Issei Sato, and Masataka Goto. 2020 · 2020
Imagic: Text-Based Real Image Editing with Diffusion Models
Bahjat Kawar, Shiran Zada, Oran Lang, Omer Tov, Huiwen Chang, Tali Dekel, Inbar Mosseri, and Michal Irani. 2022 · 2022
Later among the works it cites.
Mixplorer: Scaffolding Design Space Exploration through Genetic Recombination of Multiple Peoples’ Designs to Support Novices’ Creativity. In Proceedings of the 2022 CHI Conference on Human Factors in Computing Systems (New Orleans, LA, USA) (CHI ’22) . Association for Computing Machinery, New York, NY, USA, Article 308, 13 pages
Kevin Gonyop Kim, Richard Lee Davis, Alessia Eletta Coppi, Alberto Cattaneo, and Pierre Dillenbourg. 2022b · 2022
Later among the works it cites.
Stylette: Styling the Web with Natural Language. In Proceedings of the 2022 CHI Conference on Human Factors in Computing Systems (New Orleans, LA, USA) (CHI ’22) . Association for Computing Machinery, New York, NY, USA, Article 5, 17 pages
Tae Soo Kim, DaEun Choi, Yoonseo Choi, and Juho Kim. 2022a · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Novice-AI Music Co-Creation via AI-Steering Tools for Deep Generative Models. In Proceedings of the 2020 CHI Conference on Human Factors in Computing Systems (Honolulu, HI, USA) (CHI ’20) . Association for Computing Machinery, New York, NY, USA, 1–13
Ryan Louie, Andy Coenen, Cheng Zhi Huang, Michael Terry, and Carrie J. Cai. 2020 · 2020
Cited alongside, same era.
Swapping Autoencoder for Deep Image Manipulation. In Advances in Neural Information Processing Systems
Taesung Park, Jun-Yan Zhu, Oliver Wang, Jingwan Lu, Eli Shechtman, Alexei A. Efros, and Richard Zhang. 2020 · 2020
Cited alongside, same era.
HyperStyle: StyleGAN Inversion with HyperNetworks for Real Image Editing
Yuval Alaluf, Omer Tov, Ron Mokady, Rinon Gal, and Amit H. Bermano. 2021 · 2021
Cited alongside, same era.
Evaluating Large Language Models Trained on Code
Mark Chen, Jerry Tworek, Heewoo Jun, Qiming Yuan, Henrique Ponde de Oliveira Pinto, Jared Kaplan, Harri Edwards, Yuri Burda, Nicholas Joseph, Greg Brockman, Alex Ray, Raul Puri, Gretchen Krueger, Michael Petrov, Heidy Khlaaf, Girish Sastry, Pamela Mishkin, Brooke Chan, Scott Gray, Nick Ryder, Mikhail Pavlov, Alethea Power, Lukasz Kaiser, Mohammad Bavarian, Clemens Winter, Philippe Tillet, Felipe Petroski Such, Dave Cummings, Matthias Plappert, Fotios Chantzis, Elizabeth Barnes, Ariel Herbert-Voss, William Hebgen Guss, Alex Nichol, Alex Paino, Nikolas Tezak, Jie Tang, Igor Babuschkin, Suchir Balaji, Shantanu Jain, William Saunders, Christopher Hesse, Andrew N. Carr, Jan Leike, Josh Achiam, Vedant Misra, Evan Morikawa, Alec Radford, Matthew Knight, Miles Brundage, Mira Murati, Katie Mayer, Peter Welinder, Bob McGrew, Dario Amodei, Sam McCandlish, Ilya Sutskever, and Wojciech Zaremba. 2021 · 2021
Cited alongside, same era.
Diffusion Models Beat GANs on Image Synthesis. In Advances in Neural Information Processing Systems , M. Ranzato, A. Beygelzimer, Y. Dauphin, P.S. Liang, and J. Wortman Vaughan (Eds.), Vol. 34. Curran Associates, Inc., 8780–8794
Prafulla Dhariwal and Alexander Nichol. 2021 · 2021
Cited alongside, same era.
Zero-Shot Text-Guided Object Generation with Dream Fields
Ajay Jain, Ben Mildenhall, Jonathan T. Barron, Pieter Abbeel, and Ben Poole. 2021 · 2021
Cited alongside, same era.
CLIPstyler: Image Style Transfer with a Single Text Condition
Gihyun Kwon and Jong Chul Ye. 2021 · 2021
Cited alongside, same era.
Glide: Towards photorealistic image generation and editing with text-guided diffusion models
Alex Nichol, Prafulla Dhariwal, Aditya Ramesh, Pranav Shyam, Pamela Mishkin, Bob McGrew, Ilya Sutskever, and Mark Chen. 2021 · 2021
Cited alongside, same era.
Jun Hao Liew, Hanshu Yan, Daquan Zhou, and Jiashi Feng. 2022 · 2022
Later among the works it cites.
Design Guidelines for Prompt Engineering Text-to-Image Generative Models. In Proceedings of the 2022 CHI Conference on Human Factors in Computing Systems (New Orleans, LA, USA) (CHI ’22) . Association for Computing Machinery, New York, NY, USA, Article 384, 23 pages
Vivian Liu and Lydia B Chilton. 2022 · 2022
Later among the works it cites.
SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations. In International Conference on Learning Representations
Chenlin Meng, Yutong He, Yang Song, Jiaming Song, Jiajun Wu, Jun-Yan Zhu, and Stefano Ermon. 2022 · 2022
Later among the works it cites.
DreamFusion: Text-to-3D using 2D Diffusion
Ben Poole, Ajay Jain, Jonathan T. Barron, and Ben Mildenhall. 2022 · 2022
Later among the works it cites.
Hierarchical Text-Conditional Image Generation with CLIP Latents
Aditya Ramesh, Prafulla Dhariwal, Alex Nichol, Casey Chu, and Mark Chen. 2022 · 2022
Later among the works it cites.
Palette: Image-to-Image Diffusion Models. In ACM SIGGRAPH 2022 Conference Proceedings (Vancouver, BC, Canada) (SIGGRAPH ’22) . Association for Computing Machinery, New York, NY, USA, Article 15, 10 pages
Chitwan Saharia, William Chan, Huiwen Chang, Chris Lee, Jonathan Ho, Tim Salimans, David Fleet, and Mohammad Norouzi. 2022a · 2022
Later among the works it cites.
Make-A-Video: Text-to-Video Generation without Text-Video Data
Uriel Singer, Adam Polyak, Thomas Hayes, Xi Yin, Jie An, Songyang Zhang, Qiyuan Hu, Harry Yang, Oron Ashual, Oran Gafni, Devi Parikh, Sonal Gupta, and Yaniv Taigman. 2022 · 2022
Later among the works it cites.
Diffusion Art or Digital Forgery? Investigating Data Replication in Diffusion Models
Gowthami Somepalli, Vasu Singla, Micah Goldblum, Jonas Geiping, and Tom Goldstein. 2022 · 2022
Later among the works it cites.
Interactive and Visual Prompt Engineering for Ad-hoc Task Adaptation With Large Language Models
Hendrik Strobelt, Albert Webson, Victor Sanh, Benjamin Hoover, Johanna Beyer, Hanspeter Pfister, and Alexander M. Rush. 2022 · 2022
Later among the works it cites.
UniTune: Text-Driven Image Editing by Fine Tuning an Image Generation Model on a Single Image
Dani Valevski, Matan Kalman, Yossi Matias, and Yaniv Leviathan. 2022 · 2022
Later among the works it cites.
Phenaki: Variable Length Video Generation From Open Domain Textual Description
Ruben Villegas, Mohammad Babaeizadeh, Pieter-Jan Kindermans, Hernan Moraldo, Han Zhang, Mohammad Taghi Saffar, Santiago Castro, Julius Kunze, and Dumitru Erhan. 2022 · 2022
Later among the works it cites.
FlatMagic: Improving Flat Colorization through AI-driven Design for DigitalComic Professionals
Chuan Yan, John Joon Young Chung, Kiheon Yoon, Yotam Gingold, Eytan Adar, and Sungsoo Ray Hong. 2022 · 2022
Later among the works it cites.
Adobe Firefly (Beta)
Adobe. 2023 · 2023
Closest in time.
MusicLM: Generating Music From Text
Andrea Agostinelli, Timo I. Denk, Zalán Borsos, Jesse Engel, Mauro Verzetti, Antoine Caillon, Qingqing Huang, Aren Jansen, Adam Roberts, Marco Tagliasacchi, Matt Sharifi, Neil Zeghidour, and Christian Frank. 2023 · 2023
Closest in time.
ELEMENTS OF ART and PRINCIPLES OF DESIGN
Atlee Arts. 2023 · 2023
Closest in time.
Stable-diffusion-webui: Features
AUTOMATIC1111. 2023a · 2023
Closest in time.
Stable-diffusion-webui: Negative prompt
AUTOMATIC1111. 2023b · 2023
Closest in time.
Artinter: AI-powered Boundary Objects for Commissioning Visual Arts. In Designing Interactive Systems Conference (Pittsburgh, PA) (DIS ’23) . Association for Computing Machinery, New York, NY, USA
John Joon Young Chung, , and Eytan Adar. 2023 · 2023
Closest in time.
Colaboratory
Colaboratory. 2023 · 2023
Closest in time.
Composer: Creative and Controllable Image Synthesis with Composable Conditions
Lianghua Huang, Di Chen, Yu Liu, Yujun Shen, Deli Zhao, and Jingren Zhou. 2023 · 2023
Closest in time.
GLIGEN: Open-Set Grounded Text-to-Image Generation
Yuheng Li, Haotian Liu, Qingyang Wu, Fangzhou Mu, Jianwei Yang, Jianfeng Gao, Chunyuan Li, and Yong Jae Lee. 2023 · 2023
Closest in time.
Midjourney
Midjourney. 2023 · 2023
Closest in time.
Zero-shot Image-to-Image Translation
Gaurav Parmar, Krishna Kumar Singh, Richard Zhang, Yijun Li, Jingwan Lu, and Jun-Yan Zhu. 2023 · 2023
Closest in time.
DreamStudio
Stability.ai. 2023 · 2023
Closest in time.
RePrompt: Automatic Prompt Editing to Refine AI-Generative Art Towards Precise Expressions
Yunlong Wang, Shuyuan Shen, and Brian Y Lim. 2023 · 2023
Closest in time.
Adding Conditional Control to Text-to-Image Diffusion Models
Lvmin Zhang and Maneesh Agrawala. 2023 · 2023
Closest in time.