Fetching the paper…
Reading the bibliography…
Sketching serves as a versatile tool for externalizing ideas, enabling rapid exploration and visual communication that spans various disciplines.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel Ziegler, Jeffrey Wu, Clemens Winter, Chris Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei · 1901
Earlier work this paper cites.
Language models are few-shot learners
Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel Ziegler, Jeffrey Wu, Clemens Winter, Chris Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei · 1901
Earlier work this paper cites.
A computational approach to edge detection
John Canny · 1986
Earlier work this paper cites.
The reflective practitioner: How professionals think in action, 1986
Donald A Schon and Vincent DeSanctis · 1986
Earlier work this paper cites.
Serial sketching: visual problem solving in designing
Gabriela Goldschmidt · 1992
Earlier work this paper cites.
Collaborative plans for complex group action
Barbara J. Grosz and Sarit Kraus · 1996
Earlier work this paper cites.
When is a tool a tool? user perceptions of system agency in human–ai co-creative drawing
Tomas Lawton, Kazjon Grace, and Francisco J Ibarrola · 1996
Earlier work this paper cites.
Sketchpad—a man-machine graphical communication system , page 391–408
Ivan E. Sutherland · 1998
Earlier work this paper cites.
Scalable Vector Graphics (SVG) , 1999
World Wide Web Consortium (W3C) · 1999
Earlier work this paper cites.
Graphic thinking for architects and designers
Paul Laseau · 2000
Earlier work this paper cites.
What do sketches say about thinking?
Barbara Tversky · 2002
Earlier work this paper cites.
Sketches for design and design of sketches
Barbara Tversky, Masaki Suwa, Maneesh Agrawala, Julie Heiser, Chris Stolte, Pat Hanrahan, Doantam Phan, Jeff Klingner, Marie-Paule Daniel, Paul Lee, et al · 2003
Earlier work this paper cites.
Foundations of representation: Where might graphical symbol systems come from?
Simon Garrod, Nicolas Fay, John Lee, Jon Oberlander, and Tracy MacLeod · 2007
Earlier work this paper cites.
Ad hoc autonomous agent teams: Collaboration without pre-coordination
Peter Stone, Gal Kaminka, Sarit Kraus, and Jeffrey Rosenschein · 2010
Earlier work this paper cites.
Cogsketch: Sketch understanding for cognitive science research and for education
Kenneth Forbus, Jeffrey Usher, Andrew Lovett, Kate Lockwood, and Jon Wetzel · 2011
Earlier work this paper cites.
Chapter three - psychological research on joint action: Theory and data
Günther Knoblich, Stephen Butterfill, and Natalie Sebanz · 2011
Earlier work this paper cites.
Shadowdraw: real-time user guidance for freehand drawing
Yong Jae Lee, C. Lawrence Zitnick, and Michael F. Cohen · 2011
Earlier work this paper cites.
How do humans sketch objects?
Mathias Eitz, James Hays, and Marc Alexa · 2012
Earlier work this paper cites.
Xdog: An extended difference-of-gaussians compendium including advanced image stylization
Holger Winnemöller, Jan Eric Kyprianidis, and Sven C. Olsen · 2012
Earlier work this paper cites.
Style and abstraction in portrait sketching
Itamar Berger, Ariel Shamir, Moshe Mahler, Elizabeth Carter, and Jessica Hodgins · 2013
Earlier work this paper cites.
Collaborative drawing on a shared digital canvas in elementary science education: The effects of script and task awareness support
Hannie Gijlers, Armin Weinberger, Alieke Mattia van Dijk, Lars Bollen, and Wouter van Joolingen · 2013
Earlier work this paper cites.
Visualizing thought
Barbara Tversky · 2013
Earlier work this paper cites.
Drawing apprentice: An enactive co-creative agent for artistic collaboration
Nicholas Davis, Chih-PIn Hsiao, Kunwar Yashraj Singh, Lisa Li, Sanat Moningi, and Brian Magerko · 2015
Earlier work this paper cites.
Free-hand sketch synthesis with deformable stroke models
Yi Li, Yi-Zhe Song, Timothy M. Hospedales, and Shaogang Gong · 2015
Earlier work this paper cites.
Holistically-nested edge detection
Saining Xie and Zhuowen Tu · 2015
Earlier work this paper cites.
The Quick, Draw! - A.I. Experiment, 2016
Jongejan Jonas, Rowley Henry, Kawashima Takashi, Kim Jongmin, and Fox-Gieg Nick · 2016
Earlier work this paper cites.
The sketchy database: Learning to retrieve badly drawn bunnies
Patsorn Sangkloy, Nathan Burnell, Cusuh Ham, and James Hays · 2016
Earlier work this paper cites.
A neural representation of sketch drawings
David Ha and Douglas Eck · 2017
Earlier work this paper cites.
Synthesizing programs for images using reinforced adversarial learning
Yaroslav Ganin, Tejas D. Kulkarni, Igor Babuschkin, S. M. Ali Eslami, and Oriol Vinyals · 2018
Earlier work this paper cites.
I lead, you help but only with enough details: Understanding user experience of co-creation with artificial intelligence
Changhoon Oh, Jungwoo Song, Jinhan Choi, Seonghyeon Kim, Sungwoo Lee, and Bongwon Suh · 2018
Earlier work this paper cites.
Learning to sketch with shortcut cycle consistency, 2018
Jifei Song, Kaiyue Pang, Yi-Zhe Song, Tao Xiang, and Timothy Hospedales · 2018
Earlier work this paper cites.
Learning to sketch with deep q networks and demonstrated strokes
Tao Zhou, Chen Fang, Zhaowen Wang, Jimei Yang, Byungmoon Kim, Zhili Chen, Jonathan Brandt, and Demetri Terzopoulos · 2018
Earlier work this paper cites.
On the utility of learning about humans for human-ai coordination
Micah Carroll, Rohin Shah, Mark K Ho, Tom Griffiths, Sanjit Seshia, Pieter Abbeel, and Anca Dragan · 2019
Earlier work this paper cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova · 2019
Earlier work this paper cites.
collabdraw: An environment for collaborative sketching with an artificial agent
Judith E. Fan, Monica Dinculescu, and David Ha · 2019
Earlier work this paper cites.
Opensketch: a richly-annotated dataset of product design sketches
Yulia Gryaditskaya, Mark Sypesteyn, Jan Willem Hoftijzer, Sylvia Pont, Frédo Durand, and Adrien Bousseau · 2019
Earlier work this paper cites.
Drawing theories apart: The dispersion of Feynman diagrams in postwar physics
David Kaiser · 2019
Earlier work this paper cites.
Photo-sketching: Inferring contour drawings from images
Mengtian Li, Zhe Lin, Radomir Mech, Ersin Yumer, and Deva Ramanan · 2019
Earlier work this paper cites.
Unsupervised doodling and painting with improved spiral
John FJ Mellor, Eunbyung Park, Yaroslav Ganin, Igor Babuschkin, Tejas Kulkarni, Dan Rosenbaum, Andy Ballard, Theophane Weber, Oriol Vinyals, and SM Eslami · 2019
Earlier work this paper cites.
Observing by hand: sketching the nebulae in the nineteenth century
Omar W Nasim · 2019
Earlier work this paper cites.
Apdrawinggan: Generating artistic portrait drawings from face photos with hierarchical gans
Ran Yi, Yong-Jin Liu, Yu-Kun Lai, and Paul L Rosin · 2019
Cited alongside, same era.
Pixelor: a competitive sketching ai agent. so you think you can sketch?
Ayan Kumar Bhunia, Ayan Das, Umar Riaz Muhammad, Yongxin Yang, Timothy M. Hospedales, Tao Xiang, Yulia Gryaditskaya, and Yi-Zhe Song · 2020
Cited alongside, same era.
Deepsvg: A hierarchical generative network for vector graphics animation, 2020
Alexandre Carlier, Martin Danelljan, Alexandre Alahi, and Radu Timofte · 2020
Cited alongside, same era.
Béziersketch: A generative model for scalable vector sketches
Ayan Das, Yongxin Yang, Timothy Hospedales, Tao Xiang, and Yi-Zhe Song · 2020
Cited alongside, same era.
Pragmatic Inference and Visual Abstraction Enable Contextual Flexibility During Visual Communication
Judith E. Fan, Robert D. Hawkins, Mike Wu, and Noah D. Goodman · 2020
Cited alongside, same era.
Vectorfusion: Text-to-svg by abstracting pixel-based diffusion models
Ajay Jain, Amber Xie, and Pieter Abbeel · 2023
Later among the works it cites.
Cairosvg
Kozea · 2023
Later among the works it cites.
Visual instruction tuning
Haotian Liu, Chunyuan Li, Qingyang Wu, and Yong Jae Lee · 2023
Later among the works it cites.
Seva: Leveraging sketches to evaluate alignment between human and machine visual abstraction
Kushin Mukherjee, Holly Huey, Xuanchen Lu, Yael Vinker, Rio Aguina-Kang, Ariel Shamir, and Judith Fan · 2023
Later among the works it cites.
Sdxl: Improving latent diffusion models for high-resolution image synthesis
Dustin Podell, Zion English, Kyle Lacey, A. Blattmann, Tim Dockhorn, Jonas Muller, Joe Penna, and Robin Rombach · 2023
Later among the works it cites.
Visual chain of thought: bridging logical gaps with multimodal infillings
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Creating drawings enhances learning by teaching
Logan Fiorella and Shelbi Kuhlmann · 2020
Cited alongside, same era.
Sketchycoco: Image generation from freehand scene sketches
Chengying Gao, Qi Liu, Qi Xu, Limin Wang, Jianzhuang Liu, and Changqing Zou · 2020
Cited alongside, same era.
Songwei Ge, Vedanuj Goswami, C Lawrence Zitnick, and Devi Parikh · 2020
Cited alongside, same era.
Differentiable vector graphics rasterization for editing and learning
Tzu-Mao Li, Michal Lukáč, Gharbi Michaël, and Jonathan Ragan-Kelley · 2020
Cited alongside, same era.
Sketch-bert: Learning sketch bidirectional encoder representation from transformers by self-supervised learning of sketch gestalt
Hangyu Lin, Yanwei Fu, Yu-Gang Jiang, and X. Xue · 2020
Cited alongside, same era.
Exploring the limits of transfer learning with a unified text-to-text transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu · 2020
Cited alongside, same era.
Sketchformer: Transformer-based representation for sketched structure
Leo Sampaio Ferraz Ribeiro, Tu Bui, John P. Collomosse, and Moacir Antonelli Ponti · 2020
Cited alongside, same era.
Daniel Rose, Vaishnavi Himakunthala, Andy Ouyang, Ryan He, Alex Mei, Yujie Lu, Michael Saxon, Chinmay Sonar, Diba Mirza, and William Yang Wang · 2023
Later among the works it cites.
Frida: A collaborative robot painter with a differentiable, real2sim2real planning environment
Peter Schaldenbrand, James McCann, and Jean Oh · 2023
Later among the works it cites.
Exploring effective factors for improving visual in-context learning
Yanpeng Sun, Qiang Chen, Jian Wang, Jingdong Wang, and Zechao Li · 2023
Later among the works it cites.
Design ideation with ai - sketching, thinking and talking with generative machine learning models
Jakob Tholander and Martin Jonsson · 2023
Later among the works it cites.
Llama: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, Aurelien Rodriguez, Armand Joulin, Edouard Grave, and Guillaume Lample · 2023
Later among the works it cites.
Clipascene: Scene sketching with different types and levels of abstraction
Yael Vinker, Yuval Alaluf, Daniel Cohen-Or, and Ariel Shamir · 2023
Later among the works it cites.
Contextseg: Sketch semantic segmentation by querying the context with attention
Jiawei Wang and Changjian Li · 2023
Later among the works it cites.
Diffsketcher: Text guided vector sketch synthesis through latent diffusion models
XiMing Xing, Chuang Wang, Haitao Zhou, Jing Zhang, Qian Yu, and Dong Xu · 2023
Later among the works it cites.
Flamingo: a visual language model for few-shot learning
Jean-Baptiste Alayrac, Jeff Donahue, Pauline Luc, Antoine Miech, Iain Barr, Yana Hasson, Karel Lenc, Arthur Mensch, Katie Millicah, Malcolm Reynolds, Roman Ring, Eliza Rutherford, Serkan Cabi, Tengda Han, Zhitao Gong, Sina Samangooei, Marianne Monteiro, Jacob Menick, Sebastian Borgeaud, Andrew Brock, Aida Nematzadeh, Sahand Sharifzadeh, Mikolaj Binkowski, Ricardo Barreira, Oriol Vinyals, Andrew Zisserman, and Karen Simonyan · 2024
Closest in time.
Delving into LLMs’ visual understanding ability using SVG to bridge image and text, 2024
Mu Cai, Zeyi Huang, Yuheng Li, Haohan Wang, and Yong Jae Lee · 2024
Closest in time.
Visual chain-of-thought prompting for knowledge-based visual reasoning
Zhenfang Chen, Qinhong Zhou, Yikang Shen, Yining Hong, Zhiqing Sun, Dan Gutfreund, and Chuang Gan · 2024
Closest in time.
3doodle: Compact abstraction of objects with 3d strokes
Changwoon Choi, Jaeah Lee, Jaesik Park, and Young Min Kim · 2024
Closest in time.
Visionllama: A unified llama backbone for vision tasks, 2024
Xiangxiang Chu, Jianlin Su, Bo Zhang, and Chunhua Shen · 2024
Closest in time.
Towards multimodal in-context learning for vision & language models
Sivan Doveh, Shaked Perek, M Jehanzeb Mirza, Wei Lin, Amit Alfassy, Assaf Arbelle, Shimon Ullman, and Leonid Karlinsky · 2024
Closest in time.
Abhimanyu Dubey, Abhinav Jauhri, Abhinav Pandey, et al · 2024
Closest in time.
Clipdraw: exploring text-to-drawing synthesis through language-image encoders
Kevin Frans, L. B. Soros, and Olaf Witkowski · 2024
Closest in time.
Blink: Multimodal large language models can see but not perceive
Xingyu Fu, Yushi Hu, Bangzheng Li, Yu Feng, Haoyu Wang, Xudong Lin, Dan Roth, Noah A Smith, Wei-Chiu Ma, and Ranjay Krishna · 2024
Closest in time.
Breathing life into sketches using text-to-video priors
Rinon Gal, Yael Vinker, Yuval Alaluf, Amit Bermano, Daniel Cohen-Or, Ariel Shamir, and Gal Chechik · 2024
Closest in time.
Hallusionbench: An advanced diagnostic suite for entangled language hallucination and visual illusion in large vision-language models
Tianrui Guan, Fuxiao Liu, Xiyang Wu, Ruiqi Xian, Zongxia Li, Xiaoyu Liu, Xijun Wang, Lichang Chen, Furong Huang, Yaser Yacoob, Dinesh Manocha, and Tianyi Zhou · 2024
Closest in time.
Visual sketchpad: Sketching as a visual chain of thought for multimodal language models
Yushi Hu, Weijia Shi, Xingyu Fu, Dan Roth, Mari Ostendorf, Luke Zettlemoyer, Noah A Smith, and Ranjay Krishna · 2024
Closest in time.
Parallel developmental changes in children’s production and recognition of line drawings of visual concepts
Bria Long, Judith Fan, Holly Huey, Zixian Chai, and Michael Frank · 2024
Closest in time.
Openeqa: Embodied question answering in the era of foundation models
Arjun Majumdar, Anurag Ajay, Xiaohan Zhang, Pranav Putta, Sriram Yenamandra, Mikael Henaff, Sneha Silwal, Paul Mcvay, Oleksandr Maksymets, Sergio Arnaud, Karmesh Yadav, Qiyang Li, Ben Newman, Mohit Sharma, Vincent Berges, Shiqi Zhang, Pulkit Agrawal, Yonatan Bisk, Dhruv Batra, Mrinal Kalakrishnan, Franziska Meier, Chris Paxton, Sasha Sax, and Aravind Rajeswaran · 2024
Closest in time.
Communicating design intent using drawing and text
William P. McCarthy, Justin Matejka, Karl D.D. Willis, Judith E. Fan, and Yewen Pu · 2024
Closest in time.
Compositional chain-of-thought prompting for large multimodal models
Chancharik Mitra, Brandon Huang, Trevor Darrell, and Roei Herzig · 2024
Closest in time.
Gpt-4 technical report, 2024
OpenAI · 2024
Closest in time.
Cofrida: Self-supervised fine-tuning for human-robot co-painting
Peter Schaldenbrand, Gaurav Parmar, Jun-Yan Zhu, James McCann, and Jean Oh · 2024
Closest in time.
A multimodal automated interpretability agent
Tamar Rott Shaham, Sarah Schwettmann, Franklin Wang, Achyuta Rajaram, Evan Hernandez, Jacob Andreas, and Antonio Torralba · 2024
Closest in time.
Hao Shao, Shengju Qian, Han Xiao, Guanglu Song, Zhuofan Zong, Letian Wang, Yu Liu, and Hongsheng Li · 2024
Closest in time.
A vision check-up for language models
Pratyusha Sharma, Tamar Rott Shaham, Manel Baradad, Stephanie Fu, Adrian Rodriguez-Munoz, Shivam Duggal, Phillip Isola, and Antonio Torralba · 2024
Closest in time.
Gemini: A family of highly capable multimodal models, 2024
Gemini Team · 2024
Closest in time.
Eyes wide shut? exploring the visual shortcomings of multimodal llms, 2024
Shengbang Tong, Zhuang Liu, Yuexiang Zhai, Yi Ma, Yann LeCun, and Saining Xie · 2024
Closest in time.
Chain-of-thought prompting elicits reasoning in large language models
Jason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma, Brian Ichter, Fei Xia, Ed H. Chi, Quoc V. Le, and Denny Zhou · 2024
Closest in time.
Svgdreamer: Text guided svg generation with diffusion model
Ximing Xing, Haitao Zhou, Chuang Wang, Jing Zhang, Dong Xu, and Qian Yu · 2024
Closest in time.
Creativeseg: Semantic segmentation of creative sketches
Yixiao Zheng, Kaiyue Pang, Ayan Das, Dongliang Chang, Yi-Zhe Song, and Zhanyu Ma · 2024
Closest in time.