Fetching the paper…
Reading the bibliography…
This survey provides a comprehensive overview of the advancements in Artificial Intelligence in Graphic Design (AIGD), focusing on integrating AI techniques to support design interpretation and enhance the creative process.
Type and image: The language of graphic design
Philip B Meggs · 1992
Earlier work this paper cites.
The document spectrum for page layout analysis
Lawrence O’Gorman · 1993
Earlier work this paper cites.
Separating style and content
Joshua Tenenbaum et al · 1996
Earlier work this paper cites.
A fast algorithm for bottom-up document layout analysis
Anikó Simon et al · 1997
Earlier work this paper cites.
A structural representation for understanding line-drawing images
Jean-Yves Ramel et al · 2000
Earlier work this paper cites.
Font recognition based on global texture analysis
Yong Zhu et al · 2001
Earlier work this paper cites.
Page segmentation for manhattan and non-manhattan layout documents via selective crla
Hung-Ming Sun · 2005
Earlier work this paper cites.
Sketching cartoons by example
Daniel Sỳkora et al · 2005
Earlier work this paper cites.
The design of high-level features for photo quality assessment
Yan Ke et al · 2006
Earlier work this paper cites.
Symbol spotting using full visibility graph representation
Hervé Locteau et al · 2007
Earlier work this paper cites.
Image vectorization using optimized gradient meshes
Jian Sun et al · 2007
Earlier work this paper cites.
Calligraphic packing
Jie Xu and Craig S Kaplan · 2007
Earlier work this paper cites.
Voronoi++: A dynamic page segmentation approach based on voronoi and docstrum features
Mudit Agrawal et al · 2009
Earlier work this paper cites.
Saliency-enhanced image aesthetics class prediction
Lai-Kuan Wong et al · 2009
Earlier work this paper cites.
Patch-based image vectorization with automatic curvilinear feature alignment
Tian Xia et al · 2009
Earlier work this paper cites.
Automatic and topology-preserving gradient mesh generation for image vectorization
Yu-Kun Lai et al · 2009
Earlier work this paper cites.
Vectorizing cartoon animations
Song-Hai Zhang et al · 2009
Earlier work this paper cites.
Review of automatic document formatting
Nathan Hurst et al · 2009
Earlier work this paper cites.
Enhancing document structure analysis using visual analytics
Andreas Stoffel et al · 2010
Earlier work this paper cites.
Texture transfer based on continuous structure of texture patches for design of artistic shodo fonts
Yutaka Goda et al · 2010
Earlier work this paper cites.
Designing with interactive example galleries
Brian Lee et al · 2010
Earlier work this paper cites.
High level describable attributes for predicting aesthetics and interestingness
Sagnik Dhar et al · 2011
Earlier work this paper cites.
Easy generation of personal chinese handwritten fonts
Baoyao Zhou et al · 2011
Earlier work this paper cites.
Probabilistic document model for automated document composition
Niranjan Damera-Venkata et al · 2011
Earlier work this paper cites.
Towards category-based aesthetic models of photographs
Pere Obrador et al · 2012
Earlier work this paper cites.
Algorithm for segmentation of documents based on texture features
AM Vil’kin et al · 2013
Earlier work this paper cites.
Recommendation system for automatic design of magazine covers
Ali Jahanian et al · 2013
Earlier work this paper cites.
Predicting users’ first impressions of website aesthetics with a quantification of perceived visual complexity and colorfulness
Katharina Reinecke et al · 2013
Earlier work this paper cites.
Large-scale visual font recognition
Guang Chen et al · 2014
Earlier work this paper cites.
Rapid: Rating pictorial aesthetics using deep learning
Xin Lu et al · 2014
Earlier work this paper cites.
Learning layouts for single-page graphic designs
Peter O’Donovan et al · 2014
Earlier work this paper cites.
Deepfont: Identify your font from an image
Zhangyang Wang et al · 2015
Earlier work this paper cites.
Color sommelier: Interactive color recommendation system based on community-generated color palettes
KyoungHee Son et al · 2015
Earlier work this paper cites.
Deep multi-patch aggregation network for image aesthetics, and quality estimation
Xin Lu et al · 2015
Earlier work this paper cites.
Flexyfont: Learning transferring rules for flexible typeface synthesis
Huy Quoc Phan et al · 2015
Earlier work this paper cites.
Designscape: Design with interactive layout suggestions
Peter O’Donovan et al · 2015
Earlier work this paper cites.
Deep colorization
Zezhou Cheng et al · 2015
Earlier work this paper cites.
Palette-based photo recoloring
Huiwen Chang et al · 2015
Earlier work this paper cites.
Decomposing images into layers via rgb-space geometry
Jianchao Tan et al · 2016
Earlier work this paper cites.
Colorful image colorization
Richard Zhang et al · 2016
Earlier work this paper cites.
Ssd: Single shot multibox detector
Wei Liu et al · 2016
Earlier work this paper cites.
Detecting text in natural image with connectionist text proposal network
Zhi Tian et al · 2016
Earlier work this paper cites.
Faster r-cnn: Towards real-time object detection with region proposal networks
Shaoqing Ren et al · 2016
Earlier work this paper cites.
Reading text in the wild with convolutional neural networks
Max Jaderberg et al · 2016
Earlier work this paper cites.
Page segmentation using minimum homogeneity algorithm and adaptive mathematical morphology
Tuan Anh Tran et al · 2016
Earlier work this paper cites.
Automatic generation of visual-textual presentation layout
Xuyong Yang et al · 2016
Earlier work this paper cites.
Automatic generation of large-scale handwriting fonts via style
Zhouhui Lian et al · 2016
Earlier work this paper cites.
Japanese kanji-calligraphic font design using onomatopoeia utterance
Kenichi Murata et al · 2016
Earlier work this paper cites.
Legible compact calligrams
Changqing Zou et al · 2016
Earlier work this paper cites.
Synthetic data for text localisation in natural images
Ankush Gupta et al · 2016
Earlier work this paper cites.
Sketchplore: Sketch and explore with a layout optimiser
Kashyap Todi et al · 2016
Earlier work this paper cites.
Directing user attention via visual flow on web designs
Xufang Pang et al · 2016
Earlier work this paper cites.
Deepprop: Extracting deep features from a single image for edit propagation
Yuki Endo et al · 2016
Earlier work this paper cites.
Textboxes: A fast text detector with a single deep neural network
Minghui Liao et al · 2017
Earlier work this paper cites.
Mask r-cnn
Kaiming He et al · 2017
Earlier work this paper cites.
A font style classification system for english ocr
V Bharath et al · 2017
Earlier work this paper cites.
Attention is all you need
A Vaswani et al · 2017
Earlier work this paper cites.
A neural representation of sketch drawings
David Ha et al · 2017
Earlier work this paper cites.
zi2zi: Master chinese calligraphy with conditional adversarial networks
Yuchen Tian · 2017
Earlier work this paper cites.
Dcfont: an end-to-end deep chinese font generation system
Yue Jiang et al · 2017
Earlier work this paper cites.
Awesome typography: Statistics-based text effects transfer
Shuai Yang et al · 2017
Earlier work this paper cites.
Synthesizing ornamental typefaces
Junsong Zhang et al · 2017
Earlier work this paper cites.
Learning visual importance for graphic designs and data visualizations
Zoya Bylinskii et al · 2017
Earlier work this paper cites.
Efficient palette-based decomposition and recoloring of images via rgbxy-space geometry
Jianchao Tan et al · 2018
Earlier work this paper cites.
Char-net: A character-aware neural network for distorted scene text recognition
Wei Liu et al · 2018
Earlier work this paper cites.
Aon: Towards arbitrarily-oriented text recognition
Zhanzhan Cheng et al · 2018
Earlier work this paper cites.
Synthetically supervised feature learning for scene text recognition
Yang Liu et al · 2018
Earlier work this paper cites.
Multi-task adversarial network for disentangled feature learning
Yang Liu et al · 2018
Earlier work this paper cites.
Multi-task layout analysis for historical handwritten documents using fully convolutional networks
Yue Xu et al · 2018
Earlier work this paper cites.
Coloring with words: Guiding image colorization through text-based palette generation
Hyojin Bahng et al · 2018
Earlier work this paper cites.
Separating style and content for generalized style transfer
Yexun Zhang et al · 2018
Earlier work this paper cites.
A common framework for interactive texture transfer
Yifang Men et al · 2018
Earlier work this paper cites.
Multi-content gan for few-shot font style transfer
Samaneh Azadi et al · 2018
Earlier work this paper cites.
Verisimilar image synthesis for accurate detection and recognition of texts
Fangneng Zhan et al · 2018
Earlier work this paper cites.
Two-stage sketch colorization
Lvmin Zhang et al · 2018
Earlier work this paper cites.
Visual font pairing
Shuhui Jiang et al · 2019
Earlier work this paper cites.
An improved geometric approach for palette-based image decomposition and recoloring
Yili Wang et al · 2019
Earlier work this paper cites.
Show, attend and read: A simple and strong baseline for irregular text recognition
Hui Li et al · 2019
Earlier work this paper cites.
Large-scale tag-based font retrieval with generative feature learning
Tianlang Chen et al · 2019
Earlier work this paper cites.
A two-stage method for text line detection in historical documents
Tobias Grüning et al · 2019
Earlier work this paper cites.
A learned representation for scalable vector graphics
Raphael Gontijo Lopes et al · 2019
Earlier work this paper cites.
Svg-vae: Scalable vector graphics with variational auto encoders
Rapha Gontijo Lopes · 2019
Earlier work this paper cites.
Vectorization of line drawings via polyvector fields
Mikhail Bessmeltsev et al · 2019
Earlier work this paper cites.
Artistic glyph image synthesis via one-stage few-shot learning
Yue Gao et al · 2019
Earlier work this paper cites.
Automatic layout generation for graphical design magazines
Sou Tabata et al · 2019
Earlier work this paper cites.
Layoutgan: Generating graphic layouts with wireframe discriminators
Jianan Li, Jimei Yang, Aaron Hertzmann, Jianming Zhang, and Tingfa Xu · 2019
Cited alongside, same era.
Layoutvae: Stochastic scene layout generation from a label set
Akash Abdu Jyothi et al · 2019
Cited alongside, same era.
Content-aware generative modeling of graphic design layouts
Xinru Zheng et al · 2019
Cited alongside, same era.
Adversarial picture colorization with semantic class distribution
Patricia Vitoria et al · 2020
Cited alongside, same era.
Learning semantically enhanced feature for fine-grained image classification
Wei Luo et al · 2020
Cited alongside, same era.
Automatic chinese font generation system reflecting emotions based on generative adversarial network
Diffsketcher: Text guided vector sketch synthesis through latent diffusion models
Ximing Xing et al · 2023
Later among the works it cites.
Vectorfusion: Text-to-svg by abstracting pixel-based diffusion models
Ajay Jain et al · 2023
Later among the works it cites.
Word-as-image for semantic typography
Shir Iluz et al · 2023
Later among the works it cites.
Editable image geometric abstraction via neural primitive assembly
Ye Chen et al · 2023
Later among the works it cites.
Style transfer network for complex multi-stroke text
Fangmei Chen et al · 2023
Later among the works it cites.
An art font generation technique using pix2pix-based networks
Minghao Xue et al · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Lu Chen et al · 2020
Cited alongside, same era.
End-to-end object detection with transformers
Nicolas Carion et al · 2020
Cited alongside, same era.
Read: Recursive autoencoders for document layout generation
Akshay Gadi Patil et al · 2020
Cited alongside, same era.
Palettailor: Discriminable colorization for categorical data
Kecheng Lu et al · 2020
Cited alongside, same era.
Social-sensed image aesthetics assessment
Chaoran Cui et al · 2020
Cited alongside, same era.
Personalized image quality assessment with social-sensed aesthetic preference
Chaoran Cui et al · 2020
Cited alongside, same era.
Adaptive fractional dilated convolution network for image aesthetics assessment
Qiuyu Chen et al · 2020
Cited alongside, same era.
Anything to glyph: Artistic font synthesis via text-to-image diffusion model
Changshuo Wang et al · 2023
Later among the works it cites.
Ds-fusion: Artistic typography via discriminated and stylized diffusion
Maham Tanveer et al · 2023
Later among the works it cites.
Anytext: Multilingual visual text generation and editing
Yuxiang Tuo et al · 2023
Later among the works it cites.
Glyphdraw: Seamlessly rendering text with intricate spatial structures in text-to-image generation
Jian Ma et al · 2023
Later among the works it cites.
Layout generation for various scenarios in mobile shopping applications
Qianzhi Jing et al · 2023
Later among the works it cites.
Two-stage content-aware layout generation for poster designs
Shang Chai et al · 2023
Later among the works it cites.
Layoutdm: Transformer-based diffusion model for layout generation
Shang Chai et al · 2023
Later among the works it cites.
Layoutdiffusion: Improving graphic layout generation by discrete diffusion probabilistic models
Junyi Zhang et al · 2023
Later among the works it cites.
Unifying layout generation with a decoupled diffusion model
Mude Hui et al · 2023
Later among the works it cites.
Layoutdm: Discrete diffusion model for controllable layout generation
Naoto Inoue et al · 2023
Later among the works it cites.
Layoutdit: Layout-aware end-to-end document image translation with multi-step conductive decoder
Zhiyang Zhang et al · 2023
Later among the works it cites.
Dlt: Conditioned layout generation with joint discrete-continuous diffusion layout transformer
Elad Levi et al · 2023
Later among the works it cites.
Posterlayout: A new benchmark and approach for content-aware visual-textual presentation layout
Hsiao Yuan Hsu et al · 2023
Later among the works it cites.
Layoutnuwa: Revealing the hidden layout expertise of large language models
Zecheng Tang et al · 2023
Later among the works it cites.
Region assisted sketch colorization
Ning Wang et al · 2023
Later among the works it cites.
Diffusing colors: Image colorization with text guided diffusion
Nir Zabari et al · 2023
Later among the works it cites.
icolorit: Towards propagating local hints to the right region in interactive colorization by leveraging vision transformer
Jooyeol Yun et al · 2023
Later among the works it cites.
Unsupervised deep exemplar colorization via pyramid dual non-local attention
Hanzhang Wang et al · 2023
Later among the works it cites.
Self-driven dual-path learning for reference-based line art colorization under limited data
Shukai Wu et al · 2023
Later among the works it cites.
Flexicon: Flexible icon colorization via guided images and palettes
Shukai Wu et al · 2023
Later among the works it cites.
L-coins: Language-based colorization with instance awareness
Zheng Chang et al · 2023
Later among the works it cites.
Language-based colorization with any-level descriptions using diffusion priors
Shuchen Weng et al · 2023
Later among the works it cites.
Adding conditional control to text-to-image diffusion models
Lvmin Zhang et al · 2023
Later among the works it cites.
Dreamllm: Synergistic multimodal comprehension and creation, 2023
Shengding Dong et al · 2023
Later among the works it cites.
Openflamingo: An open-source framework for training large autoregressive vision-language models, 2023
Anas Awadalla et al · 2023
Later among the works it cites.
Visual instruction tuning, 2023
Haotian Liu et al · 2023
Later among the works it cites.
Layoutllm-t2i: Eliciting layout guidance from llm for text-to-image generation
Leigang Qu et al · 2023
Later among the works it cites.
Layoutgpt: Compositional visual planning and generation with large language models
Weixi Feng et al · 2023
Later among the works it cites.
Webgpt: Browser-assisted question-answering with human feedback, 2023
Nisan Lee et al · 2023
Later among the works it cites.
Layoutgpt: Compositional visual planning and generation with llms
Weixi Feng et al · 2024
Later among the works it cites.
Visual instruction tuning
Haotian Liu et al · 2024
Later among the works it cites.
Instruct-imagen: Image generation with multi-modal instruction
Hexiang Hu et al · 2024
Later among the works it cites.
Intelligent graphic layout generation: Current status and future perspectives
Zhixiong Liu et al · 2024
Later among the works it cites.
What’s next? exploring utilization, challenges, and future directions of ai-generated image tools in graphic design
Yuying Tang et al · 2024
Later among the works it cites.
Vision-language models for vision tasks: A survey
Jingyi Zhang et al · 2024
Later among the works it cites.
Designprobe: Design rational and intent guided layout generation, 2024
Zhengxia Lin et al · 2024
Later among the works it cites.
Graphicllm: Unleashing the power of large language models for graphic design, 2024
Yuxuan Cheng et al · 2024
Later among the works it cites.
Omnigen: Unified image generation
Shitao Xiao et al · 2024
Later among the works it cites.
Transfusion: Universal transformer-based models for transparent image fusion
Yanchun Zhou et al · 2024
Later among the works it cites.
Hierarchical recognizing vector graphics and a new chart-based vector graphics dataset
Shuguang Dou et al · 2024
Later among the works it cites.
Graphimind: Lmms for graphic design, 2024
Yifeng Huang et al · 2024
Later among the works it cites.
Designprobe: Design rational and intent guided layout generation, 2024
Zhengxia Lin et al · 2024
Later among the works it cites.
Desigen: A pipeline for controllable design template generation
Haohan Weng et al · 2024
Later among the works it cites.
Layoutllm: Layout instruction tuning with llms for document understanding
Chuwei Luo et al · 2024
Later among the works it cites.
Rodla: Benchmarking the robustness of document layout analysis models
Yufan Chen et al · 2024
Later among the works it cites.
M2doc: A multi-modal fusion approach for document layout analysis
Ning Zhang et al · 2024
Later among the works it cites.
A fusion framework of whitespace smear cutting and swin transformer for document layout analysis
Ran Chen et al · 2024
Later among the works it cites.
Diff-font: Diffusion model for robust one-shot font generation
Haibin He et al · 2024
Later among the works it cites.
Omniparser: A unified framework for text spotting key information extraction and table recognition
Jianqiang Wan et al · 2024
Later among the works it cites.
Text grouping adapter: Adapting pre-trained text detector for layout analysis
Tianci Bi et al · 2024
Later among the works it cites.
Extending clip for text-to-font retrieval
Qinghua Sun et al · 2024
Later among the works it cites.
Local interpretations for explainable natural language processing: A survey
Siwen Luo et al · 2024
Later among the works it cites.
Breathing life into sketches using text-to-video priors
Rinon Gal et al · 2024
Later among the works it cites.
Svgdreamer: Text guided svg generation with diffusion model
Ximing Xing et al · 2024
Later among the works it cites.
Text-to-vector generation with neural path representation
Peiying Zhang et al · 2024
Later among the works it cites.
Samvg: A multi-stage image vectorization model with the segment-anything model
Haokun Zhu et al · 2024
Later among the works it cites.
Supersvg: Superpixel-based scalable vector graphics synthesis
Teng Hu et al · 2024
Later among the works it cites.
Textdiffuser: Diffusion models as text painters
Jingye Chen et al · 2024
Later among the works it cites.
Scaling rectified flow transformers for high-resolution image synthesis
Patrick Esser et al · 2024
Later among the works it cites.
Glyphcontrol: Glyph conditional control for visual text generation
Yukang Yang et al · 2024
Later among the works it cites.
Brush your text: Synthesize any scene text on images via diffusion model
Lingjun Zhang et al · 2024
Later among the works it cites.
First creating backgrounds then rendering texts: A new paradigm for visual text blending
Zhenhang Li et al · 2024
Later among the works it cites.
Visual layout composer: Image-vector dual diffusion model for design layout generation
Mohammad Amin Shabani et al · 2024
Later among the works it cites.
Mulan: Multimodal-llm agent for progressive multi-object diffusion
Sen Li et al · 2024
Later among the works it cites.
Textlap: Customizing language models for text-to-layout planning
Jian Chen et al · 2024
Later among the works it cites.
Layoutprompter: Awaken the design ability of large language models
Jiawei Lin et al · 2024
Later among the works it cites.
Gldesigner: Leveraging multi-modal llms as designer for enhanced aesthetic text glyph layouts
Junwen He et al · 2024
Later among the works it cites.
Refining text-to-image generation: Towards accurate training-free glyph-enhanced image generation
Sanyam Lakhanpal et al · 2024
Later among the works it cites.
Lightweight exemplar colorization via semantic attention-guided laplacian pyramid
Chengyi Zou et al · 2024
Later among the works it cites.
Multimodal markup document models for graphic design completion, 2024
Kotaro Kikuchi et al · 2024
Later among the works it cites.
Opencole: Towards reproducible automatic graphic design generation
Naoto Inoue et al · 2024
Later among the works it cites.
Layoutnuwa: Revealing the hidden layout design capability within a multimodal llm, 2024
Bin Wang et al · 2024
Later among the works it cites.
Clip-layout: Language-driven placement modeling for data-free layout generation, 2024
Zhaoyun Gao et al · 2024
Later among the works it cites.
Vascar: Content-aware layout generation via visual-aware self-correction
Jiahao Zhang et al · 2024
Later among the works it cites.
Design-o-meter: Towards evaluating and refining graphic designs
Sahil Goyal et al · 2024
Later among the works it cites.
Magicbrush: A multimodal assistant for visual design editing, 2024
Shihong Zhang et al · 2024
Later among the works it cites.
One diffusion to generate them all
Duong H Le et al · 2024
Later among the works it cites.
Visually descriptive language model for vector graphics reasoning
Zhenhailong Wang et al · 2024
Later among the works it cites.
Layoutdetr: detection transformer is a good multimodal layout designer
Ning Yu et al · 2025
Closest in time.