Fetching the paper…
Reading the bibliography…
In the field of graphic design, automating the integration of design elements into a cohesive multi-layered artwork not only boosts productivity but also paves the way for the democratization of graphic design.
Jobling, P., Crowley, D.: Graphic design: reproduction and representation since 1800. Manchester University Press (1996)
1996
Earlier work this paper cites.
Shibata, T., Kurohashi, S.: Automatic slide generation based on discourse structure analysis. In: ICON. pp. 754–766. Springer (2005)
2005
Earlier work this paper cites.
Bauerly, M., Liu, Y.: Computational modeling and experimental investigation of effects of compositional elements on interface and design aesthetics. International journal of human-computer studies 64
2006
Earlier work this paper cites.
Hurst, N., Li, W., Marriott, K.: Review of automatic document formatting. In: Proceedings of the 9th ACM symposium on Document engineering. pp. 99–108 (2009)
2009
Earlier work this paper cites.
Jahanian, A., Liu, J., Lin, Q., Tretter, D., O’Brien-Strain, E., Lee, S.C., Lyons, N., Allebach, J.: Recommendation system for automatic design of magazine covers. In: Proceedings of the 2013 international conference on Intelligent user interfaces. pp. 95–106 (2013)
2013
Earlier work this paper cites.
Yin, W., Mei, T., Chen, C.W.: Automatic generation of social media snippets for mobile browsing. In: Proceedings of the 21st ACM international conference on Multimedia. pp. 927–936 (2013)
2013
Earlier work this paper cites.
Hu, Y., Wan, X.: Ppsgen: Learning-based presentation slides generation for academic papers. IEEE transactions on knowledge and data engineering 27
2014
Earlier work this paper cites.
Plummer, B.A., Wang, L., Cervantes, C.M., Caicedo, J.C., Hockenmaier, J., Lazebnik, S.: Flickr30k entities: Collecting region-to-phrase correspondences for richer image-to-sentence models. In: ICCV. pp. 2641–2649 (2015)
2015
Earlier work this paper cites.
Yang, X., Mei, T., Xu, Y.Q., Rui, Y., Li, S.: Automatic generation of visual-textual presentation layout. ACM Transactions on Multimedia Computing, Communications, and Applications (TOMM) 12
2016
Earlier work this paper cites.
Deka, B., Huang, Z., Franzen, C., Hibschman, J., Afergan, D., Li, Y., Nichols, J., Kumar, R.: Rico: A mobile app dataset for building data-driven design applications. In: Proceedings of the 30th annual ACM symposium on user interface software and technology. pp. 845–854 (2017)
2017
Earlier work this paper cites.
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A.N., Kaiser, Ł., Polosukhin, I.: Attention is all you need. Advances in neural information processing systems 30
2017
Earlier work this paper cites.
Beltramelli, T.: pix2code: Generating code from a graphical user interface screenshot. In: Proceedings of the ACM SIGCHI Symposium on Engineering Interactive Computing Systems. pp. 1–6 (2018)
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
2019
Earlier work this paper cites.
Zheng, X., Qiao, X., Cao, Y., Lau, R.W.: Content-aware generative modeling of graphic design layouts. ACM Transactions on Graphics (TOG) 38
2019
Earlier work this paper cites.
Zhong, X., Tang, J., Yepes, A.J.: Publaynet: largest dataset ever for document layout analysis. In: 2019 International Conference on Document Analysis and Recognition (ICDAR). pp. 1015–1022. IEEE (2019)
2019
Earlier work this paper cites.
Brown, T., Mann, B., Ryder, N., Subbiah, M., Kaplan, J.D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al.: Language models are few-shot learners. Advances in neural information processing systems 33
2020
Earlier work this paper cites.
Carion, N., Massa, F., Synnaeve, G., Usunier, N., Kirillov, A., Zagoruyko, S.: End-to-end object detection with transformers. In: Computer Vision–ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part I 16. pp. 213–229. Springer (2020)
2020
Earlier work this paper cites.
Carlier, A., Danelljan, M., Alahi, A., Timofte, R.: Deepsvg: A hierarchical generative network for vector graphics animation. Advances in Neural Information Processing Systems 33
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
2020
Earlier work this paper cites.
Lee, H.Y., Jiang, L., Essa, I., Le, P.B., Gong, H., Yang, M.H., Yang, W.: Neural design network: Graphic layout generation with constraints. In: Computer Vision–ECCV 2020: 16th European Conference, Glasgow, UK, August 23–28, 2020, Proceedings, Part III 16. pp. 491–506. Springer (2020)
2020
Earlier work this paper cites.
Li, J., Yang, J., Zhang, J., Liu, C., Wang, C., Xu, T.: Attribute-conditioned layout gan for automatic graphic design. IEEE Transactions on Visualization and Computer Graphics 27
2020
Earlier work this paper cites.
Arroyo, D.M., Postels, J., Tombari, F.: Variational transformer networks for layout generation. In: CVPR. pp. 13642–13652 (2021)
2021
Earlier work this paper cites.
2021
Earlier work this paper cites.
Ganin, Y., Bartunov, S., Li, Y., Keller, E., Saliceti, S.: Computer-aided design as language. Advances in Neural Information Processing Systems 34
2021
Earlier work this paper cites.
Para, W., Bhat, S., Guerrero, P., Kelly, T., Mitra, N., Guibas, L.J., Wonka, P.: Sketchgen: Generating constrained cad sketches. Advances in Neural Information Processing Systems 34
2021
Earlier work this paper cites.
Radford, A., Kim, J.W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., et al.: Learning transferable visual models from natural language supervision. In: ICML. pp. 8748–8763. PMLR (2021)
2021
Earlier work this paper cites.
Yamaguchi, K.: Canvasvae: Learning to generate vector graphic documents. In: ICCV. pp. 5481–5489 (2021)
2021
Earlier work this paper cites.
Alayrac, J.B., Donahue, J., Luc, P., Miech, A., Barr, I., Hasson, Y., Lenc, K., Mensch, A., Millican, K., Reynolds, M., et al.: Flamingo: a visual language model for few-shot learning. Advances in Neural Information Processing Systems 35
2022
Earlier work this paper cites.
Feng, S., Jiang, M., Zhou, T., Zhen, Y., Chen, C.: Auto-icon+: An automated end-to-end code generation tool for icon designs in ui development. ACM Transactions on Interactive Intelligent Systems 12
2022
Earlier work this paper cites.
Fu, T.J., Wang, W.Y., McDuff, D., Song, Y.: DOC2PPT: automatic presentation slides generation from scientific documents. In: AAAI. pp. 634–642 (2022)
2022
Earlier work this paper cites.
Jiang, Z., Sun, S., Zhu, J., Lou, J.G., Zhang, D.: Coarse-to-fine generative modeling for graphic layouts. In: AAAI. pp. 1096–1103 (2022)
2022
Earlier work this paper cites.
Kong, X., Jiang, L., Chang, H., Zhang, H., Hao, Y., Gong, H., Essa, I.: Blt: bidirectional layout transformer for controllable layout generation. In: European Conference on Computer Vision. pp. 474–490. Springer (2022)
2022
Earlier work this paper cites.
Wang, P., Yang, A., Men, R., Lin, J., Bai, S., Li, Z., Ma, J., Zhou, C., Zhou, J., Yang, H.: Ofa: Unifying architectures, tasks, and modalities through a simple sequence-to-sequence learning framework. In: ICML. pp. 23318–23340. PMLR (2022)
2022
Earlier work this paper cites.
2022
Cited alongside, same era.
2022
Cited alongside, same era.
Zhou, M., Xu, C., Ma, Y., Ge, T., Jiang, Y., Xu, W.: Composition-aware graphic layout GAN for visual-textual presentation designs. In: IJCAI. pp. 4995–5001. ijcai.org (2022)
2022
Cited alongside, same era.
OpenAI: Gpt-4 technical report (2023)
2023
Later among the works it cites.
OpenAI: Gpt-4v(ision) system card (2023), https://openai.com/research/gpt-4v-system-card
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2023
Cited alongside, same era.
2023
Cited alongside, same era.
Betker, J., Goh, G., Jing, L., Brooks, T., Wang, J., Li, L., Ouyang, L., Zhuang, J., Lee, J., Guo, Y., et al.: Improving image generation with better captions. Computer Science. https://cdn. openai. com/papers/dall-e-3. pdf 2
2023
Cited alongside, same era.
2023
Cited alongside, same era.
Chai, S., Zhuang, L., Yan, F., Zhou, Z.: Two-stage content-aware layout generation for poster designs. In: ACM MM. pp. 8415–8423 (2023)
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
2023
Cited alongside, same era.
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
Team, I.: Internlm: A multilingual language model with progressively enhanced capabilities. https://github.com/InternLM/InternLM (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
Weng, H., Huang, D., Zhang, T., Lin, C.Y.: Learn and sample together: Collaborative generation for graphic design layout. In: IJCAI. pp. 5851–5859 (2023)
2023
Later among the works it cites.
Wu, R., Su, W., Ma, K., Liao, J.: Iconshop: Text-guided vector icon synthesis with autoregressive transformers. ACM Transactions on Graphics (TOG) 42
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
Zhang, J., Guo, J., Sun, S., Lou, J.G., Zhang, D.: LayoutDiffusion: Improving graphic layout generation by discrete diffusion probabilistic models. In: ICCV (2023)
2023
Later among the works it cites.
2023
Later among the works it cites.
2023
Later among the works it cites.
Cui, C., Ma, Y., Cao, X., Ye, W., Zhou, Y., Liang, K., Chen, J., Lu, J., Yang, Z., Liao, K.D., et al.: A survey on multimodal large language models for autonomous driving. In: WACV. pp. 958–979 (2024)
2024
Closest in time.
2024
Closest in time.
Li, X., Yuan, H., Li, W., Ding, H., Wu, S., Zhang, W., Li, Y., Chen, K., Loy, C.C.: Omg-seg: Is one model good enough for all segmentation? arXiv (2024)
2024
Closest in time.
Mu, Y., Zhang, Q., Hu, M., Wang, W., Ding, M., Jin, J., Wang, B., Dai, J., Qiao, Y., Luo, P.: Embodiedgpt: Vision-language pre-training via embodied chain of thought. NeurIPS 36
2024
Closest in time.
Tai, Y., Fan, W., Zhang, Z., Liu, Z.: Link-context learning for multimodal llms. In: CVPR (2024)
2024
Closest in time.
2024
Closest in time.
Wu, J., Li, X., Xu, S., Yuan, H., Ding, H., Yang, Y., Li, X., Zhang, J., Tong, Y., Jiang, X., et al.: Towards open vocabulary learning: A survey. IEEE Transactions on Pattern Analysis and Machine Intelligence (2024)
2024
Closest in time.
2024
Closest in time.
2024
Closest in time.