Fetching the paper…
Reading the bibliography…
Designs and artworks are ubiquitous across various creative fields, requiring graphic design skills and dedicated software to create compositions that include many graphical elements, such as logos, icons, symbols, and art scenes, which are integral to visual storytelling.
Ben-Ezra, M.: Segmentation with Invisible Keying Signal. In: CVPR (2000)
2000
Earlier work this paper cites.
Rother, C., Kolmogorov, V., Blake, A.: "GrabCut": interactive foreground extraction using iterated graph cuts. ACM Trans. Graphics 23
2004
Earlier work this paper cites.
Bird, S., Loper, E., Klein, E.: Natural Language Processing with Python. O’Reilly Media Inc. (2009)
2009
Earlier work this paper cites.
Rhemann, C., Rother, C., Wang, J., Gelautz, M., Kohli, P., Rott, P.: A perceptually motivated online benchmark for image matting. In: CVPR (2009)
2009
Earlier work this paper cites.
Sohl-Dickstein, J., Weiss, E., Maheswaranathan, N., Ganguli, S.: Deep Unsupervised Learning using Nonequilibrium Thermodynamics. In: ICML (2015)
2015
Earlier work this paper cites.
Zhu, Q., Shao, L., Li, X., Wang, L.: Targeting Accurate Object Extraction From an Image: A Comprehensive Study of Natural Image Matting. IEEE Trans. Neural Netw. Learn. Syst. p. 185 (2015)
2015
Earlier work this paper cites.
Van Den Oord, A., Vinyals, O., et al.: Neural Discrete Representation Learning. NeurIPS (2017)
2017
Earlier work this paper cites.
Xu, N., Price, B., Cohen, S., Huang, T.: Deep image matting. In: CVPR (2017)
2017
Earlier work this paper cites.
Song, Y., Ermon, S.: Generative Modeling by Estimating Gradients of the Data Distribution. NeurIPS (2019)
2019
Earlier work this paper cites.
ZHENG, X., Xiaotian, Q., Ying, C., Rynson, W.: Content-aware generative modeling of graphic design layouts. ACM Trans. Graphics 38
2019
Earlier work this paper cites.
Ho, J., Jain, A., Abbeel, P.: Denoising Diffusion Probabilistic Models. NeurIPS (2020)
2020
Earlier work this paper cites.
Raffel, C., Shazeer, N., Roberts, A., Lee, K., Narang, S., Matena, M., Zhou, Y., Li, W., Liu, P.J.: Exploring the Limits of Transfer Learning with a Unified Text-To-Text Transformer. Journal of Machine Learning Research 21
2020
Earlier work this paper cites.
Song, Y., Ermon, S.: Improved Techniques for Training Score-Based Generative Models. NeurIPS (2020)
2020
Earlier work this paper cites.
Dhariwal, P., Nichol, A.: Diffusion Models Beat GANs on Image Synthesis. NeurIPS (2021)
2021
Earlier work this paper cites.
Esser, P., Rombach, R., Ommer, B.: Taming Transformers for High-Resolution Image Synthesis. In: CVPR (2021)
2021
Earlier work this paper cites.
Hessel, J., Holtzman, A., Forbes, M., Le Bras, R., Choi, Y.: CLIPScore: A Reference-free Evaluation Metric for Image Captioning. In: EMNLP (2021)
2021
Earlier work this paper cites.
Ho, J., Salimans, T.: Classifier-Free Diffusion Guidance. In: NeurIPS Workshop (2021)
2021
Earlier work this paper cites.
Li, D.W., Huang, D., Ma, T., Lin, C.Y.: Towards topic-aware slide generation for academic papers with unsupervised mutual learning. In: AAAI (2021)
2021
Earlier work this paper cites.
Nichol, A.Q., Dhariwal, P.: Improved Denoising Diffusion Probabilistic Models. In: ICML (2021)
2021
Earlier work this paper cites.
Radford, A., Kim, J.W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., et al.: Learning Transferable Visual Models from Natural Language Supervision. In: ICML (2021)
2021
Earlier work this paper cites.
Song, J., Meng, C., Ermon, S.: Denoising Diffusion Implicit Models. In: ICLR (2021)
2021
Earlier work this paper cites.
Song, Y., Sohl-Dickstein, J., Kingma, D.P., Kumar, A., Ermon, S., Poole, B.: Score-Based Generative Modeling through Stochastic Differential Equations. In: ICLR (2021)
2021
Earlier work this paper cites.
Sun, Y., Tang, C.K., Tai, Y.W.: Semantic image matting. In: CVPR (2021)
2021
Earlier work this paper cites.
Yamaguchi, K.: CanvasVAE: Learning to Generate Vector Graphic Documents. In: ICCV. IEEE (2021)
2021
Earlier work this paper cites.
2022
Earlier work this paper cites.
Couairon, G., Verbeek, J., Schwenk, H., Cord, M.: DiffEdit: Diffusion-based semantic image editing with mask guidance. In: ICLR (2022)
2022
Earlier work this paper cites.
Gafni, O., Polyak, A., Ashual, O., Sheynin, S., Parikh, D., Taigman, Y.: Make-a-Scene: Scene-Based Text-to-Image Generation with Human Priors. In: ECCV (2022)
2022
Cited alongside, same era.
2022
Cited alongside, same era.
Karras, T., Aittala, M., Aila, T., Laine, S.: Elucidating the Design Space of Diffusion-Based Generative Models. NeurIPS (2022)
2022
Cited alongside, same era.
Lüddecke, T., Ecker, A.: Image Segmentation Using Text and Image Prompts. In: CVPR (2022)
2022
Cited alongside, same era.
Lugmayr, A., Danelljan, M., Romero, A., Yu, F., Timofte, R., Van Gool, L.: RePaint: Inpainting using Denoising Diffusion Probabilistic Models. In: CVPR (2022)
Kirillov, A., Mintun, E., Ravi, N., Mao, H., Rolland, C., Gustafson, L., Xiao, T., Whitehead, S., Berg, A.C., Lo, W.Y., et al.: Segment anything. In: ICCV (2023)
2023
Later among the works it cites.
Lee, Y., Kim, K., Kim, H., Sung, M.: Syncdiffusion: Coherent montage via Synchronized Joint Diffusions. In: NeurIPS (2023)
2023
Later among the works it cites.
Li, F., Liu, A., Feng, W., Zhu, H., Li, Y., Zhang, Z., Lv, J., Zhu, X., Shen, J., Lin, Z., et al.: Relation-aware diffusion model for controllable poster layout generation. In: CIKM (2023)
2023
Later among the works it cites.
Li, Z., Zhou, Q., Zhang, X., Zhang, Y., Wang, Y., Xie, W.: Open-vocabulary object segmentation with diffusion models. In: ICCV (2023)
2023
Later among the works it cites.
Liu, H., Li, C., Wu, Q., Lee, Y.J.: Visual Instruction Tuning. In: NeurIPS (2023)
2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2022
Cited alongside, same era.
Meng, C., He, Y., Song, Y., Song, J., Wu, J., Zhu, J.Y., Ermon, S.: SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations. In: ICLR (2022)
2022
Cited alongside, same era.
Qin, X., Dai, H., Hu, X., Fan, D.P., Shao, L., Van Gool, L.: Highly Accurate Dichotomous Image Segmentation. In: ECCV (2022)
2022
Cited alongside, same era.
Ramesh, A., Dhariwal, P., Nichol, A., Chu, C., Chen, M.: Hierarchical Text-Conditional Image Generation with CLIP Latents. arXiv e-prints pp. arXiv–2204 (2022)
2022
Cited alongside, same era.
Rombach, R., Blattmann, A., Lorenz, D., Esser, P., Ommer, B.: High-resolution Image Synthesis with Latent Diffusion Models. In: CVPR (2022)
2022
Cited alongside, same era.
Schuhmann, C., Beaumont, R., Vencu, R., Gordon, C., Wightman, R., Cherti, M., Coombes, T., Katta, A., Mullis, C., Wortsman, M., et al.: Laion-5b: An open large-scale dataset for training next generation image-text models. NeurIPS (2022)
2022
Cited alongside, same era.
Von Platen, P., Patil, S., Lozhkov, A., Cuenca, P., Lambert, N., Rasul, K., Davaadorj, M., Wolf, T.: Diffusers: State-of-the-art diffusion models. https://github.com/huggingface/diffusers (2022)
2022
Cited alongside, same era.
Avrahami, O., Fried, O., Lischinski, D.: Blended Latent Diffusion. ACM Trans. Graphics 42
2023
Cited alongside, same era.
Mokady, R., Hertz, A., Aberman, K., Pritch, Y., Cohen-Or, D.: Null-text Inversion for Editing Real Images using Guided Diffusion Models. In: CVPR (2023)
2023
Later among the works it cites.
Peebles, W., Xie, S.: Scalable Diffusion Models with Transformers. In: ICCV (2023)
2023
Later among the works it cites.
Tang, R., Liu, L., Pandey, A., Jiang, Z., Yang, G., Kumar, K., Stenetorp, P., Lin, J., Ture, F.: What the daam: Interpreting stable diffusion using cross attention. In: Proceedings of the Annual Conference of the North American Chapter of the Association for Computational Linguistics. pp. 5644–5659 (2023)
2023
Later among the works it cites.
Tumanyan, N., Geyer, M., Bagon, S., Dekel, T.: Plug-and-Play Diffusion Features for Text-Driven Image-to-Image Translation. In: CVPR (2023)
2023
Later among the works it cites.
Wu, Q., Liu, Y., Zhao, H., Kale, A., Bui, T., Yu, T., Lin, Z., Zhang, Y., Chang, S.: Uncovering the disentanglement capability in text-to-image diffusion models. In: CVPR (2023)
2023
Later among the works it cites.
Zhang, L., Rao, A., Agrawala, M.: Adding Conditional Control to Text-to-Image Diffusion Models. In: ICCV (2023)
2023
Later among the works it cites.
Zhang, Q., Song, J., Huang, X., Chen, Y., Liu, M.Y.: DiffCollage: Parallel Generation of Large Content With Diffusion Models. In: CVPR (2023)
2023
Later among the works it cites.
Burgert, R.D., Price, B.L., Kuen, J., Li, Y., Ryoo, M.S.: MAGICK: A Large-scale Captioned Dataset from Matting Generated Images using Chroma Keying. In: CVPR (2024)
2024
Closest in time.
2024
Closest in time.
Chen, J., YU, J., GE, C., Yao, L., Xie, E., Wang, Z., Kwok, J., Luo, P., Lu, H., Li, Z.: PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis. In: ICLR (2024)
2024
Closest in time.
Crowson, K., Baumann, S.A., Birch, A., Abraham, T.M., Kaplan, D.Z., Shippole, E.: Scalable high-resolution pixel-space image synthesis with hourglass diffusion transformers. In: ICML (2024)
2024
Closest in time.
Liu, H., Li, C., Wu, Q., Lee, Y.J.: Visual Instruction Tuning. NeurIPS (2024)
2024
Closest in time.
2024
Closest in time.
Nguyen, Q., Vu, T., Tran, A., Nguyen, K.: Dataset diffusion: Diffusion-based synthetic data generation for pixel-level semantic segmentation. NeurIPS (2024)
2024
Closest in time.
Podell, D., English, Z., Lacey, K., Blattmann, A., Dockhorn, T., Müller, J., Penna, J., Rombach, R.: SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis. In: ICLR (2024)
2024
Closest in time.
Sarukkai, V., Li, L., Ma, A., Ré, C., Fatahalian, K.: Collage Diffusion. In: WACV (2024)
2024
Closest in time.
Weng, H., Huang, D., Qiao, Y., Hu, Z., Lin, C.Y., Zhang, T., Chen, C.: Desigen: A pipeline for controllable design template generation. CVPR (2024)
2024
Closest in time.
Yao, J., Wang, X., Yang, S., Wang, B.: ViTMatte: Boosting image matting with pre-trained plain vision transformers. Inform. Fusion 103
2024
Closest in time.
Yao, J., Wang, X., Ye, L., Liu, W.: Matte anything: Interactive natural image matting with segment anything model. Image and Vision Computing 147
2024
Closest in time.