Fetching the paper…
Reading the bibliography…
We present LayerDiffuse, an approach enabling large-scale pretrained latent diffusion models to generate transparent images.
Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer
Colin Raffel, Noam Shazeer, Adam Roberts, Katherine Lee, Sharan Narang, Michael Matena, Yanqi Zhou, Wei Li, and Peter J. Liu. 2019 · 1910
Earlier work this paper cites.
Plug-and-Play Diffusion Features for Text-Driven Image-to-Image Translation. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) . 1921–1930
Narek Tumanyan, Michal Geyer, Shai Bagon, and Tali Dekel. 2023 · 1930
Earlier work this paper cites.
Ian J. Goodfellow, Jonathon Shlens, and Christian Szegedy. 2015 · 2015
Earlier work this paper cites.
Deep unsupervised learning using nonequilibrium thermodynamics
Jascha Sohl-Dickstein, Eric A. Weiss, Niru Maheswaranathan, and Surya Ganguli. 2015 · 2015
Earlier work this paper cites.
Decomposing Time-Lapse Paintings into Layers
Jianchao Tan, Marek Dvorožňák, Daniel Sýkora, and Yotam Gingold. 2015 · 2015
Earlier work this paper cites.
Learning a discriminative model for the perception of realism in composite images. In Proceedings of the IEEE International Conference on Computer Vision . 3943–3951
Jun-Yan Zhu, Philipp Krahenbuhl, Eli Shechtman, and Alexei A Efros. 2015 · 2015
Earlier work this paper cites.
Interactive High-Quality Green-Screen Keying via Color Unmixing
Yağız Aksoy, Tunç Ozan Aydın, Marc Pollefeys, and Aljoša Smolić. 2016 · 2016
Earlier work this paper cites.
Decomposing Images into Layers via RGB-space Geometry
Jianchao Tan, Jyh-Ming Lien, and Yotam Gingold. 2016 · 2016
Earlier work this paper cites.
Unmixing-Based Soft Color Segmentation for Image Manipulation
Yağız Aksoy, Tunç Ozan Aydın, Aljoša Smolić, and Marc Pollefeys. 2017b · 2017
Earlier work this paper cites.
Deep image harmonization. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition . 3789–3797
Yi-Hsuan Tsai, Xiaohui Shen, Zhe Lin, Kalyan Sunkavalli, Xin Lu, and Ming-Hsuan Yang. 2017 · 2017
Earlier work this paper cites.
Ning Xu, Brian Price, Scott Cohen, and Thomas Huang. 2017 · 2017
Earlier work this paper cites.
Unpaired Image-to-Image Translation using Cycle-Consistent Adversarial Networks. In Computer Vision (ICCV), 2017 IEEE International Conference on
Jun-Yan Zhu, Taesung Park, Phillip Isola, and Alexei A Efros. 2017 · 2017
Earlier work this paper cites.
Semantic Soft Segmentation
Yağız Aksoy, Tae-Hyun Oh, Sylvain Paris, Marc Pollefeys, and Wojciech Matusik. 2018 · 2018
Earlier work this paper cites.
Decomposing Images into Layers with Advanced Color Blending
Yuki Koyama and Masataka Goto. 2018 · 2018
Earlier work this paper cites.
Pigmento: Pigment-Based Image Analysis and Editing
Jianchao Tan, Stephen DiVerdi, Jingwan Lu, and Yotam Gingold. 2019 · 2018
Earlier work this paper cites.
Efficient palette-based decomposition and recoloring of images via RGBXY-space geometry
Jianchao Tan, Jose Echevarria, and Yotam Gingold. 2018 · 2018
Earlier work this paper cites.
Invertible Grayscale
Menghan Xia, Xueting Liu, and Tien-Tsin Wong. 2018 · 2018
Earlier work this paper cites.
Learning-based Sampling for Natural Image Matting. In Proc. CVPR
Jingwei Tang, Yağız Aksoy, Cengiz Öztireli, Markus Gross, and Tunç Ozan Aydın. 2019 · 2019
Earlier work this paper cites.
Dovenet: Deep image harmonization via domain verification. In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition . 8394–8403
Wenyan Cong, Jianfu Zhang, Li Niu, Liu Liu, Zhixin Ling, Weiyuan Li, and Liqing Zhang. 2020 · 2020
Earlier work this paper cites.
Denoising diffusion probabilistic models
Jonathan Ho, Ajay Jain, and Pieter Abbeel. 2020 · 2020
Earlier work this paper cites.
Score-based generative modeling through stochastic differential equations
Yang Song, Jascha Sohl-Dickstein, Diederik P. Kingma, Abhishek Kumar, Stefano Ermon, and Ben Poole. 2020 · 2020
Earlier work this paper cites.
Invertible Image Rescaling
Mingqing Xiao, Shuxin Zheng, Chang Liu, Yaolong Wang, Di He, Guolin Ke, Jiang Bian, Zhouchen Lin, and Tie-Yan Liu. 2020 · 2020
Cited alongside, same era.
Intrinsic image harmonization. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 16367–16376
Zonghui Guo, Haiyong Zheng, Yufeng Jiang, Zhaorui Gu, and Bing Zheng. 2021 · 2021
Cited alongside, same era.
LoRA: Low-Rank Adaptation of Large Language Models
Edward J Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen. 2021 · 2021
Cited alongside, same era.
OpenCLIP
Gabriel Ilharco, Mitchell Wortsman, Ross Wightman, Cade Gordon, Nicholas Carlini, Rohan Taori, Achal Dave, Vaishaal Shankar, Hongseok Namkoong, John Miller, Hannaneh Hajishirzi, Ali Farhadi, and Ludwig Schmidt. 2021 · 2021
Cited alongside, same era.
On fast sampling of diffusion probabilistic models
Zhifeng Kong and Wei Ping. 2021 · 2021
Cited alongside, same era.
Image vectorization and editing via linear gradient layer decomposition
Zheng-Jun Du, Liang-Fu Kang, Jianchao Tan, Yotam Gingold, and Kun Xu. 2023 · 2023
Later among the works it cites.
PCT-Net: Full Resolution Image Harmonization Using Pixel-Wise Color Transformations. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) . 5917–5926
Julian Jorge Andrade Guerreiro, Mitsuru Nakazawa, and Björn Stenger. 2023 · 2023
Later among the works it cites.
Prompt-to-Prompt Image Editing with Cross-Attention Control. In International Conference on Learning Representations (ICLR)
Amir Hertz, Ron Mokady, Jay Tenenbaum, Kfir Aberman, Yael Pritch, and Daniel Cohen-Or. 2023 · 2023
Later among the works it cites.
Imagic: Text-Based Real Image Editing with Diffusion Models. In Conference on Computer Vision and Pattern Recognition (CVPR)
Bahjat Kawar, Shiran Zada, Oran Lang, Omer Tov, Huiwen Chang, Tali Dekel, Inbar Mosseri, and Michal Irani. 2023 · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Alex Nichol, Prafulla Dhariwal, Aditya Ramesh, Pranav Shyam, Pamela Mishkin, Bob McGrew, Ilya Sutskever, and Mark Chen. 2021 · 2021
Cited alongside, same era.
Noise estimation for generative diffusion models
Robin San-Roman, Eliya Nachmani, and Lior Wolf. 2021 · 2021
Cited alongside, same era.
Denoising diffusion implicit models
Jiaming Song, Chenlin Meng, and Stefano Ermon. 2021 · 2021
Cited alongside, same era.
Blended diffusion for text-driven editing of natural images. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition . 18208–18218
Omri Avrahami, Dani Lischinski, and Ohad Fried. 2022 · 2022
Cited alongside, same era.
ediffi: Text-to-image diffusion models with an ensemble of expert denoisers
Yogesh Balaji, Seungjun Nah, Xun Huang, Arash Vahdat, Jiaming Song, Karsten Kreis, Miika Aittala, Timo Aila, Samuli Laine, Bryan Catanzaro, et al · 2022
Cited alongside, same era.
PP-Matting: High-Accuracy Natural Image Matting
Guowei Chen, Yi Liu, Jian Wang, Juncai Peng, Yuying Hao, Lutao Chu, Shiyu Tang, Zewu Wu, Zeyu Chen, Zhiliang Yu, Yuning Du, Qingqing Dang, Xiaoguang Hu, and Dianhai Yu. 2022 · 2022
Cited alongside, same era.
An image is worth one word: Personalizing text-to-image generation using textual inversion
Rinon Gal, Yuval Alaluf, Yuval Atzmon, Or Patashnik, Amit H Bermano, Gal Chechik, and Daniel Cohen-Or. 2022 · 2022
Cited alongside, same era.
Alexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao, Chloe Rolland, Laura Gustafson, Tete Xiao, Spencer Whitehead, Alexander C. Berg, Wan-Yen Lo, Piotr Dollár, and Ross Girshick. 2023 · 2023
Later among the works it cites.
Pick-a-Pic: An Open Dataset of User Preferences for Text-to-Image Generation
Yuval Kirstain, Adam Polyak, Uriel Singer, Shahbuland Matiana, Joe Penna, and Omer Levy. 2023 · 2023
Later among the works it cites.
Matting Anything
Jiachen Li, Jitesh Jain, and Humphrey Shi. 2023b · 2023
Later among the works it cites.
Visual Instruction Tuning. In NeurIPS
Haotian Liu, Chunyuan Li, Qingyang Wu, and Yong Jae Lee. 2023 · 2023
Later among the works it cites.
Null-text inversion for editing real images using guided diffusion models. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) . 6038–6047
Ron Mokady, Amir Hertz, Kfir Aberman, Yael Pritch, and Daniel Cohen-Or. 2023 · 2023
Later among the works it cites.
Chong Mou, Xintao Wang, Liangbin Xie, Jian Zhang, Zhongang Qi, Ying Shan, and Xiaohu Qie. 2023 · 2023
Later among the works it cites.
Deep Image Harmonization with Learnable Augmentation. In Proceedings of the IEEE/CVF International Conference on Computer Vision . 7482–7491
Li Niu, Junyan Cao, Wenyan Cong, and Liqing Zhang. 2023 · 2023
Later among the works it cites.
SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis
Dustin Podell, Zion English, Kyle Lacey, Andreas Blattmann, Tim Dockhorn, Jonas Müller, Joe Penna, and Robin Rombach. 2023 · 2023
Later among the works it cites.
LAION POP: 600,000 High-Resolution Images With Detailed Descriptions
Christoph Schuhmann and Peter Bevan. 2023 · 2023
Later among the works it cites.
Deep image harmonization in dual color spaces. In Proceedings of the 31st ACM International Conference on Multimedia . 2159–2167
Linfeng Tan, Jiangtong Li, Li Niu, and Liqing Zhang. 2023 · 2023
Later among the works it cites.
IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models
Hu Ye, Jun Zhang, Sibo Liu, Xiao Han, and Wei Yang. 2023 · 2023
Later among the works it cites.
Adding conditional control to text-to-image diffusion models
Lvmin Zhang and Maneesh Agrawala. 2023 · 2023
Later among the works it cites.
Text2Layer: Layered Image Generation using Latent Diffusion Model
Xinyang Zhang, Wentian Zhao, Xin Lu, and Jeff Chien. 2023 · 2023
Later among the works it cites.
animagine-xl-3.0
cagliostrolab. 2024 · 2024
Closest in time.
stable-diffusion-xl-1.0-inpainting-0.1
diffusers. 2024 · 2024
Closest in time.
ViTMatte: Boosting image matting with pre-trained plain vision transformers
Jingfeng Yao, Xinggang Wang, Shusheng Yang, and Baoyuan Wang. 2024 · 2024
Closest in time.