Fetching the paper…
Reading the bibliography…
Stable diffusion, a generative model used in text-to-image synthesis, frequently encounters resolution-induced composition problems when generating images of varying sizes.
Image Quality Assessment: from Error Visibility to Structural Similarity
Wang, Z.; Bovik, A.; Sheikh, H.; and Simoncelli, E. 2004 · 2004
Earlier work this paper cites.
Generative Adversarial Nets
Goodfellow, I.; Pouget-Abadie, J.; Mirza, M.; Xu, B.; Warde-Farley, D.; Ozair, S.; Courville, A.; and Bengio, Y. 2014 · 2014
Earlier work this paper cites.
Adam: A Method for Stochastic Optimization
Kingma, D. P.; and Ba, J. 2014 · 2014
Earlier work this paper cites.
Microsoft COCO: Common Objects in Context
Lin, T.-Y.; Maire, M.; Belongie, S.; Hays, J.; Perona, P.; Ramanan, D.; Dollár, P.; and Zitnick, C. L. 2014 · 2014
Earlier work this paper cites.
Improved Techniques for Training GANs
Salimans, T.; Goodfellow, I.; Zaremba, W.; Cheung, V.; Radford, A.; and Chen, X. 2016 · 2016
Earlier work this paper cites.
GANs Trained by a Two Time-Scale Update Rule Converge to a Local Nash Equilibrium
Heusel, M.; Ramsauer, H.; Unterthiner, T.; Nessler, B.; and Hochreiter, S. 2017 · 2017
Earlier work this paper cites.
The Unreasonable Effectiveness of Deep Features as a Perceptual Metric
Zhang, R.; Isola, P.; Efros, A. A.; Shechtman, E.; and Wang, O. 2018 · 2018
Earlier work this paper cites.
Pytorch: An imperative style, high-performance deep learning library
Paszke, A.; Gross, S.; Massa, F.; Lerer, A.; Bradbury, J.; Chanan, G.; Killeen, T.; Lin, Z.; Gimelshein, N.; Antiga, L.; et al. 2019 · 2019
Earlier work this paper cites.
Denoising Diffusion Probabilistic Models
Ho, J.; Jain, A.; and Abbeel, P. 2020 · 2020
Earlier work this paper cites.
Denoising Diffusion Implicit Models
Song, J.; Meng, C.; and Ermon, S. 2020 · 2020
Earlier work this paper cites.
Diffusion Models Beat Gans on Image Synthesis
Dhariwal, P.; and Nichol, A. 2021 · 2021
Earlier work this paper cites.
LoRA: Low-Rank Adaptation of Large Language Models
Hu, E. J.; Shen, Y.; Wallis, P.; Allen-Zhu, Z.; Li, Y.; Wang, S.; Wang, L.; and Chen, W. 2021 · 2021
Cited alongside, same era.
Improved Denoising Diffusion Probabilistic models
Nichol, A. Q.; and Dhariwal, P. 2021 · 2021
Cited alongside, same era.
Learning Transferable Visual Models From Natural Language Supervision
Radford, A.; Wook Kim, J.; Hallacy, C.; Ramesh, A.; Goh, G.; Agarwal, S.; Sastry, G.; Askell, A.; Mishkin, P.; Clark, J.; Krueger, G.; and Sutskever, I. 2021 · 2021
Cited alongside, same era.
Score-Based Generative Modeling through Stochastic Differential Equations
Song, Y.; Sohl-Dickstein, J.; Kingma, D. P.; Kumar, A.; Ermon, S.; and Poole, B. 2021 · 2021
Cited alongside, same era.
TediGAN: Text-Guided Diverse Face Image Generation and Manipulation
Xia, W.; Yang, Y.; Xue, J.-H.; and Wu, B. 2021 · 2021
Cited alongside, same era.
MultiDiffusion: Fusing Diffusion Paths for Controlled Image Generation
Bar-Tal, O.; Yariv, L.; Lipman, Y.; and Dekel, T. 2023 · 2023
Closest in time.
Reproducible Scaling Laws for Contrastive Language-Image Learning
Cherti, M.; Beaumont, R.; Wightman, R.; Wortsman, M.; Ilharco, G.; Gordon, C.; Schuhmann, C.; Schmidt, L.; and Jitsev, J. 2023 · 2023
Closest in time.
Dissecting Arbitrary-scale Super-resolution Capability from Pre-trained Diffusion Generative Models
Li, R.; Zhou, Q.; Guo, S.; Zhang, J.; Guo, J.; Jiang, X.; Shen, Y.; and Han, Z. 2023 · 2023
Closest in time.
Solving Diffusion ODEs with Optimal Boundary Conditions for Better Image Super-Resolution
Ma, Y.; Yang, H.; Yang, W.; Fu, J.; and Liu, J. 2023 · 2023
Closest in time.
On Distillation of Guided Diffusion Models
Meng, C.; Rombach, R.; Gao, R.; Kingma, D.; Ermon, S.; Ho, J.; and Salimans, T. 2023 · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models
Nichol, A. Q.; Dhariwal, P.; Ramesh, A.; Shyam, P.; Mishkin, P.; Mcgrew, B.; Sutskever, I.; and Chen, M. 2022 · 2022
Cited alongside, same era.
Hierarchical Text-Conditional Image Generation with CLIP Latents
Ramesh, A.; Dhariwal, P.; Nichol, A.; Chu, C.; and Chen, M. 2022 · 2022
Cited alongside, same era.
High-resolution Image Synthesis with Latent Diffusion Models
Rombach, R.; Blattmann, A.; Lorenz, D.; Esser, P.; and Ommer, B. 2022 · 2022
Cited alongside, same era.
Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding
Saharia, C.; Chan, W.; Saxena, S.; Li, L.; Whang, J.; Denton, E. L.; Ghasemipour, K.; Gontijo Lopes, R.; Karagol Ayan, B.; Salimans, T.; et al. 2022 · 2022
Cited alongside, same era.
LAION-AESTHETICS
Schuhmann, C. 2022 · 2022
Cited alongside, same era.
LAION COCO: 600M Synthetic Captions from LAION2B-EN
Schuhmann, C.; Köpf, A.; Vencu, R.; Coombes, T.; and Beaumont, R. 2022 · 2022
Cited alongside, same era.
Mm-diffusion: Learning Multi-modal Diffusion Models for Joint Audio and Video Generation
Ruan, L.; Ma, Y.; Yang, H.; He, H.; Liu, B.; Fu, J.; Yuan, N. J.; Jin, Q.; and Guo, B. 2023 · 2023
Closest in time.
DreamBooth: Fine Tuning Text-to-Image Diffusion Models for Subject-Driven Generation
Ruiz, N.; Li, Y.; Jampani, V.; Pritch, Y.; Rubinstein, M.; and Aberman, K. 2023 · 2023
Closest in time.
Denoising Diffusion Probabilistic Models for Robust Image Super-Resolution in the Wild
Sahak, H.; Watson, D.; Saharia, C.; and Fleet, D. 2023 · 2023
Closest in time.
Image Super-Resolution via Iterative Refinement
Saharia, C.; Ho, J.; Chan, W.; Salimans, T.; Fleet, D. J.; and Norouzi, M. 2023 · 2023
Closest in time.
Exploiting Diffusion Prior for Real-World Image Super-Resolution
Wang, J.; Yue, Z.; Zhou, S.; Chan, K. C.; and Loy, C. C. 2023 · 2023
Closest in time.
Mixture of Diffusers for Scene Composition and High Resolution Image Generation
Álvaro Barbero Jiménez. 2023 · 2023
Closest in time.