Fetching the paper…
Reading the bibliography…
Diffusion models have demonstrated superior performance across various generative tasks including images, videos, and audio.
Learning multiple layers of features from tiny images
Krizhevsky, A., Hinton, G., et al · 2009
Earlier work this paper cites.
Denoising diffusion implicit models
Song, J., Meng, C., and Ermon, S · 2010
Earlier work this paper cites.
Score-based generative modeling through stochastic differential equations
Song, Y., Sohl-Dickstein, J., Kingma, D. P., Kumar, A., Ermon, S., and Poole, B · 2011
Earlier work this paper cites.
Making a “completely blind” image quality analyzer
Mittal, A., Soundararajan, R., and Bovik, A. C · 2012
Earlier work this paper cites.
Auto-encoding variational bayes
Kingma, D. P. and Welling, M · 2013
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation
Ronneberger, O., Fischer, P., and Brox, T · 2015
Earlier work this paper cites.
Gans trained by a two time-scale update rule converge to a local nash equilibrium
Heusel, M., Ramsauer, H., Unterthiner, T., Nessler, B., and Hochreiter, S · 2017
Earlier work this paper cites.
Progressive growing of gans for improved quality, stability, and variation
Karras, T., Aila, T., Laine, S., and Lehtinen, J · 2018
Earlier work this paper cites.
Denoising diffusion probabilistic models
Ho, J., Jain, A., and Abbeel, P · 2020
Earlier work this paper cites.
Lora: Low-rank adaptation of large language models
Hu, E. J., Shen, Y., Wallis, P., Allen-Zhu, Z., Li, Y., Wang, S., Wang, L., and Chen, W · 2021
Earlier work this paper cites.
Sdedit: Image synthesis and editing with stochastic differential equations
Meng, C., Song, Y., Song, J., Wu, J., Zhu, J.-Y., and Ermon, S · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
Radford, A., Kim, J. W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., et al · 2021
Cited alongside, same era.
Stable diffusion web ui
Automatic1111 · 2022
Cited alongside, same era.
ediffi: Text-to-image diffusion models with an ensemble of expert denoisers
Balaji, Y., Nah, S., Huang, X., Vahdat, A., Song, J., Kreis, K., Aittala, M., Aila, T., Laine, S., Catanzaro, B., et al · 2022
Cited alongside, same era.
Classifier-free diffusion guidance
Ho, J. and Salimans, T · 2022
Cited alongside, same era.
Pseudo numerical methods for diffusion models on manifolds
Liu, L., Ren, Y., Lin, Z., and Zhao, Z · 2022
Cited alongside, same era.
Gu, J., Zhai, S., Zhang, Y., Susskind, J., and Jaitly, N · 2023
Later among the works it cites.
Animatediff: Animate your personalized text-to-image diffusion models without specific tuning
Guo, Y., Yang, C., Rao, A., Wang, Y., Qiao, Y., Lin, D., and Dai, B · 2023
Later among the works it cites.
simple diffusion: End-to-end diffusion for high resolution images
Hoogeboom, E., Heek, J., and Salimans, T · 2023
Later among the works it cites.
Resolution chromatography of diffusion models
Hwang, J., Park, Y.-H., and Jo, J · 2023
Later among the works it cites.
Latent consistency models: Synthesizing high-resolution images with few-step inference
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Dreamfusion: Text-to-3d using 2d diffusion
Poole, B., Jain, A., Barron, J. T., and Mildenhall, B · 2022
Cited alongside, same era.
Hierarchical text-conditional image generation with clip latents
Ramesh, A., Dhariwal, P., Nichol, A., Chu, C., and Chen, M · 2022
Cited alongside, same era.
High-resolution image synthesis with latent diffusion models
Rombach, R., Blattmann, A., Lorenz, D., Esser, P., and Ommer, B · 2022
Cited alongside, same era.
Low-rank adaptation for fast text-to-image diffusion fine-tuning
Ryu, S · 2022
Cited alongside, same era.
Towards robust blind face restoration with codebook lookup transformer
Zhou, S., Chan, K. C., Li, C., and Loy, C. C · 2022
Cited alongside, same era.
Multidiffusion: Fusing diffusion paths for controlled image generation
Bar-Tal, O., Yariv, L., Lipman, Y., and Dekel, T · 2023
Cited alongside, same era.
Stable video diffusion: Scaling latent video diffusion models to large datasets
Blattmann, A., Dockhorn, T., Kulal, S., Mendelevitch, D., Kilian, M., Lorenz, D., Levi, Y., English, Z., Voleti, V., Letts, A., et al
Cited in the paper.
Luo, S., Tan, Y., Huang, L., Li, J., and Zhao, H · 2023
Later among the works it cites.
Sdxl: Improving latent diffusion models for high-resolution image synthesis
Podell, D., English, Z., Lacey, K., Blattmann, A., Dockhorn, T., Müller, J., Penna, J., and Rombach, R · 2023
Later among the works it cites.
If by deepfloyd lab at stabilityai
StabilityAI · 2023
Later among the works it cites.
Ip-adapter: Text compatible image prompt adapter for text-to-image diffusion models
Ye, H., Zhang, J., Liu, S., Han, X., and Yang, W · 2023
Later among the works it cites.
Adding conditional control to text-to-image diffusion models
Zhang, L., Rao, A., and Agrawala, M · 2023
Later among the works it cites.
Any-size-diffusion: Toward efficient text-driven synthesis for any-size hd images
Zheng, Q., Guo, Y., Deng, J., Han, J., Li, Y., Xu, S., and Xu, H · 2023
Later among the works it cites.
Dreamshaper
Lykon · 2024
Closest in time.