Fetching the paper…
Reading the bibliography…
Diffusion models have achieved remarkable success in generative tasks but suffer from high computational costs due to their iterative sampling process and quadratic attention costs.
An introduction to probability theory and its applications
Feller, W. et al · 1971
Earlier work this paper cites.
A first course in the numerical analysis of differential equations
Iserles, A · 2009
Earlier work this paper cites.
Auto-encoding variational bayes
Kingma, D. P. and Welling, M · 2014
Earlier work this paper cites.
Microsoft coco: Common objects in context
Lin, T.-Y., Maire, M., Belongie, S., Hays, J., Perona, P., Ramanan, D., Dollár, P., and Zitnick, C. L · 2014
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation
Ronneberger, O., Fischer, P., and Brox, T · 2015
Earlier work this paper cites.
Deep unsupervised learning using nonequilibrium thermodynamics
Sohl-Dickstein, J., Weiss, E., Maheswaranathan, N., and Ganguli, S · 2015
Earlier work this paper cites.
Gans trained by a two time-scale update rule converge to a local nash equilibrium
Heusel, M., Ramsauer, H., Unterthiner, T., Nessler, B., and Hochreiter, S · 2017
Earlier work this paper cites.
Attention is all you need
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, Ł., and Polosukhin, I · 2017
Earlier work this paper cites.
The unreasonable effectiveness of deep features as a perceptual metric
Zhang, R., Isola, P., Efros, A. A., Shechtman, E., and Wang, O · 2018
Earlier work this paper cites.
Generative modeling by estimating gradients of the data distribution
Song, Y. and Ermon, S · 2019
Earlier work this paper cites.
Denoising diffusion probabilistic models
Ho, J., Jain, A., and Abbeel, P · 2020
Earlier work this paper cites.
Improved denoising diffusion probabilistic models
Nichol, A. Q. and Dhariwal, P · 2021
Earlier work this paper cites.
Learning transferable visual models from natural language supervision
Radford, A., Kim, J. W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., et al · 2021
Earlier work this paper cites.
Building normalizing flows with stochastic interpolants
Albergo, M. S. and Vanden-Eijnden, E · 2022
Earlier work this paper cites.
Classifier-free diffusion guidance
Ho, J. and Salimans, T · 2022
Earlier work this paper cites.
Elucidating the design space of diffusion-based generative models
Karras, T., Aittala, M., Aila, T., and Laine, S · 2022
Earlier work this paper cites.
Flow straight and fast: Learning to generate and transfer data with rectified flow
Liu, X., Gong, C., and Liu, Q · 2022
Cited alongside, same era.
High-resolution image synthesis with latent diffusion models
Rombach, R., Blattmann, A., Lorenz, D., Esser, P., and Ommer, B · 2022
Cited alongside, same era.
Token merging for fast stable diffusion
Bolya, D. and Hoffman, J · 2023
Cited alongside, same era.
Restoration-degradation beyond linear diffusions: A non-asymptotic analysis for ddim-type samplers
Chen, S., Daras, G., and Dimakis, A · 2023
Cited alongside, same era.
Understanding diffusion objectives as the elbo with simple data augmentation
Kingma, D. and Gao, R · 2023
Cited alongside, same era.
Flow matching for generative modeling
Lipman, Y., Chen, R. T., Ben-Hamu, H., Nickel, M., and Le, M · 2023
Token fusion: Bridging the gap between token pruning and token merging
Kim, M., Gao, S., Hsu, Y.-C., Shen, Y., and Jin, H · 2024
Later among the works it cites.
Timestep embedding tells: It’s time to cache for video diffusion model
Liu, F., Zhang, S., Wang, X., Wei, Y., Qiu, H., Zhao, Y., Zhang, Y., Ye, Q., and Wan, F · 2024
Later among the works it cites.
Simplifying, stabilizing and scaling continuous-time consistency models
Lu, C. and Song, Y · 2024
Later among the works it cites.
Dinov2: Learning robust visual features without supervision
Oquab, M., Darcet, T., Moutakanni, T., Vo, H., Szafraniec, M., Khalidov, V., Fernandez, P., Haziza, D., Massa, F., El-Nouby, A., et al · 2024
Later among the works it cites.
Pfdiff: Training-free acceleration of diffusion models combining past and future scores
Wang, G., Cai, Y., Peng, W., Su, S.-Z., et al · 2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
A unified sampling framework for solver searching of diffusion probabilistic models
Liu, E., Ning, X., Yang, H., and Wang, Y · 2023
Cited alongside, same era.
Scalable diffusion models with transformers
Peebles, W. and Xie, S · 2023
Cited alongside, same era.
Consistency models
Song, Y., Dhariwal, P., Chen, M., and Sutskever, I · 2023
Cited alongside, same era.
Lipschitz singularities in diffusion models
Yang, Z., Feng, R., Zhang, H., Shen, Y., Zhu, K., Huang, L., Zhang, Y., Liu, Y., Zhao, D., Zhou, J., et al · 2023
Cited alongside, same era.
Adding conditional control to text-to-image diffusion models
Zhang, L., Rao, A., and Agrawala, M · 2023
Cited alongside, same era.
Dpm-solver-v3: Improved diffusion ode solver with empirical model statistics
Zheng, K., Lu, C., Chen, J., and Zhu, J · 2023
Cited alongside, same era.
Cache me if you can: Accelerating diffusion models through block caching
Wimbauer, F., Wu, B., Schoenfeld, E., Dai, X., Hou, J., He, Z., Sanakoyeu, A., Zhang, P., Tsai, S., Kohler, J., et al · 2024
Later among the works it cites.
Training-free adaptive diffusion with bounded difference approximation strategy
Ye, H., Yuan, J., Xia, R., Yan, X., Chen, T., Yan, J., Shi, B., and Zhang, B · 2024
Later among the works it cites.
Ditfastattn: Attention compression for diffusion transformer models
Yuan, Z., Zhang, H., Lu, P., Ning, X., Zhang, L., Zhao, T., Yan, S., Dai, G., and Wang, Y · 2024
Later among the works it cites.
Token pruning for caching better: 9 times acceleration on stable diffusion for free
Zhang, E., Xiao, B., Tang, J., Ma, Q., Zou, C., Ning, X., Hu, X., and Zhang, L · 2024
Later among the works it cites.
Real-time video generation with pyramid attention broadcast
Zhao, X., Jin, X., Wang, K., and You, Y · 2024
Later among the works it cites.
Accelerating diffusion transformers with token-wise feature caching
Zou, C., Liu, X., Liu, T., Huang, S., and Zhang, L · 2024
Later among the works it cites.
Timestep embedding tells: It’s time to cache for video diffusion model
Liu, F., Zhang, S., Wang, X., Wei, Y., Qiu, H., Zhao, Y., Zhang, Y., Ye, Q., and Wan, F · 2025
Closest in time.
Saghatchian, O., Moghadam, A. G., and Nickabadi, A · 2025
Closest in time.
Ditfastattnv2: Head-wise attention compression for multi-modality diffusion transformers, 2025
Zhang, H., Su, R., Yuan, Z., Chen, P., Fan, M. S. Y., Yan, S., Dai, G., and Wang, Y · 2025
Closest in time.
Token-aware and step-aware acceleration for stable diffusion
Zhen, T., Cao, J., Sun, X., Pan, J., Ji, Z., and Pang, Y · 2025
Closest in time.