Fetching the paper…
Reading the bibliography…
Generative models, particularly diffusion models, have made significant success in data synthesis across various modalities, including images, videos, and 3D assets.
An empirical bayes approach to statistics
H. E. Robbins · 1992
Earlier work this paper cites.
Denoising diffusion implicit models
J. Song, C. Meng, and S. Ermon · 2010
Earlier work this paper cites.
Tweedie’s formula and selection bias
B. Efron · 2011
Earlier work this paper cites.
Score-based generative modeling through stochastic differential equations
Y. Song, J. Sohl-Dickstein, D. P. Kingma, A. Kumar, S. Ermon, and B. Poole · 2011
Earlier work this paper cites.
Auto-encoding variational bayes
D. P. Kingma and M. Welling · 2013
Earlier work this paper cites.
Deep unsupervised learning using nonequilibrium thermodynamics
J. Sohl-Dickstein, E. Weiss, N. Maheswaranathan, and S. Ganguli · 2015
Earlier work this paper cites.
Density estimation using real nvp
L. Dinh, J. Sohl-Dickstein, and S. Bengio · 2016
Earlier work this paper cites.
Towards principled methods for training generative adversarial networks
M. Arjovsky and L. Bottou · 2017
Earlier work this paper cites.
Wasserstein generative adversarial networks
M. Arjovsky, S. Chintala, and L. Bottou · 2017
Earlier work this paper cites.
Towards accurate generative models of video: A new metric & challenges
T. Unterthiner, S. Van Steenkiste, K. Kurach, R. Marinier, M. Michalski, and S. Gelly · 2018
Earlier work this paper cites.
A style-based generator architecture for generative adversarial networks
T. Karras, S. Laine, and T. Aila · 2019
Earlier work this paper cites.
Generating diverse high-fidelity images with vq-vae-2
A. Razavi, A. Van den Oord, and O. Vinyals · 2019
Earlier work this paper cites.
Generative adversarial networks
I. Goodfellow, J. Pouget-Abadie, M. Mirza, B. Xu, D. Warde-Farley, S. Ozair, A. Courville, and Y. Bengio · 2020
Earlier work this paper cites.
Denoising diffusion probabilistic models
J. Ho, A. Jain, and P. Abbeel · 2020
Earlier work this paper cites.
Diffwave: A versatile diffusion model for audio synthesis
Z. Kong, W. Ping, J. Huang, K. Zhao, and B. Catanzaro · 2020
Earlier work this paper cites.
Frozen in time: A joint video and image encoder for end-to-end retrieval
M. Bain, A. Nagrani, G. Varol, and A. Zisserman · 2021
Earlier work this paper cites.
Diffusion models beat gans on image synthesis
P. Dhariwal and A. Nichol · 2021
Earlier work this paper cites.
Clipscore: A reference-free evaluation metric for image captioning
J. Hessel, A. Holtzman, M. Forbes, R. L. Bras, and Y. Choi · 2021
Earlier work this paper cites.
Knowledge distillation in iterative generative models for improved sampling speed
E. Luhman and T. Luhman · 2021
Earlier work this paper cites.
Improved denoising diffusion probabilistic models
A. Q. Nichol and P. Dhariwal · 2021
Earlier work this paper cites.
Stable diffusion
StabilityAI · 2021
Earlier work this paper cites.
Rectified flow: A marginal preserving approach to optimal transport
Q. Liu · 2022
Cited alongside, same era.
Dpm-solver: A fast ode solver for diffusion probabilistic model sampling in around 10 steps
C. Lu, Y. Zhou, F. Bao, J. Chen, C. Li, and J. Zhu · 2022
Cited alongside, same era.
Dreamfusion: Text-to-3d using 2d diffusion
B. Poole, A. Jain, J. T. Barron, and B. Mildenhall · 2022
Cited alongside, same era.
Hierarchical text-conditional image generation with clip latents
A. Ramesh, P. Dhariwal, A. Nichol, C. Chu, and M. Chen · 2022
Cited alongside, same era.
Progressive distillation for fast sampling of diffusion models
T. Salimans and J. Ho · 2022
Se (3) diffusion model with application to protein backbone generation
J. Yim, B. L. Trippe, V. De Bortoli, E. Mathieu, A. Doucet, R. Barzilay, and T. Jaakkola · 2023
Later among the works it cites.
Video probabilistic diffusion models in projected latent space
S. Yu, K. Sohn, S. Kim, and J. Shin · 2023
Later among the works it cites.
I2vgen-xl: High-quality image-to-video synthesis via cascaded diffusion models
S. Zhang, J. Wang, Y. Zhang, K. Zhao, H. Yuan, Z. Qin, X. Wang, D. Zhao, and J. Zhou · 2023
Later among the works it cites.
Accurate structure prediction of biomolecular interactions with alphafold 3
J. Abramson, J. Adler, J. Dunger, R. Evans, T. Green, A. Pritzel, O. Ronneberger, L. Willmore, A. J. Ballard, J. Bambrick, et al · 2024
Closest in time.
Unifying gans and score-based diffusion as generative particle models
J.-Y. Franceschi, M. Gartrell, L. Dos Santos, T. Issenhuth, E. de Bézenac, M. Chen, and A. Rakotomamonjy · 2024
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Make-a-video: Text-to-video generation without text-video data
U. Singer, A. Polyak, T. Hayes, X. Yin, J. An, S. Zhang, Q. Hu, H. Yang, O. Ashual, O. Gafni, et al · 2022
Cited alongside, same era.
Diffusion-gan: Training gans with diffusion
Z. Wang, H. Zheng, P. He, W. Chen, and M. Zhou · 2022
Cited alongside, same era.
Fast sampling of diffusion models with exponential integrator
Q. Zhang and Y. Chen · 2022
Cited alongside, same era.
Magicvideo: Efficient video generation with latent diffusion models
D. Zhou, W. Wang, H. Yan, W. Lv, Y. Zhu, and J. Feng · 2022
Cited alongside, same era.
Improving image generation with better captions
J. Betker, G. Goh, L. Jing, T. Brooks, J. Wang, L. Li, L. Ouyang, J. Zhuang, J. Lee, Y. Guo, et al · 2023
Cited alongside, same era.
Videocrafter1: Open diffusion models for high-quality video generation
H. Chen, M. Xia, Y. He, Y. Zhang, X. Cun, S. Yang, J. Xing, Y. Liu, Q. Chen, X. Wang, et al · 2023
Cited alongside, same era.
Diffusion policy: Visuomotor policy learning via action diffusion
C. Chi, Z. Xu, S. Feng, E. Cousineau, Y. Du, B. Burchfiel, R. Tedrake, and S. Song · 2023
Cited alongside, same era.
Closest in time.
Imagine flash: Accelerating emu diffusion models with backward distillation
J. Kohler, A. Pumarola, E. Schönfeld, A. Sanakoyeu, R. Sumbaly, P. Vajda, and A. Thabet · 2024
Closest in time.
Act-diffusion: Efficient adversarial consistency training for one-step diffusion models
F. Kong, J. Duan, L. Sun, H. Cheng, R. Xu, H. Shen, X. Zhu, X. Shi, and K. Xu · 2024
Closest in time.
T2v-turbo: Breaking the quality bottleneck of video consistency model with mixed reward feedback
J. Li, W. Feng, T.-J. Fu, X. Wang, S. Basu, W. Chen, and W. Y. Wang · 2024
Closest in time.
Animatediff-lightning: Cross-model diffusion distillation
S. Lin and X. Yang · 2024
Closest in time.
Sdxl-lightning: Progressive adversarial diffusion distillation
S. Lin, A. Wang, and X. Yang · 2024
Closest in time.
One-step diffusion distillation through score implicit matching
W. Luo, Z. Huang, Z. Geng, J. Z. Kolter, and G.-J. Qi · 2024
Closest in time.
Osv: One step is enough for high-quality image to video generation
X. Mao, Z. Jiang, F.-Y. Wang, W. Zhu, J. Zhang, H. Chen, M. Chi, and Y. Wang · 2024
Closest in time.
Openvid-1m: A large-scale high-quality dataset for text-to-video generation
K. Nan, R. Xie, P. Zhou, T. Fan, Z. Yang, Z. Chen, X. Li, J. Yang, and Y. Tai · 2024
Closest in time.
Swiftbrush: One-step text-to-image diffusion model with variational score distillation
T. H. Nguyen and A. Tran · 2024
Closest in time.
Open-sora-plan
PKU-Yuan Lab and Tuzhan AI · 2024
Closest in time.
Perflow: Piecewise rectified flow as universal plug-and-play accelerator
H. Yan, X. Liu, J. Pan, J. H. Liew, Q. Liu, and J. Feng · 2024
Closest in time.
Cogvideox: Text-to-video diffusion models with an expert transformer
Z. Yang, J. Teng, W. Zheng, M. Ding, S. Huang, J. Xu, Y. Yang, W. Hong, X. Zhang, G. Feng, et al · 2024
Closest in time.
Y. Zhai, K. Lin, Z. Yang, L. Li, J. Wang, C.-C. Lin, D. Doermann, J. Yuan, and L. Wang · 2024
Closest in time.
Swiftbrush v2: Make your one-step diffusion model better than its teacher
T. Dao, T. H. Nguyen, T. Le, D. Vu, K. Nguyen, C. Pham, and A. Tran · 2025
Closest in time.
Slimflow: Training smaller one-step diffusion models with rectified flow
Y. Zhu, X. Liu, and Q. Liu · 2025
Closest in time.