Fetching the paper…
Reading the bibliography…
This paper introduces Bespoke Non-Stationary (BNS) Solvers, a solver distillation approach to improve sample efficiency of Diffusion and Flow models.
Some practical runge-kutta formulas
Shampine, L. F · 1986
Earlier work this paper cites.
Switchboard: Telephone speech corpus for research and development
Godfrey, J. J., Holliman, E. C., and McDaniel, J · 1992
Earlier work this paper cites.
Numerical optimization
Nocedal, J. and Wright, S. J · 1999
Earlier work this paper cites.
Fisher English training speech parts 1 and 2 LDC200{4,5}S13
Cieri, Christopher, et al. · 2005
Earlier work this paper cites.
Imagenet: A large-scale hierarchical image database
Deng, J., Dong, W., Socher, R., Li, L.-J., Li, K., and Li, F.-F · 2009
Earlier work this paper cites.
A first course in the numerical analysis of differential equations
Iserles, A · 2009
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Krizhevsky, A. and Hinton, G · 2009
Earlier work this paper cites.
Microsoft coco: Common objects in context, 2015
Lin, T.-Y., Maire, M., Belongie, S., Bourdev, L., Girshick, R., Hays, J., Perona, P., Ramanan, D., Zitnick, C. L., and Dollár, P · 2015
Earlier work this paper cites.
Librispeech: An asr corpus based on public domain audio books
Panayotov, V., Chen, G., Povey, D., and Khudanpur, S · 2015
Earlier work this paper cites.
A downsampled variant of imagenet as an alternative to the cifar datasets
Chrabaszcz, P., Loshchilov, I., and Hutter, F · 2017
Earlier work this paper cites.
Gans trained by a two time-scale update rule converge to a local nash equilibrium
Heusel, M., Ramsauer, H., Unterthiner, T., Nessler, B., and Hochreiter, S · 2017
Earlier work this paper cites.
Adam: A method for stochastic optimization, 2017
Kingma, D. P. and Ba, J · 2017
Earlier work this paper cites.
Common voice: A massively-multilingual speech corpus
Ardila, R., Branson, M., Davis, K., Henretty, M., Kohler, M., Meyer, J., Morais, R., Saunders, L., Tyers, F. M., and Weber, G · 2019
Earlier work this paper cites.
Audiocaps: Generating captions for audios in the wild
Kim, C. D., Kim, B., Lee, H., and Kim, G · 2019
Earlier work this paper cites.
Clifton, A., Pappu, A., Reddy, S., Yu, Y., Karlgren, J., Carterette, B., and Jones, R · 2020
Earlier work this paper cites.
Denoising diffusion probabilistic models
Ho, J., Jain, A., and Abbeel, P · 2020
Earlier work this paper cites.
Exploring the limits of transfer learning with a unified text-to-text transformer
Raffel, C., Shazeer, N., Roberts, A., Lee, K., Narang, S., Matena, M., Zhou, Y., Li, W., and Liu, P. J · 2020
Cited alongside, same era.
Score-based generative modeling through stochastic differential equations
Song, Y., Sohl-Dickstein, J., Kingma, D. P., Kumar, A., Ermon, S., and Poole, B · 2020
Cited alongside, same era.
Diffusion models beat gans on image synthesis
Dhariwal, P. and Nichol, A · 2021
Cited alongside, same era.
Variational diffusion models
Kingma, D., Salimans, T., Poole, B., and Ho, J · 2021
Cited alongside, same era.
Knowledge distillation in iterative generative models for improved sampling speed
Luhman, E. and Luhman, T · 2021
Cited alongside, same era.
Hierarchical text-conditional image generation with clip latents, 2022
Ramesh, A., Dhariwal, P., Nichol, A., Chu, C., and Chen, M · 2022
Later among the works it cites.
Progressive distillation for fast sampling of diffusion models
Salimans, T. and Ho, J · 2022
Later among the works it cites.
Make-a-video: Text-to-video generation without text-video data, 2022
Singer, U., Polyak, A., Hayes, T., Yin, X., An, J., Zhang, S., Hu, Q., Yang, H., Ashual, O., Gafni, O., Parikh, D., Gupta, S., and Taigman, Y · 2022
Later among the works it cites.
Denoising diffusion implicit models, 2022
Song, J., Meng, C., and Ermon, S · 2022
Later among the works it cites.
Fast sampling of diffusion models with exponential integrator
Zhang, Q. and Chen, Y · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
High-resolution image synthesis with latent diffusion models, 2021
Rombach, R., Blattmann, A., Lorenz, D., Esser, P., and Ommer, B · 2021
Cited alongside, same era.
Learning fast samplers for diffusion models by differentiating through sample quality
Watson, D., Chan, W., Ho, J., and Norouzi, M · 2021
Cited alongside, same era.
Building normalizing flows with stochastic interpolants, 2022
Albergo, M. S. and Vanden-Eijnden, E · 2022
Cited alongside, same era.
Wavlm: Large-scale self-supervised pre-training for full stack speech processing
Chen, S., Wang, C., Chen, Z., Wu, Y., Liu, S., Chen, Z., Li, J., Kanda, N., Yoshioka, T., Xiao, X., et al · 2022
Cited alongside, same era.
High fidelity neural audio compression, 2022
Défossez, A., Copet, J., Synnaeve, G., and Adi, Y · 2022
Cited alongside, same era.
Classifier-free diffusion guidance
Ho, J. and Salimans, T · 2022
Cited alongside, same era.
Equivariant diffusion for molecule generation in 3d
Hoogeboom, E., Satorras, V. G., Vignac, C., and Welling, M · 2022
Cited alongside, same era.
Pick-a-pic: An open dataset of user preferences for text-to-image generation, 2023
Kirstain, Y., Polyak, A., Singer, U., Matiana, S., Penna, J., and Levy, O · 2023
Later among the works it cites.
Dpm-solver++: Fast solver for guided sampling of diffusion probabilistic models, 2023
Lu, C., Zhou, Y., Bao, F., Chen, J., Li, C., and Zhu, J · 2023
Later among the works it cites.
On distillation of guided diffusion models
Meng, C., Rombach, R., Gao, R., Kingma, D., Ermon, S., Ho, J., and Salimans, T · 2023
Later among the works it cites.
Expresso: A benchmark and analysis of discrete expressive speech resynthesis
Nguyen, T. A., Hsu, W.-N., d’Avirro, A., Shi, B., Gat, I., Fazel-Zarani, M., Remez, T., Copet, J., Synnaeve, G., Hassid, M., et al · 2023
Later among the works it cites.
Training-free linear image inversion via flows
Pokle, A., Muckley, M. J., Chen, R. T., and Karrer, B · 2023
Later among the works it cites.
Bespoke solvers for generative flow models, 2023
Shaul, N., Perez, J., Chen, R. T. Q., Thabet, A., Pumarola, A., and Lipman, Y · 2023
Later among the works it cites.
Audiobox: Unified audio generation with natural language prompts, 2023
Vyas, A., Shi, B., Le, M., Tjandra, A., Wu, Y.-C., Guo, B., Zhang, J., Zhang, X., Adkins, R., Ngan, W., Wang, J., Cruz, I., Akula, B., Akinyemi, A., Ellis, B., Moritz, R., Yungster, Y., Rakotoarison, A., Tan, L., Summers, C., Wood, C., Lane, J., Williamson, M., and Hsu, W.-N · 2023
Later among the works it cites.
Mosaic-sdf for 3d generative models, 2023
Yariv, L., Puny, O., Neverova, N., Gafni, O., and Lipman, Y · 2023
Later among the works it cites.
One-step diffusion with distribution matching distillation, 2023
Yin, T., Gharbi, M., Zhang, R., Shechtman, E., Durand, F., Freeman, W. T., and Park, T · 2023
Later among the works it cites.
Fast sampling of diffusion models with exponential integrator, 2023
Zhang, Q. and Chen, Y · 2023
Later among the works it cites.
Guided flows for generative modeling and decision making
Zheng, Q., Le, M., Shaul, N., Lipman, Y., Grover, A., and Chen, R. T · 2023
Later among the works it cites.