Fetching the paper…
Reading the bibliography…
Diffusion models excel at capturing complex data distributions, such as those of natural images and proteins.
Schr \ \backslash ” odinger bridge samplers
Bernton, E., J. Heng, A. Doucet, and P. E. Jacob (2019) · 1912
Earlier work this paper cites.
Über die umkehrung der naturgesetze
Schrödinger, E. (1931) · 1931
Earlier work this paper cites.
Stochastic calculus for finance II: Continuous-time models
Shreve, S. E. et al. (2004) · 2004
Earlier work this paper cites.
An introduction to stochastic control theory, path integrals and reinforcement learning
Kappen, H. J. (2007) · 2007
Earlier work this paper cites.
Denoising diffusion implicit models
Song, J., C. Meng, and S. Ermon (2020) · 2010
Earlier work this paper cites.
A generalized path integral control approach to reinforcement learning
Theodorou, E., J. Buchli, and S. Schaal (2010) · 2010
Earlier work this paper cites.
Riemann manifold langevin and hamiltonian monte carlo methods
Girolami, M. and B. Calderhead (2011) · 2011
Earlier work this paper cites.
Score-based generative modeling through stochastic differential equations
Song, Y., J. Sohl-Dickstein, D. P. Kingma, A. Kumar, S. Ermon, and B. Poole (2020) · 2011
Earlier work this paper cites.
Brownian motion and stochastic calculus
Karatzas, I. and S. Shreve (2012) · 2012
Earlier work this paper cites.
Ava: A large-scale database for aesthetic visual analysis
Murray, N., L. Marchesotti, and F. Perronnin (2012) · 2012
Earlier work this paper cites.
Relative entropy and free energy dualities: Connections to path integral and kl control
Theodorou, E. A. and E. Todorov (2012) · 2012
Earlier work this paper cites.
Sinkhorn distances: Lightspeed computation of optimal transport
Cuturi, M. (2013) · 2013
Earlier work this paper cites.
Deep unsupervised learning using nonequilibrium thermodynamics
Sohl-Dickstein, J., E. Weiss, N. Maheswaranathan, and S. Ganguli (2015) · 2015
Earlier work this paper cites.
Survey of variation in human transcription factors reveals prevalent dna binding changes
Barrera, L. A., A. Vedenko, J. V. Kurland, J. M. Rogers, S. S. Gisselbrecht, E. J. Rossin, J. Woodard, L. Mariani, K. H. Kock, S. Inukai, et al. (2016) · 2016
Earlier work this paper cites.
Training deep nets with sublinear memory cost
Chen, T., B. Xu, C. Zhang, and C. Guestrin (2016) · 2016
Earlier work this paper cites.
Memory-efficient backpropagation through time
Gruslys, A., R. Munos, I. Danihelka, M. Lanctot, and A. Graves (2016) · 2016
Earlier work this paper cites.
Local fitness landscape of the green fluorescent protein
Sarkisyan, K. S., D. A. Bolotin, M. V. Meer, D. R. Usmanova, A. S. Mishin, G. V. Sharonov, D. N. Ivankov, N. G. Bozhanova, M. S. Baranov, O. Soylemez, et al. (2016) · 2016
Earlier work this paper cites.
Proximal policy optimization algorithms
Schulman, J., F. Wolski, P. Dhariwal, A. Radford, and O. Klimov (2017) · 2017
Earlier work this paper cites.
Neural ordinary differential equations
Chen, R. T., Y. Rubanova, J. Bettencourt, and D. K. Duvenaud (2018) · 2018
Earlier work this paper cites.
Reinforcement learning and control as probabilistic inference: Tutorial and review
Levine, S. (2018) · 2018
Earlier work this paper cites.
Decoupled weight decay regularization
Loshchilov, I. and F. Hutter (2019) · 2019
Earlier work this paper cites.
Sampling can be faster than optimization
Ma, Y.-A., Y. Chen, C. Jin, N. Flammarion, and M. I. Jordan (2019) · 2019
Earlier work this paper cites.
Theoretical guarantees for sampling and inference in generative models with latent diffusions
Tzen, B. and M. Raginsky (2019) · 2019
Earlier work this paper cites.
High-dimensional statistics: A non-asymptotic viewpoint
Wainwright, M. J. (2019) · 2019
Earlier work this paper cites.
Controlled sequential monte carlo
Heng, J., A. N. Bishop, G. Deligiannidis, and A. Doucet (2020) · 2020
Cited alongside, same era.
Denoising diffusion probabilistic models
Ho, J., A. Jain, and P. Abbeel (2020) · 2020
Cited alongside, same era.
Scalable gradients for stochastic differential equations
Li, X., T.-K. L. Wong, R. T. Chen, and D. Duvenaud (2020) · 2020
Cited alongside, same era.
Learning to summarize with human feedback
Stiennon, N., L. Ouyang, J. Wu, D. Ziegler, R. Lowe, C. Voss, A. Radford, D. Amodei, and P. F. Christiano (2020) · 2020
Cited alongside, same era.
Diffusion schrodinger bridge with applications to score-based generative modeling
De Bortoli, V., J. Thornton, J. Heng, and A. Doucet (2021) · 2021
Cited alongside, same era.
Diffusion models beat gans on image synthesis
Dhariwal, P. and A. Nichol (2021) · 2021
Cited alongside, same era.
Universal guidance for diffusion models
Bansal, A., H.-M. Chu, A. Schwarzschild, S. Sengupta, M. Goldblum, J. Geiping, and T. Goldstein (2023) · 2023
Later among the works it cites.
Gflownet foundations
Bengio, Y., S. Lahlou, T. Deleu, E. J. Hu, M. Tiwari, and E. Bengio (2023) · 2023
Later among the works it cites.
Training diffusion models with reinforcement learning
Black, K., M. Janner, Y. Du, I. Kostrikov, and S. Levine (2023) · 2023
Later among the works it cites.
Open problems and fundamental limitations of reinforcement learning from human feedback
Casper, S., X. Davies, C. Shi, T. K. Gilbert, J. Scheurer, J. Rando, R. Freedman, T. Korbak, D. Lindner, P. Freire, et al. (2023) · 2023
Later among the works it cites.
Directly fine-tuning diffusion models on differentiable rewards
Clark, K., P. Vicol, K. Swersky, and D. J. Fleet (2023) · 2023
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Lora: Low-rank adaptation of large language models
Hu, E. J., Y. Shen, P. Wallis, Z. Allen-Zhu, Y. Li, S. Wang, L. Wang, and W. Chen (2021) · 2021
Cited alongside, same era.
Efficient and accurate gradients for neural sdes
Kidger, P., J. Foster, X. C. Li, and T. Lyons (2021) · 2021
Cited alongside, same era.
Normalizing flows for probabilistic modeling and inference
Papamakarios, G., E. Nalisnick, D. J. Rezende, S. Mohamed, and B. Lakshminarayanan (2021) · 2021
Cited alongside, same era.
Learning transferable visual models from natural language supervision
Radford, A., J. W. Kim, C. Hallacy, A. Ramesh, G. Goh, S. Agarwal, G. Sastry, A. Askell, P. Mishkin, J. Clark, G. Krueger, and I. Sutskever (2021) · 2021
Cited alongside, same era.
Path integral sampler: a stochastic control approach for sampling
Zhang, Q. and Y. Chen (2021) · 2021
Cited alongside, same era.
Constitutional ai: Harmlessness from ai feedback
Bai, Y., S. Kadavath, S. Kundu, A. Askell, J. Kernion, A. Jones, A. Chen, A. Goldie, A. Mirhoseini, C. McKinnon, et al. (2022) · 2022
Cited alongside, same era.
Later among the works it cites.
Inversion by direct iteration: An alternative to denoising diffusion for image restoration
Delbracio, M. and P. Milanfar (2023) · 2023
Later among the works it cites.
Dpok: Reinforcement learning for fine-tuning text-to-image diffusion models
Fan, Y., O. Watkins, Y. Du, H. Liu, M. Ryu, C. Boutilier, P. Abbeel, M. Ghavamzadeh, K. Lee, and K. Lee (2023) · 2023
Later among the works it cites.
Scaling laws for reward model overoptimization
Gao, L., J. Schulman, and J. Hilton (2023) · 2023
Later among the works it cites.
Generative flow networks assisted biological sequence editing
Ghari, P. M., A. Tseng, G. Eraslan, R. Lopez, T. Biancalani, G. Scalia, and E. Hajiramezanali (2023) · 2023
Later among the works it cites.
Protein design with guided discrete diffusion
Gruver, N., S. Stanton, N. C. Frey, T. G. Rudner, I. Hotzel, J. Lafrance-Vanasse, A. Rajpal, K. Cho, and A. G. Wilson (2023) · 2023
Later among the works it cites.
A theory of continuous generative flow networks
Lahlou, S., T. Deleu, P. Lemos, D. Zhang, A. Volokhova, A. Hernández-Garcıa, L. N. Ezzine, Y. Bengio, and N. Malkin (2023) · 2023
Later among the works it cites.
Aligning text-to-image models using human feedback
Lee, K., H. Liu, M. Ryu, O. Watkins, Y. Du, C. Boutilier, P. Abbeel, M. Ghavamzadeh, and S. S. Gu (2023) · 2023
Later among the works it cites.
Flow matching for generative modeling
Lipman, Y., R. T. Chen, H. Ben-Hamu, M. Nickel, and M. Le (2023) · 2023
Later among the works it cites.
I2sb: Image-to-image schrödinger bridge
Liu, G.-H., A. Vahdat, D.-A. Huang, E. A. Theodorou, W. Nie, and A. Anandkumar (2023) · 2023
Later among the works it cites.
Aligning text-to-image diffusion models with reward backpropagation
Prabhudesai, M., A. Goyal, D. Pathak, and K. Fragkiadaki (2023) · 2023
Later among the works it cites.
Diffusion schr \ \backslash ” odinger bridge matching
Shi, Y., V. De Bortoli, A. Campbell, and A. Doucet (2023) · 2023
Later among the works it cites.
Aligned diffusion schr \ \backslash ” odinger bridges
Somnath, V. R., M. Pariset, Y.-P. Hsieh, M. R. Martinez, A. Krause, and C. Bunne (2023) · 2023
Later among the works it cites.
Conditional flow matching: Simulation-free dynamic optimal transport
Tong, A., N. Malkin, G. Huguet, Y. Zhang, J. Rector-Brooks, K. Fatras, G. Wolf, and Y. Bengio (2023) · 2023
Later among the works it cites.
De novo design of protein structure and function with rfdiffusion
Watson, J. L., D. Juergens, N. R. Bennett, B. L. Trippe, J. Yim, H. E. Eisenach, W. Ahern, A. J. Borst, R. J. Ragotte, L. F. Milles, et al. (2023) · 2023
Later among the works it cites.
Better aligning text-to-image models with human preference
Wu, X., K. Sun, F. Zhu, R. Zhao, and H. Li (2023) · 2023
Later among the works it cites.
Imagereward: Learning and evaluating human preferences for text-to-image generation
Xu, J., X. Liu, Y. Wu, Y. Tong, Q. Li, M. Ding, J. Tang, and Y. Dong (2023) · 2023
Later among the works it cites.
Reward-directed conditional diffusion: Provable distribution estimation and reward improvement
Yuan, H., K. Huang, C. Ni, M. Chen, and M. Wang (2023) · 2023
Later among the works it cites.
Zhang, D., R. T. Q. Chen, C.-H. Liu, A. Courville, and Y. Bengio (2023) · 2023
Later among the works it cites.