Fetching the paper…
Reading the bibliography…
Representing probability distributions by the gradient of their density functions has proven effective in modeling a wide range of continuous data modalities.
Equation of state calculations by fast computing machines
Nicholas Metropolis, Arianna W Rosenbluth, Marshall N Rosenbluth, Augusta H Teller, and Edward Teller · 1953
Earlier work this paper cites.
Monte carlo sampling methods using markov chains and their applications
W Keith Hastings · 1970
Earlier work this paper cites.
Concrete mathematics: a foundation for computer science
Ronald L Graham, Donald E Knuth, Oren Patashnik, and Stanley Liu · 1989
Earlier work this paper cites.
The eigenvalues of mega-dimensional matrices
John Skilling · 1989
Earlier work this paper cites.
A stochastic estimator of the trace of the influence matrix for laplacian smoothing splines
Michael F Hutchinson · 1989
Earlier work this paper cites.
Understanding the metropolis-hastings algorithm
Siddhartha Chib and Edward Greenberg · 1995
Earlier work this paper cites.
The mnist database of handwritten digits
Yann LeCun · 1998
Earlier work this paper cites.
Annealed importance sampling
Radford M Neal · 2001
Earlier work this paper cites.
Estimation of non-normalized statistical models by score matching
Aapo Hyvärinen and Peter Dayan · 2005
Earlier work this paper cites.
Connections between score matching, contrastive divergence, and pseudolikelihood for continuous-valued variables
Aapo Hyvarinen · 2007
Earlier work this paper cites.
Some extensions of score matching
Aapo Hyvärinen · 2007
Earlier work this paper cites.
A connection between score matching and denoising autoencoders
Pascal Vincent · 2011
Earlier work this paper cites.
Sum-product networks: A new deep architecture
Hoifung Poon and Pedro Domingos · 2011
Earlier work this paper cites.
Interpretation and generalization of score matching
Siwei Lyu · 2012
Earlier work this paper cites.
Markov network structure learning: A randomized feature generation approach
Jan Van Haaren and Jesse Davis · 2012
Earlier work this paper cites.
Auto-encoding variational bayes
Diederik P Kingma and Max Welling · 2013
Earlier work this paper cites.
Stochastic backpropagation and approximate inference in deep generative models
Danilo Jimenez Rezende, Shakir Mohamed, and Daan Wierstra · 2014
Earlier work this paper cites.
Generative adversarial nets
Ian Goodfellow, Jean Pouget-Abadie, Mehdi Mirza, Bing Xu, David Warde-Farley, Sherjil Ozair, Aaron Courville, and Yoshua Bengio · 2014
Earlier work this paper cites.
Expectation-maximization for learning determinantal point processes
Jennifer A Gillenwater, Alex Kulesza, Emily Fox, and Ben Taskar · 2014
Earlier work this paper cites.
Made: Masked autoencoder for distribution estimation
Mathieu Germain, Karol Gregor, Iain Murray, and Hugo Larochelle · 2015
Earlier work this paper cites.
Variational inference with normalizing flows
Danilo Rezende and Shakir Mohamed · 2015
Earlier work this paper cites.
Deep unsupervised learning using nonequilibrium thermodynamics
Jascha Sohl-Dickstein, Eric Weiss, Niru Maheswaranathan, and Surya Ganguli · 2015
Earlier work this paper cites.
Accurate and conservative estimates of mrf log-likelihood using reverse annealing
Yuri Burda, Roger Grosse, and Ruslan Salakhutdinov · 2015
Earlier work this paper cites.
U-net: Convolutional networks for biomedical image segmentation
Olaf Ronneberger, Philipp Fischer, and Thomas Brox · 2015
Cited alongside, same era.
beta-vae: Learning basic visual concepts with a constrained variational framework
Irina Higgins, Loic Matthey, Arka Pal, Christopher Burgess, Xavier Glorot, Matthew Botvinick, Shakir Mohamed, and Alexander Lerchner · 2016
Cited alongside, same era.
Conditional image generation with pixelcnn decoders
Aaron Van den Oord, Nal Kalchbrenner, Lasse Espeholt, Oriol Vinyals, Alex Graves, et al · 2016
Cited alongside, same era.
Improved variational inference with inverse autoregressive flow
Durk P Kingma, Tim Salimans, Rafal Jozefowicz, Xi Chen, Ilya Sutskever, and Max Welling · 2016
Cited alongside, same era.
Neural discrete representation learning
Aaron Van Den Oord, Oriol Vinyals, et al · 2017
Cited alongside, same era.
Wavegrad: Estimating gradients for waveform generation
Nanxin Chen, Yu Zhang, Heiga Zen, Ron J Weiss, Mohammad Norouzi, and William Chan · 2020
Later among the works it cites.
Diffwave: A versatile diffusion model for audio synthesis
Zhifeng Kong, Wei Ping, Jiaji Huang, Kexin Zhao, and Bryan Catanzaro · 2020
Later among the works it cites.
Efficient learning of generative models via finite-difference score matching
Tianyu Pang, Kun Xu, Chongxuan Li, Yang Song, Stefano Ermon, and Jun Zhu · 2020
Later among the works it cites.
Block neural autoregressive flow
Nicola De Cao, Wilker Aziz, and Ivan Titov · 2020
Later among the works it cites.
Permutation invariant graph generation via score-based generative modeling
Chenhao Niu, Yang Song, Jiaming Song, Shengjia Zhao, Aditya Grover, and Stefano Ermon · 2020
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Tim Salimans, Andrej Karpathy, Xi Chen, and Diederik P Kingma · 2017
Cited alongside, same era.
Masked autoregressive flow for density estimation
George Papamakarios, Theo Pavlakou, and Iain Murray · 2017
Cited alongside, same era.
Wasserstein generative adversarial networks
Martin Arjovsky, Soumith Chintala, and Léon Bottou · 2017
Cited alongside, same era.
Improved training of wasserstein gans
Ishaan Gulrajani, Faruk Ahmed, Martin Arjovsky, Vincent Dumoulin, and Aaron C Courville · 2017
Cited alongside, same era.
Glow: Generative flow with invertible 1x1 convolutions
Durk P Kingma and Prafulla Dhariwal · 2018
Cited alongside, same era.
Ffjord: Free-form continuous dynamics for scalable reversible generative models
Will Grathwohl, Ricky TQ Chen, Jesse Bettencourt, Ilya Sutskever, and David Duvenaud · 2018
Cited alongside, same era.
Flow++: Improving flow-based generative models with variational dequantization and architecture design
Jonathan Ho, Xi Chen, Aravind Srinivas, Yan Duan, and Pieter Abbeel · 2019
Cited alongside, same era.
Bi-level score matching for learning energy-based latent variable models
Fan Bao, Chongxuan Li, Kun Xu, Hang Su, Jun Zhu, and Bo Zhang · 2020
Later among the works it cites.
Improved autoregressive modeling with distribution smoothing
Chenlin Meng, Jiaming Song, Yang Song, Shengjia Zhao, and Stefano Ermon · 2021
Later among the works it cites.
Accelerating feedforward computation via parallel nonlinear equation solving
Yang Song, Chenlin Meng, Renjie Liao, and Stefano Ermon · 2021
Later among the works it cites.
D2c: Diffusion-decoding models for few-shot conditional generation
Abhishek Sinha, Jiaming Song, Chenlin Meng, and Stefano Ermon · 2021
Later among the works it cites.
Denoising diffusion implicit models
Jiaming Song, Chenlin Meng, and Stefano Ermon · 2021
Later among the works it cites.
Sdedit: Guided image synthesis and editing with stochastic differential equations
Chenlin Meng, Yutong He, Yang Song, Jiaming Song, Jiajun Wu, Jun-Yan Zhu, and Stefano Ermon · 2021
Later among the works it cites.
Estimating high order gradients of the data distribution by denoising
Chenlin Meng, Yang Song, Wenzhe Li, and Stefano Ermon · 2021
Later among the works it cites.
Gradient-guided importance sampling for learning discrete energy-based models
Meng Liu, Haoran Liu, and Shuiwang Ji · 2021
Later among the works it cites.
Argmax flows and multinomial diffusion: Learning categorical distributions
Emiel Hoogeboom, Didrik Nielsen, Priyank Jaini, Patrick Forré, and Max Welling · 2021
Later among the works it cites.
Structured denoising diffusion models in discrete state-spaces
Jacob Austin, Daniel D Johnson, Jonathan Ho, Daniel Tarlow, and Rianne van den Berg · 2021
Later among the works it cites.
Variational (gradient) estimate of the score function in energy-based latent variable models
Fan Bao, Kun Xu, Chongxuan Li, Lanqing Hong, Jun Zhu, and Bo Zhang · 2021
Later among the works it cites.
Oops i took a gradient: Scalable sampling for discrete distributions
Will Grathwohl, Kevin Swersky, Milad Hashemi, David Duvenaud, and Chris Maddison · 2021
Later among the works it cites.
Self-similarity priors: Neural collages as differentiable fractal representations
Michael Poli, Winnie Xu, Stefano Massaroli, Chenlin Meng, Kuno Kim, and Stefano Ermon · 2022
Closest in time.
Butterflyflow: Building invertible layers with butterfly matrices
Chenlin Meng, Linqi Zhou, Kristy Choi, Tri Dao, and Stefano Ermon · 2022
Closest in time.
Dual diffusion implicit bridges for image-to-image translation
Xuan Su, Jiaming Song, Chenlin Meng, and Stefano Ermon · 2022
Closest in time.
Density ratio estimation via infinitesimal classification
Kristy Choi, Chenlin Meng, Yang Song, and Stefano Ermon · 2022
Closest in time.
Generative flow networks for discrete probabilistic modeling
Dinghuai Zhang, Nikolay Malkin, Zhen Liu, Alexandra Volokhova, Aaron Courville, and Yoshua Bengio · 2022
Closest in time.