Fetching the paper…
Reading the bibliography…
Nowozin \textit{et al} showed last year how to extend the GAN \textit{principle} to all $f$-divergences.
Risk aversion in the small and in the large
J.W. Pratt · 1964
Earlier work this paper cites.
A general class of coefficients of divergence of one distribution from another
S.-M. Ali and S.-D.-S. Silvey · 1966
Earlier work this paper cites.
Information-type measures of difference of probability distributions and indirect observation
I. Csiszár · 1967
Earlier work this paper cites.
Differential-Geometrical Methods in Statistics
S.-I. Amari · 1985
Earlier work this paper cites.
Certainty equivalents and information measures: Duality and extremal principles
A. Ben-Tal, A. Ben-Israel, and M. Teboulle · 1991
Earlier work this paper cites.
About distances of discrete distributions satisfying the data processing Theorem of Information Theory
M.-C. Pardo and I. Vajda · 1997
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner · 1998
Earlier work this paper cites.
On the boosting ability of top-down decision tree learning algorithms
M.J. Kearns and Y. Mansour · 1999
Earlier work this paper cites.
Methods of Information Geometry
S.-I. Amari and H. Nagaoka · 2000
Earlier work this paper cites.
Incorporating second-order functional knowledge for better option pricing
C. Dugas, Y. Bengio, F. Bélisle, C. Nadeau, and R. Garcia · 2000
Earlier work this paper cites.
Digital selection and analogue amplification coexist in a cortex-inspired silicon circuit
R.-H.-R. Hahnloser, R. Sarpeshkar, M.-A. Mahowald, R.-J. Douglas, and H.-S. Seung · 2000
Earlier work this paper cites.
Relative loss bounds for on-line density estimation with the exponential family of distributions
K. S. Azoury and M. K. Warmuth · 2001
Earlier work this paper cites.
Convex optimization
S. Boyd and L. Vandenberghe · 2004
Earlier work this paper cites.
Risk analysis in theory and practice
J.-P. Chavas · 2004
Earlier work this paper cites.
On the optimality of conditional expectation as a bregman predictor
A. Banerjee, X. Guo, and H. Wang · 2005
Earlier work this paper cites.
Clustering with Bregman divergences
A. Banerjee, S. Merugu, I. Dhillon, and J. Ghosh · 2005
Earlier work this paper cites.
Convex Functions and their Applications, A Contemporary Approach
C. Niculescu and L.-E. Persson · 2006
Earlier work this paper cites.
Generalized exponential families and associated entropy functions
J. Naudts · 2008
Earlier work this paper cites.
On the efficient minimization of classification-calibrated surrogates
R. Nock and F. Nielsen · 2008
Earlier work this paper cites.
Bregman divergences and surrogates for learning
R. Nock and F. Nielsen · 2009
Earlier work this paper cites.
Bregman voronoi diagrams
J.-D. Boissonnat, F. Nielsen, and R. Nock · 2010
Earlier work this paper cites.
Rectified linear units improve restricted Boltzmann machines
V. Nair and G. Hinton · 2010
Cited alongside, same era.
Estimating divergence functionals and the likelihood ratio by convex risk minimization
X. Nguyen, M. J. Wainwright, and M. I. Jordan · 2010
Cited alongside, same era.
Composite binary losses
M.-D. Reid and R.-C. Williamson · 2010
Cited alongside, same era.
Deep sparse rectifier neural networks
X. Glorot, A. Bordes, and Y. Bengio · 2011
Cited alongside, same era.
Generalized thermostatistics
J. Naudts · 2011
Cited alongside, same era.
Information, divergence and risk for binary experiments
M.-D. Reid and R.-C. Williamson · 2011
Cited alongside, same era.
InfoGAN: Interpretable representation learning by information maximizing generative adversarial nets
X. Chen, Y. Duan, R. Houthooft, J. Schulman, I. Sutskever, and P. Abbeel · 2016
Later among the works it cites.
Fast and accurate deep network learning by exponential linear units (ELUs)
D.-A. Clevert, T. Unterthiner, and S. Hochreiter · 2016
Later among the works it cites.
Generative adversarial networks, 2016
I. Goodfellow · 2016
Later among the works it cites.
On conformal divergences and their population minimizers
R. Nock, F. Nielsen, and S.-I. Amari · 2016
Later among the works it cites.
f f -GAN: training generative neural samplers using variational divergence minimization
S. Nowozin, B. Cseke, and R. Tomioka · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
R.-F. Vigelis and C.-C. Cavalcante · 2011
Cited alongside, same era.
Geometry of deformed exponential families: Invariant, dually-flat and conformal geometries
S.-I. Amari, A. Ohara, and H. Matsuzoe · 2012
Cited alongside, same era.
Agglomerative Bregman clustering
M. Telgarsky and S. Dasgupta · 2012
Cited alongside, same era.
Lecture 6.5-rmsprop: Divide the gradient by a running average of its recent magnitude
T. Tieleman and G. Hinton · 2012
Cited alongside, same era.
Rectifier nonlinearities improve neural network acoustic models
A.-L. Maas, A.-Y. Hannun, and A.-Y. Ng · 2013
Cited alongside, same era.
On the difficulty of training recurrent neural networks
R. Pascanu, T. Mikolov, and Y. Bengio · 2013
Cited alongside, same era.
B. Poole, A.-A. Alemi, J. Sohl-Dickstein, and A. Angelova · 2016
Later among the works it cites.
unsupervised representation learning with deep convolutional generative adversarial networks
A. Radford, L. Metz, and S. Chintala · 2016
Later among the works it cites.
Improved techniques for training gans
T. Salimans, I.-J. Goodfellow, W. Zaremba, V. Cheung, A. Radford, and X. Chen · 2016
Later among the works it cites.
Personnal communication, 2017
S.-I. Amari · 2017
Closest in time.
M. Arjovsky, S. Chintala, and L. Bottou · 2017
Closest in time.
Generalization and equilibrium in generative adversarial nets (GANs)
S. Arora, R. Ge, Y. Liang, T. Ma, and Y. Zhang · 2017
Closest in time.
Do GANs actually learn the distribution? an empirical study
S. Arora and Y. Zhang · 2017
Closest in time.
Mode regularized generative adversarial networks
T. Che, Y. Li, A.-P. Jacob, Y. Bengio, and W. Li · 2017
Closest in time.
Density estimation using real NVP
L. Dinh, J. Sohl-Dickstein, and S. Bengio · 2017
Closest in time.
Sinkhorn-autodiff: Tractable Wasserstein learning of generative models
A. Genevay, G. Peyré, and M. Cuturi · 2017
Closest in time.
Improved training of wasserstein GANs
I. Gulrajani, F. Ahmed, M. Arjovsky, V. Dumoulin, and A.-C. Courville · 2017
Closest in time.
On the ability of neural nets to express distributions
H. Lee, R. Ge, T. Ma, A. Risteski, and S. Arora · 2017
Closest in time.
Approximation and convergence properties of generative adversarial learning
S. Liu, O. Bousquet, and K. Chaudhuri · 2017
Closest in time.
Unsupervised creation of parameterized avatars
L. Wolf, Y. Taigman, and A. Polyak · 2017
Closest in time.
Energy-based generative adversarial networks
J. Zhao, M. Mathieu, and Y. LeCun · 2017
Closest in time.