Fetching the paper…
Reading the bibliography…
We argue that the estimation of mutual information between high dimensional continuous random variables can be achieved by gradient descent over neural networks.
An algorithm for computing the capacity of arbitrary discrete memoryless channels
Arimoto, S · 1972
Earlier work this paper cites.
Asymptotic evaluation of certain markov process expectations for large time, iv
Donsker, M. and Varadhan, S · 1983
Earlier work this paper cites.
Independent coordinates for strange attractors from mutual information
Fraser, A. M. and Swinney, H. L · 1986
Earlier work this paper cites.
Density-free convergence properties of various estimators of entropy
Györfi, L. and van der Meulen, E. C · 1987
Earlier work this paper cites.
Multilayer feedforward networks are universal approximators
Hornik, K · 1989
Earlier work this paper cites.
Estimation of mutual information using kernel density estimators
Moon, Y.-I., Rajagopalan, B., and Lall, U · 1995
Earlier work this paper cites.
Information theory and statistics
Kullback, S · 1997
Earlier work this paper cites.
Multimodality image registration by maximization of mutual information
Maes, F., Collignon, A., Vandermeulen, D., Marchal, G., and Suetens, P · 1997
Earlier work this paper cites.
The mnist database of handwritten digits
LeCun, Y · 1998
Earlier work this paper cites.
Estimation of the information by an adaptive partitioning of the observation space
Darbellay, G. A. and Vajda, I · 1999
Earlier work this paper cites.
Mutual information relevance networks: functional genomic clustering using pairwise entropy measurements
Butte, A. J. and Kohane, I. S · 2000
Earlier work this paper cites.
The information bottleneck method
Tishby, N., Pereira, F. C., and Bialek, W · 2000
Earlier work this paper cites.
Empirical Processes in M-estimation
Van de Geer, S · 2000
Earlier work this paper cites.
Input feature selection by mutual information based on parzen window
Kwak, N. and Choi, C.-H · 2002
Earlier work this paper cites.
The im algorithm: a variational approach to information maximization
Barber, D. and Agakov, F · 2003
Earlier work this paper cites.
Dual representation of
Keziou, A · 2003
Earlier work this paper cites.
Estimation of entropy and mutual information
Paninski, L · 2003
Earlier work this paper cites.
Independent component analysis , volume 46
Hyvärinen, A., Karhunen, J., and Oja, E · 2004
Earlier work this paper cites.
Estimating mutual information
Kraskov, A., Stögbauer, H., and Grassberger, P · 2004
Earlier work this paper cites.
Image quality assessment: from error visibility to structural similarity
Wang, Z., Bovik, A. C., Sheikh, H. R., and Simoncelli, E. P · 2004
Earlier work this paper cites.
Information bottleneck for gaussian variables
Chechik, G., Globerson, A., Tishby, N., and Weiss, Y · 2005
Earlier work this paper cites.
Feature selection based on mutual information criteria of max-dependency, max-relevance, and min-redundancy
Peng, H., Long, F., and Ding, C · 2005
Cited alongside, same era.
Edgeworth approximation of multivariate differential entropy
Van Hulle, M. M · 2005
Cited alongside, same era.
On baysian bounds
Banerjee, A · 2006
Cited alongside, same era.
Approximating mutual information by maximum likelihood density ratio estimation
Suzuki, T., Sugiyama, M., Sese, J., and Kanamori, T · 2008
Cited alongside, same era.
Estimating divergence functionals and the likelihood ratio by convex risk minimization
Nguyen, X., Wainwright, M. J., and Jordan, M. I · 2010
Cited alongside, same era.
Tighter variational representations of f-divergences via restriction to probability measures
Adversarially learned inference
Dumoulin, V., Belghazi, I., Poole, B., Lamb, A., Arjovsky, M., Mastropietro, O., and Courville, A · 2016
Later among the works it cites.
f-gan: Training generative neural samplers using variational divergence minimization
Nowozin, S., Cseke, B., and Tomioka, R · 2016
Later among the works it cites.
Improved techniques for training gans
Salimans, T., Goodfellow, I. J., Zaremba, W., Cheung, V., Radford, A., and Chen, X · 2016
Later among the works it cites.
Finite-sample analysis of fixed-k nearest neighbor density functional estimators
Singh, S. and Póczos, B · 2016
Later among the works it cites.
Emergence of invariance and disentanglement in deep representations
Achille, A. and Soatto, S · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ruderman, A., Reid, M., García-García, D., and Petterson, J · 2012
Cited alongside, same era.
Efficient estimation of mutual information for strongly dependent variables
Gao, S., Ver Steeg, G., and Galstyan, A · 2014
Cited alongside, same era.
Generative adversarial nets
Goodfellow, I., Pouget-Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair, S., Courville, A., and Bengio, Y · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Kingma, D. P. and Ba, J · 2014
Cited alongside, same era.
Equitability, mutual information, and the maximal information coefficient
Kinney, J. B. and Atwal, G. S · 2014
Cited alongside, same era.
Understanding Machine Learning - from Theory to Algorithms
Shalev-Schwartz, S. and Ben-David, S · 2014
Cited alongside, same era.
Fast and accurate deep network learning by exponential linear units (elus)
Clevert, D., Unterthiner, T., and Hochreiter, S · 2015
Cited alongside, same era.
Ghosh, A., Kulharia, V., Namboodiri, V., Torr, P. H., and Dokania, P. K · 2017
Later among the works it cites.
Nonparametric von mises estimators for entropies, divergences and mutual informations
Kandasamy, K., Krishnamurthy, A., Poczos, B., Wasserman, L., and Robins, J · 2017
Later among the works it cites.
Nonlinear information bottleneck
Kolchinsky, A., Tracey, B. D., and Wolpert, D. H · 2017
Later among the works it cites.
Improved semi-supervised learning with gans using manifold invariances
Kumar, A., Sattigeri, P., and Fletcher, P. T · 2017
Later among the works it cites.
Towards understanding adversarial learning for joint distribution matching
Li, C., Liu, H., Chen, C., Pu, Y., Chen, L., Henao, R., and Carin, L · 2017
Later among the works it cites.
Pacgan: The power of two samples in generative adversarial networks
Lin, Z., Khetan, A., Fanti, G., and Oh, S · 2017
Later among the works it cites.
Unrolled generative adversarial networks
Metz, L., Poole, B., Pfau, D., and Sohl-Dickstein, J · 2017
Later among the works it cites.
Ensemble estimation of mutual information
Moon, K., Sricharan, K., and Hero III, A. O · 2017
Later among the works it cites.
Dual discriminator generative adversarial nets
Nguyen, T., Le, T., Vu, H., and Phung, D · 2017
Later among the works it cites.
Regularizing neural networks by penalizing confident output distributions
Pereyra, G., Tucker, G., Chorowski, J., Kaiser, Ł., and Hinton, G · 2017
Later among the works it cites.
Bayesian gan
Saatchi, Y. and Wilson, A. G · 2017
Later among the works it cites.
Veegan: Reducing mode collapse in gans using implicit variational learning
Srivastava, A., Valkov, L., Russell, C., Gutmann, M., and Sutton, C · 2017
Later among the works it cites.
Adversarial generator-encoder networks
Ulyanov, D., Vedaldi, A., and Lempitsky, V · 2017
Later among the works it cites.
Unpaired image-to-image translation using cycle-consistent adversarial networks
Zhu, J.-Y., Park, T., Isola, P., and Efros, A. A · 2017
Later among the works it cites.
Hierarchical adversarially learned inference
Belghazi, M. I., Rajeswar, S., Mastropietro, O., Mitrovic, J., Rostamzadeh, N., and Courville, A · 2018
Closest in time.