Joint measures and cross-covariance operators
Charles R Baker · 1973
Earlier work this paper cites.
Estimating optimal transformations for multiple regression and correlation
Leo Breiman and Jerome H Friedman · 1985
Earlier work this paper cites.
Remarks on functional canonical variates, alternating least squares methods and ace
Andreas Buja · 1990
Earlier work this paper cites.
Universal approximation bounds for superpositions of a sigmoidal function
Andrew R Barron · 1993
Earlier work this paper cites.
Combining labeled and unlabeled data with co-training
Avrim Blum and Tom Mitchell · 1998
Earlier work this paper cites.
Probabilistic latent semantic indexing
Thomas Hofmann · 1999
Earlier work this paper cites.
Latent semantic indexing: A probabilistic analysis
Christos H Papadimitriou, Prabhakar Raghavan, Hisao Tamaki, and Santosh Vempala · 2000
Earlier work this paper cites.
Rademacher and gaussian complexities: Risk bounds and structural results
Peter L Bartlett and Shahar Mendelson · 2002
Earlier work this paper cites.
Latent dirichlet allocation
David M Blei, Andrew Y Ng, and Michael I Jordan · 2003
Earlier work this paper cites.
Dimensionality reduction for supervised learning with reproducing kernel hilbert spaces
Kenji Fukumizu, Francis R Bach, and Michael I Jordan · 2004
Earlier work this paper cites.
Canonical correlation analysis: An overview with application to learning methods
David R Hardoon, Sandor Szedmak, and John Shawe-Taylor · 2004
Earlier work this paper cites.
Measuring statistical dependence with hilbert-schmidt norms
Arthur Gretton, Olivier Bousquet, Alex Smola, and Bernhard Schölkopf · 2005
Earlier work this paper cites.
Universal kernels
Charles A Micchelli, Yuesheng Xu, and Haizhang Zhang · 2006
Earlier work this paper cites.
Two-view feature generation model for semi-supervised learning
Rie Kubota Ando and Tong Zhang · 2007
Earlier work this paper cites.
Multi-view regression via canonical correlation analysis
Sham M Kakade and Dean P Foster · 2007
Earlier work this paper cites.
Extracting and composing robust features with denoising autoencoders
Pascal Vincent, Hugo Larochelle, Yoshua Bengio, and Pierre-Antoine Manzagol · 2008
Earlier work this paper cites.
Kernel dimension reduction in regression
Kenji Fukumizu, Francis R Bach, Michael I Jordan, et al · 2009
Earlier work this paper cites.
Noise-contrastive estimation: A new estimation principle for unnormalized statistical models
Michael Gutmann and Aapo Hyvärinen · 2010
Earlier work this paper cites.
Testing conditional independence using maximal nonlinear conditional correlation
Tzee-Ming Huang · 2010
Earlier work this paper cites.
Recovering low-rank matrices from few coefficients in any basis
David Gross · 2011
Earlier work this paper cites.
A connection between score matching and denoising autoencoders
Pascal Vincent · 2011
Earlier work this paper cites.
Learning topic models–going beyond svd
Sanjeev Arora, Rong Ge, and Ankur Moitra · 2012
Earlier work this paper cites.
Random design analysis of ridge regression
Daniel Hsu, Sham M Kakade, and Tong Zhang · 2012
Earlier work this paper cites.
Methods of modern mathematical physics: Functional analysis
Michael Reed · 2012
Earlier work this paper cites.
A practical algorithm for topic modeling with provable guarantees
Sanjeev Arora, Rong Ge, Yonatan Halpern, David Mimno, Ankur Moitra, David Sontag, Yichen Wu, and Michael Zhu · 2013
Earlier work this paper cites.