Fetching the paper…
Reading the bibliography…
In science and engineering, intelligent processing of complex signals such as images, sound or language is often performed by a parameterized hierarchy of nonlinear processing layers, sometimes biologically inspired.
Beyond Regression: New Tools for Prediction and Analysis in the Behavioral Sciences
P. J. Werbos · 1974
Earlier work this paper cites.
Monotone operators and the proximal point algorithm
R. T. Rockafellar · 1976
Earlier work this paper cites.
Maximum likelihood from incomplete data via the EM algorithm
A. P. Dempster, N. M. Laird, and D. B. Rubin · 1977
Earlier work this paper cites.
Learning by choice of internal representations
T. Grossman, R. Meir, and E. Domany · 1988
Earlier work this paper cites.
Nonlinear Multivariate Analysis
A. Gifi · 1990
Earlier work this paper cites.
A cost function for internal representations
A. Krogh, C. J. Thorbergsson, and J. A. Hertz · 1990
Earlier work this paper cites.
The ‘moving targets’ training algorithm
R. Rohwer · 1990
Earlier work this paper cites.
Learning by choice of internal representations: An energy minimization approach
D. Saad and E. Marom · 1990
Earlier work this paper cites.
Sliced inverse regression for dimension reduction
K.-C. Li · 1991
Earlier work this paper cites.
A database for handwritten text recognition research
J. J. Hull · 1994
Earlier work this paper cites.
Adaptive principal surfaces
M. LeBlanc and R. Tibshirani · 1994
Earlier work this paper cites.
On Langevin updating in multilayer perceptrons
T. Rögnvaldsson · 1994
Earlier work this paper cites.
Reducing data dimensionality through optimizing neural network inputs
S. Tan and M. L. Mavrovouniotis · 1995
Earlier work this paper cites.
Columbia object image library (COIL-20)
S. A. Nene, S. K. Nayar, and H. Murase · 1996
Earlier work this paper cites.
Emergence of simple-cell receptive field properties by learning a sparse code for natural images
B. A. Olshausen and D. J. Field · 1996
Earlier work this paper cites.
An efficient EM-based training algorithm for feedforward neural networks
S. Ma, C. Ji, and J. Farmer · 1997
Earlier work this paper cites.
Sparse coding with an overcomplete basis set: A strategy employed by V1?
B. A. Olshausen and D. J. Field · 1997
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner · 1998
Cited alongside, same era.
Neural Networks: Tricks of the Trade , volume 1524 of Lecture Notes in Computer Science
G. B. Orr and K.-R. Müller, editors · 1998
Cited alongside, same era.
Nonlinear component analysis as a kernel eigenvalue problem
B. Schölkopf, A. Smola, and K.-R. Müller · 1998
Cited alongside, same era.
Hierarchical models of object recognition in cortex
M. Riesenhuber and T. Poggio · 1999
Cited alongside, same era.
Speech and Audio Signal Processing: Processing and Perception of Speech and Music
B. Gold and N. Morgan · 2000
Cited alongside, same era.
Nonlinear dimensionality reduction by locally linear embedding
S. T. Roweis and L. K. Saul · 2000
Cited alongside, same era.
Scaling learning algorithms toward AI
Y. Bengio and Y. LeCun · 2007
Later among the works it cites.
Greedy layer-wise training of deep networks
Y. Bengio, P. Lamblin, D. Popovici, and H. Larochelle · 2007
Later among the works it cites.
Unsupervised learning of invariant feature hierarchies with applications to object recognition
M. Ranzato, F. J. Huang, Y. L. Boureau, and Y. LeCun · 2007
Later among the works it cites.
Robust object recognition with cortex-like mechanisms
T. Serre, L. Wolf, S. Bileschi, M. Riesenhuber, and T. Poggio · 2007
Later among the works it cites.
Dimensionality reduction by unsupervised regression
M. Á. Carreira-Perpiñán and Z. Lu · 2008
Later among the works it cites.
Fast inference in sparse coding algorithms with applications to object recognition
K. Kavukcuoglu, M. Ranzato, and Y. LeCun · 2008
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A global geometric framework for nonlinear dimensionality reduction
J. B. Tenenbaum, V. de Silva, and J. C. Langford · 2000
Cited alongside, same era.
Learning with Kernels. Support Vector Machines, Regularization, Optimization, and Beyond
B. Schölkopf and A. J. Smola · 2001
Cited alongside, same era.
Regularized principal manifolds
A. J. Smola, S. Mika, B. Schölkopf, and R. C. Williamson · 2001
Cited alongside, same era.
Introduction to Robotics. Mechanics and Control
J. J. Craig · 2004
Cited alongside, same era.
Linear-least-squares initialization of multilayer perceptrons through backpropagation of the desired response
D. Erdogmus, O. Fontenla-Romero, J. C. Principe, A. Alonso-Betanzos, and E. Castillo · 2005
Cited alongside, same era.
Probabilistic non-linear principal component analysis with Gaussian process latent variable models
N. Lawrence · 2005
Cited alongside, same era.
The difficulty of training deep architectures and the effect of unsupervised pre-training
D. Erhan, P. A. Manzagol, Y. Bengio, S. Bengio, and P. Vincent · 2009
Later among the works it cites.
The Elements of Statistical Learning—Data Mining, Inference and Prediction
T. J. Hastie, R. J. Tibshirani, and J. H. Friedman · 2009
Later among the works it cites.
The elastic embedding algorithm for dimensionality reduction
M. Á. Carreira-Perpiñán · 2010
Later among the works it cites.
Parametric dimensionality reduction by unsupervised regression
M. Á. Carreira-Perpiñán and Z. Lu · 2010
Later among the works it cites.
Distributed optimization and statistical learning via the alternating direction method of multipliers
S. Boyd, N. Parikh, E. Chu, B. Peleato, and J. Eckstein · 2011
Later among the works it cites.
Manifold learning and missing data recovery through unsupervised regression
M. Á. Carreira-Perpiñán and Z. Lu · 2011
Later among the works it cites.
Proximal splitting methods in signal processing
P. L. Combettes and J.-C. Pesquet · 2011
Later among the works it cites.
Learning deep energy models
J. Ngiam, Z. Chen, P. W. Koh, and A. Ng · 2011
Later among the works it cites.
Large-vocabulary continuous speech recognition systems: A look at some recent advances
G. Saon and J.-T. Chien · 2012
Closest in time.
Nonlinear low-dimensional regression using auxiliary coordinates
W. Wang and M. Á. Carreira-Perpiñán · 2012
Closest in time.