Fetching the paper…
Reading the bibliography…
In this paper, we propose a novel model for high-dimensional data, called the Hybrid Orthogonal Projection and Estimation (HOPE) model, which combines a linear orthogonal projection and a finite mixture model under a unified generative modeling framework.
The tradeoffs of large scale learning
L. Bottou and O. Bousquet · 1964
Earlier work this paper cites.
Speaker-independent phone recognition using hidden Markov models
K.-L. Lee and H.-W. Hon · 1989
Earlier work this paper cites.
Universal approximation using radial-basis-function networks
J. Park and I. W. Sandberg · 1991
Earlier work this paper cites.
Modelling the manifolds of images of handwritten digits
G. E. Hinton, P. Dayan, and M. Revow · 1997
Earlier work this paper cites.
Dimension reduction by local principal component analysis
N. Kambhatla and T. K. Leen · 1997
Earlier work this paper cites.
Heteroscedastic discriminant analysis and reduced rank HMMs for improved speech recognition
N. Kumar and A. G. Andreou · 1998
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner · 1998
Earlier work this paper cites.
Neural Networks and Brain Function
E. T. Rolls and A. Treves · 1998
Earlier work this paper cites.
EM algorithms for PCA and SPCA
S. Roweis · 1998
Earlier work this paper cites.
Mixtures of probabilistic principal component analyzers
M. E. Tipping and C. M. Bishop · 1999
Earlier work this paper cites.
Probabilistic principle component analysis
M. E. Tipping and C. M. Bishop · 1999
Earlier work this paper cites.
Stochastic learning
L. Bottou · 2004
Cited alongside, same era.
Clustering on the unit hypersphere using von mises-fisher distributions
A. Banerjee, I. S. Dhillon, and J. Ghosh · 2005
Cited alongside, same era.
Handbook of mathematical functions with formulas, graphs, and mathematical tables
M. Abramowitz and I. A. Stegun · 2006
Cited alongside, same era.
Pattern Recognition and Machine Learning
C. M. Bishop · 2006
Cited alongside, same era.
A fast learning algorithm for deep belief nets
G. E. Hinton, S. Osindero, and Y. W. Teh · 2006
Cited alongside, same era.
Reducing the dimensionality of data with neural networks
G. E. Hinton and R. R. Salakhutdinov · 2006
Cited alongside, same era.
Parameter estimation of statistical models using convex optimization: An advanced method of discriminative training for speech and language processing
H. Jiang and X. Li · 2010
Later among the works it cites.
An analysis of single-layer networks in unsupervised feature learning
A. Coates, A. Y. Ng, and H. Lee · 2011
Later among the works it cites.
Improving neural networks by preventing co-adaptation of feature detectors
G. E. Hinton, N. Srivastava, A. Krizhevsky, I. Sutskever, and R. R. Salakhutdinov · 2012
Later among the works it cites.
Investigations of deep neural networks for large vocabulary continuous speech recognition: Why DNN surpasses GMMs in acoustic modelling
J. Pan, C. Liu, Z. Wang, Y. Hu, and H. Jiang · 2012
Later among the works it cites.
Incoherent training of deep neural networks to de-correlate bottleneck features for speech recognition
Y. Bao, H. Jiang, L. Dai, and C. Liu · 2013
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Y. Bengio, P. Lamblin, D. Popovici, and H. Larochelle · 2007
Cited alongside, same era.
Extracting and composing robust features with denoising autoencoders
P. Vincent, H. Larochelle, Y. Bengio, and P. Manzagol · 2008
Cited alongside, same era.
Convolutional deep belief networks for scalable unsupervised learning of hierarchical representations
H. Lee, R. Grosse, R. Ranganath, and A. Ng · 2009
Cited alongside, same era.
Understanding the difficulty of training deep feedforward neural networks
X. Glorot and Y. Bengio · 2010
Cited alongside, same era.
Discriminative training for automatic speech recognition: A survey
H. Jiang · 2010
Cited alongside, same era.
Low-rank matrix factorization for deep neural network training with high-dimensional output targets
T. Sainath, B. Kingsbury, V. Sindhwani, E. Arisoy, and B. Ramabhadran · 2013
Later among the works it cites.
Restructuring of deep neural network acoustic models with singular value decomposition
J. Xue, J. Li, and Y. Gong · 2013
Later among the works it cites.
Do deep nets really need to be deep?
J. Ba and R. Caruana · 2014
Later among the works it cites.
Discriminative learning of generative models: large margin multinomial mixture models for document classification
H. Jiang, Z. Pan, and P. Hu · 2014
Later among the works it cites.
Fast adaptation of deep neural network based on discriminant codes for speech recognition
S. Xue, O. Abdel-Hamid, H. Jiang, L. Dai, and Q. Liu · 2014
Later among the works it cites.