Fetching the paper…
Reading the bibliography…
We describe the class of convexified convolutional neural networks (CCNNs), which capture the parameter sharing of convolutional neural networks in a convex manner.
An empirical study of learning speed in back-propagation networks
S. E. Fahlman · 1988
Earlier work this paper cites.
Training a 3-node neural network is NP-complete
A. L. Blum and R. L. Rivest · 1992
Earlier work this paper cites.
Face recognition: A convolutional neural-network approach
S. Lawrence, C. L. Giles, A. C. Tsoi, and A. D. Back · 1997
Earlier work this paper cites.
Online learning and stochastic approximations
L. Bottou · 1998
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner · 1998
Earlier work this paper cites.
Neural networks with periodic and monotonic activation functions: a comparative study in classification problems
J. M. Sopena, E. Romero, and R. Alquezar · 1999
Earlier work this paper cites.
Rademacher and Gaussian complexities: Risk bounds and structural results
P. L. Bartlett and S. Mendelson · 2003
Earlier work this paper cites.
Convex neural networks
Y. Bengio, N. L. Roux, P. Vincent, O. Delalleau, and P. Marcotte · 2005
Earlier work this paper cites.
On the Nyström method for approximating a Gram matrix for improved kernel-based learning
P. Drineas and M. W. Mahoney · 2005
Earlier work this paper cites.
Random features for large-scale kernel machines
A. Rahimi and B. Recht · 2007
Earlier work this paper cites.
Variations on the MNIST digits
VariationsMNIST · 2007
Earlier work this paper cites.
Efficient projections onto the ℓ 1 \ell_{1} -ball for learning in high dimensions
J. Duchi, S. Shalev-Shwartz, Y. Singer, and T. Chandra · 2008
Earlier work this paper cites.
Support vector machines
I. Steinwart and A. Christmann · 2008
Earlier work this paper cites.
Learning multiple layers of features from tiny images
A. Krizhevsky and G. Hinton · 2009
Earlier work this paper cites.
An analysis of single-layer networks in unsupervised feature learning
A. Coates, H. Lee, and A. Y. Ng · 2010
Earlier work this paper cites.
Suitable mlp network activation functions for breast cancer and thyroid disease detection
I. Isa, Z. Saad, S. Omar, M. Osman, K. Ahmad, and H. M. Sakim · 2010
Earlier work this paper cites.
Stacked denoising autoencoders: Learning useful representations in a deep network with a local denoising criterion
P. Vincent, H. Larochelle, I. Lajoie, Y. Bengio, and P.-A. Manzagol · 2010
Cited alongside, same era.
Adaptive subgradient methods for online learning and stochastic optimization
J. Duchi, E. Hazan, and Y. Singer · 2011
Cited alongside, same era.
On some extensions of Bernstein’s inequality for self-adjoint operators
S. Minsker · 2011
Cited alongside, same era.
Learning kernel-based halfspaces with the 0-1 loss
S. Shalev-Shwartz, O. Shamir, and K. Sridharan · 2011
Cited alongside, same era.
Compressed sensing: theory and applications
Y. C. Eldar and G. Kutyniok · 2012
Cited alongside, same era.
Deep neural networks for acoustic modeling in speech recognition: The shared views of four research groups
The loss surface of multilayer networks
A. Choromanska, M. Henaff, M. Mathieu, G. B. Arous, and Y. LeCun · 2014
Later among the works it cites.
Identifying and attacking the saddle point problem in high-dimensional non-convex optimization
Y. N. Dauphin, R. Pascanu, C. Gulcehre, K. Cho, S. Ganguli, and Y. Bengio · 2014
Later among the works it cites.
On the computational efficiency of training neural networks
R. Livni, S. Shalev-Shwartz, and O. Shamir · 2014
Later among the works it cites.
Convolutional kernel networks
J. Mairal, P. Koniusz, Z. Harchaoui, and C. Schmid · 2014
Later among the works it cites.
Provable methods for training neural networks with sparse connectivity
H. Sedghi and A. Anandkumar · 2014
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
G. Hinton, L. Deng, D. Yu, G. E. Dahl, A.-r. Mohamed, N. Jaitly, A. Senior, V. Vanhoucke, P. Nguyen, T. N. Sainath, et al · 2012
Cited alongside, same era.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Cited alongside, same era.
Learning invariant representations with local transformations
K. Sohn and H. Lee · 2012
Cited alongside, same era.
End-to-end text recognition with convolutional neural networks
T. Wang, D. J. Wu, A. Coates, and A. Y. Ng · 2012
Cited alongside, same era.
Convex two-layer modeling
Ö. Aslan, H. Cheng, X. Zhang, and D. Schuurmans · 2013
Cited alongside, same era.
Invariant scattering convolution networks
J. Bruna and S. Mallat · 2013
Cited alongside, same era.
Fastfood-approximating kernel expansions in loglinear time
Q. Le, T. Sarlós, and A. Smola · 2013
Cited alongside, same era.
Dropout: A simple way to prevent neural networks from overfitting
N. Srivastava, G. Hinton, A. Krizhevsky, I. Sutskever, and R. Salakhutdinov · 2014
Later among the works it cites.
A proximal stochastic gradient method with progressive variance reduction
L. Xiao and T. Zhang · 2014
Later among the works it cites.
Pcanet: A simple deep learning baseline for image classification?
T.-H. Chan, K. Jia, S. Gao, J. Lu, Z. Zeng, and Y. Ma · 2015
Later among the works it cites.
Global optimality in tensor factorization, deep learning, and beyond
B. D. Haeffele and R. Vidal · 2015
Later among the works it cites.
Generalization bounds for neural networks through tensor factorization
M. Janzamin, H. Sedghi, and A. Anandkumar · 2015
Later among the works it cites.
Human-level control through deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, et al · 2015
Later among the works it cites.
On the quality of the initial basin in overspecified neural networks
I. Safran and O. Shamir · 2015
Later among the works it cites.
Learning halfspaces and neural networks with random initialization
Y. Zhang, J. D. Lee, M. J. Wainwright, and M. I. Jordan · 2015
Later among the works it cites.
A. Daniely, R. Frostig, and Y. Singer · 2016
Closest in time.
Mastering the game of Go with deep neural networks and tree search
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot, et al · 2016
Closest in time.
ℓ 1 \ell_{1} -regularized neural networks are improperly learnable in polynomial time
Y. Zhang, J. D. Lee, and M. I. Jordan · 2016
Closest in time.