Fetching the paper…
Reading the bibliography…
Kernel normalization methods have been employed to improve robustness of optimization methods to reparametrization of convolution kernels, covariate shift, and to accelerate training of Convolutional Neural Networks (CNNs).
Stochastic estimation of the maximum of a regression function
J. Kiefer and J. Wolfowitz · 1952
Earlier work this paper cites.
Quasi-martingales
D. L. Fisk · 1965
Earlier work this paper cites.
The relationship between variable selection and data agumentation and a method for prediction
D. M. Allen · 1974
Earlier work this paper cites.
A rotation method for computing the qr-decomposition
F. T. Luk · 1986
Earlier work this paper cites.
Optimization by Vector Space Methods
D. G. Luenberger · 1997
Earlier work this paper cites.
A schur–fréchet algorithm for computing the logarithm and exponential of a matrix
C. S. Kenney and A. J. Laub · 1998
Earlier work this paper cites.
Efficient backprop
Y. LeCun · 1998
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner · 1998
Earlier work this paper cites.
Online algorithms and stochastic approximations
D. Saad · 1998
Earlier work this paper cites.
Introduction to smooth manifolds
J. M. Lee · 2003
Earlier work this paper cites.
Nineteen dubious ways to compute the exponential of a matrix, twenty-five years later
C. Moler and C. V. Loan · 2003
Earlier work this paper cites.
Multivariate regression via stiefel manifold constraints
G. H. Bakır, A. Gretton, M. Franz, and B. Schölkopf · 2004
Earlier work this paper cites.
Feature selection, l1 vs. l2 regularization, and rotational invariance
A. Y. Ng · 2004
Earlier work this paper cites.
The scaling and squaring method for the matrix exponential revisited
N. J. Higham · 2005
Earlier work this paper cites.
Covariance, subspace, and intrinsic cramer-rao bounds
S. T. Smith · 2005
Earlier work this paper cites.
Joint diagonalization on the oblique manifold for independent component analysis
P. A. Absil and K. A. Gallivan · 2006
Earlier work this paper cites.
Optimization Algorithms on Matrix Manifolds
P.-A. Absil, R. Mahony, and R. Sepulchre · 2007
Earlier work this paper cites.
Estimation of high-dimensional prior and posterior covariance matrices in kalman filter variants
R. Furrer and T. Bengtsson · 2007
Earlier work this paper cites.
Pattern Theory: From Representation to Inference
U. Grenander and M. Miller · 2007
Earlier work this paper cites.
Regularized estimation of large covariance matrices
P. J. Bickel and E. Levina · 2008
Earlier work this paper cites.
Statistical analysis on stiefel and grassmann manifolds with applications in computer vision
P. K. Turaga, A. Veeraraghavan, and R. Chellappa · 2008
Earlier work this paper cites.
Shrinkage-based diagonal discriminant analysis and its applications in high-dimensional data
H. Pang, T. Tong, and H. Zhao · 2009
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
X. Glorot and Y. Bengio · 2010
Earlier work this paper cites.
The generalized trace-norm and its application to structure-from-motion problems
R. Angst, C. Zach, and M. Pollefeys · 2011
Earlier work this paper cites.
Trace lasso: a trace norm regularization for correlated designs
E. Grave, G. R. Obozinski, and F. R. Bach · 2011
Earlier work this paper cites.
Shape analysis of elastic curves in euclidean spaces
A. Srivastava, E. Klassen, S. H. Joshi, and I. H. Jermyn · 2011
Cited alongside, same era.
Statistical computations on grassmann and stiefel manifolds for image and video-based recognition
P. K. Turaga, A. Veeraraghavan, A. Srivastava, and R. Chellappa · 2011
Cited alongside, same era.
Nonparametric Inference on Manifolds: With Applications to Shape Spaces
A. Bhattacharya and R. Bhattacharya · 2012
Cited alongside, same era.
Large-scale image classification with trace-norm regularization
Z. Harchaoui, M. Douze, M. Paulin, M. Dudik, and J. Malick · 2012
Cited alongside, same era.
Imagenet classification with deep convolutional neural networks
A. Krizhevsky, I. Sutskever, and G. E. Hinton · 2012
Cited alongside, same era.
Advances in matrix manifolds for computer vision
Y. M. Lui · 2012
The information geometry of mirror descent
G. Raskutti and S. Mukherjee · 2015
Later among the works it cites.
Discriminative shape from shading in uncalibrated illumination
S. R. Richter and S. Roth · 2015
Later among the works it cites.
Very deep convolutional networks for large-scale image recognition
K. Simonyan and A. Zisserman · 2015
Later among the works it cites.
Striving for simplicity: The all convolutional net
J. Springenberg, A. Dosovitskiy, T. Brox, and M. Riedmiller · 2015
Later among the works it cites.
Going deeper with convolutions
C. Szegedy, W. Liu, Y. Jia, P. Sermanet, S. Reed, D. Anguelov, D. Erhan, V. Vanhoucke, and A. Rabinovich · 2015
Later among the works it cites.
Riemannian Computing in Computer Vision
P. Turaga and A. Srivastava · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Stochastic gradient descent on riemannian manifolds
S. Bonnabel · 2013
Cited alongside, same era.
Learning hierarchical features for scene labeling
C. Farabet, C. Couprie, L. Najman, and Y. LeCun · 2013
Cited alongside, same era.
Matrix computations
G. Golub and C. Van Loan · 2013
Cited alongside, same era.
Network in network
M. Lin, Q. Chen, and S. Yan · 2013
Cited alongside, same era.
Stochastic geometry and its applications
D. Stoyan, W. S. Kendall, and J. Mecke · 2013
Cited alongside, same era.
On the importance of initialization and momentum in deep learning
I. Sutskever, J. Martens, G. Dahl, and G. Hinton · 2013
Cited alongside, same era.
3d shape estimation from 2d landmarks: A convex relaxation approach
X. Zhou, S. Leonardos, X. Hu, and K. Daniilidis · 2015
Later among the works it cites.
Unitary evolution recurrent neural networks
M. Arjovsky, A. Shah, and Y. Bengio · 2016
Closest in time.
Normalization propagation: A parametric technique for removing internal covariate shift in deep networks
D. Arpit, Y. Zhou, B. U. Kota, and V. Govindaraju · 2016
Closest in time.
Normalization propagation: A parametric technique for removing internal covariate shift in deep networks
D. Arpit, Y. Zhou, B. U. Kota, and V. Govindaraju · 2016
Closest in time.
Deep residual learning for image recognition
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Closest in time.
Densely connected convolutional networks
G. Huang, Z. Liu, and K. Q. Weinberger · 2016
Closest in time.
Deep Networks with Stochastic Depth
G. Huang, Y. Sun, Z. Liu, D. Sedra, and K. Q. Weinberger · 2016
Closest in time.
Data-dependent initializations of convolutional neural networks
P. Krähenbühl, C. Doersch, J. Donahue, and T. Darrell · 2016
Closest in time.
Trading accuracy for numerical stability: Orthogonalization, biorthogonalization and regularization
T. A. Lahlou and A. V. Oppenheim · 2016
Closest in time.
Visualizing and understanding deep texture representations
T.-Y. Lin and S. Maji · 2016
Closest in time.
Algorithmic Advances in Riemannian Geometry and Applications: For Machine Learning, Computer Vision, Statistics, and Optimization
H. Minh and V. Murino · 2016
Closest in time.
All you need is a good init
D. Mishkin and J. Matas · 2016
Closest in time.
Riemannian preconditioning
B. Mishra and R. Sepulchre · 2016
Closest in time.
Data-dependent path normalization in neural networks
B. Neyshabur, R. Salakhutdinov, and N. Srebro · 2016
Closest in time.
Weight normalization: A simple reparameterization to accelerate training of deep neural networks
T. Salimans and D. P. Kingma · 2016
Closest in time.
Gradual dropin of layers to train very deep neural networks
L. N. Smith, E. M. Hand, and T. Doster · 2016
Closest in time.
Hcp: A flexible cnn framework for multi-label image classification
Y. Wei, W. Xia, M. Lin, J. Huang, B. Ni, J. Dong, Y. Zhao, and S. Yan · 2016
Closest in time.
Metric learning as convex combinations of local models with generalization guarantees
V. Zantedeschi, R. Emonet, and M. Sebban · 2016
Closest in time.