Fetching the paper…
Reading the bibliography…
We characterize the singular values of the linear transformation associated with a standard 2D multi-channel convolutional layer, enabling their efficient computation.
Proximity maps for convex sets
W. Cheney and A. A. Goldstein · 1959
Earlier work this paper cites.
A note on block circulant matrices
C. Chao · 1974
Earlier work this paper cites.
A method for finding projections onto the intersection of convex sets in hilbert spaces
J. P. Boyle and R. L. Dykstra · 1986
Earlier work this paper cites.
Fundamentals of digital image processing
A. K. Jain · 1989
Earlier work this paper cites.
Untersuchungen zu dynamischen neuronalen netzen
S. Hochreiter · 1991
Earlier work this paper cites.
Improving generalization performance using double backpropagation
H. Drucker and Y. Le Cun · 1992
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Y. LeCun, L. Bottou, Y. Bengio, and P. Haffner · 1998
Earlier work this paper cites.
Gradient flow in recurrent nets: the difficulty of learning long-term dependencies, 2001
S. Hochreiter, Y. Bengio, P. Frasconi, J. Schmidhuber, et al · 2001
Earlier work this paper cites.
Alternating projections, 2003
S. Boyd and J. Dattorro · 2003
Cited alongside, same era.
Toeplitz and circulant matrices: A review
R. M. Gray · 2006
Cited alongside, same era.
Matrix Analysis
R. A. Horn and C. R. Johnson · 2012
Cited alongside, same era.
Hessian Schatten-norm regularization for linear inverse problems
S. Lefkimmiatis, J. P. Ward, and M. Unser · 2013
Cited alongside, same era.
Exact solutions to the nonlinear dynamics of learning in deep linear neural networks
A. M. Saxe, J. L. McClelland, and S. Ganguli · 2013
Cited alongside, same era.
Deep Learning
I. Goodfellow, Y. Bengio, and A. Courville · 2016
Cited alongside, same era.
Spectrally-normalized margin bounds for neural networks
P. L. Bartlett, D. J. Foster, and M. J. Telgarsky · 2017
Later among the works it cites.
Parseval networks: Improving robustness to adversarial examples
M. Cisse, P. Bojanowski, E. Grave, Y. Dauphin, and N. Usunier · 2017
Later among the works it cites.
Formal guarantees on the robustness of a classifier against adversarial manipulation
M. Hein and M. Andriushchenko · 2017
Later among the works it cites.
Resurrecting the sigmoid in deep learning through dynamical isometry: theory and practice
J. Pennington, S. Schoenholz, and S. Ganguli · 2017
Later among the works it cites.
Spectral norm regularization for improving the generalizability of deep learning
Y. Yoshida and T. Miyato · 2017
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Cited alongside, same era.
Regularisation of neural networks by enforcing lipschitz continuity
H. Gouk, E. Frank, B. Pfahringer, and M. Cree
Cited in the paper.
MaxGain: Regularisation of neural networks by constraining activation magnitudes
H. Gouk, B. Pfahringer, E. Frank, and M. Cree
Cited in the paper.
T. Miyato, T. Kataoka, M. Koyama, and Y. Yoshida · 2018
Closest in time.
Deep layers as stochastic solvers
Adel Bibi, Bernard Ghanem, Vladlen Koltun, and Rene Ranftl · 2019
Closest in time.