Fetching the paper…
Reading the bibliography…
Nonlinearity is crucial to the performance of a deep (neural) network (DN).
Reorthogonalization and stable algorithms for updating the Gram-Schmidt QR factorization
J. W. Daniel, W. B. Gragg, L. Kaufman, and G. W. Stewart · 1976
Earlier work this paper cites.
Image coding using vector quantization: A review
N. M. Nasrabadi and R. A. King · 1988
Earlier work this paper cites.
Assessing a mixture model for clustering with the integrated completed likelihood
C. Biernacki, G. Celeux, and G. Govaert · 2000
Earlier work this paper cites.
CRC Concise Encyclopedia of Mathematics
E. W. Weisstein · 2002
Earlier work this paper cites.
Pattern Recognition and Machine Learning
C. M. Bishop · 2006
Earlier work this paper cites.
Convex piecewise-linear fitting
A. Magnani and S. P. Boyd · 2009
Earlier work this paper cites.
A theoretical analysis of feature pooling in visual recognition
Y. Boureau, J. Ponce, and Y. LeCun · 2010
Cited alongside, same era.
Understanding the difficulty of training deep feedforward neural networks
X. Glorot and Y. Bengio · 2010
Cited alongside, same era.
Vector Quantization and Signal Compression
A. Gersho and R. M. Gray · 2012
Cited alongside, same era.
Multivariate convex regression with adaptive partitioning
L. A. Hannah and D. B. Dunson · 2013
Cited alongside, same era.
Understanding locally competitive networks
R. K. Srivastava, J. Masci, F. Gomez, and J. Schmidhuber · 2014
Cited alongside, same era.
Neural photo editing with introspective adversarial networks
Deep Learning , volume 1
I. Goodfellow, Y. Bengio, and A. Courville · 2016
Later among the works it cites.
L. Huang, X. Liu, B. Lang, A. W. Yu, Y. Wang, and B. Li · 2017
Later among the works it cites.
Searching for activation functions
P. Ramachandran, B. Zoph, and Q. Le · 2017
Later among the works it cites.
Parametric exponential linear unit for deep convolutional neural networks
L. Trottier, P. Gigu, and B. Chaib-draa · 2017
Later among the works it cites.
Deep learning using rectified linear units (ReLU)
A. F. Agarap · 2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A. Brock, T. Lim, J. M. Ritchie, and N. Weston · 2016
Cited alongside, same era.
Mad max: Affine spline insights into deep learning
R. Balestriero and R. Baraniuk
Cited in the paper.
A spline theory of deep networks
R. Balestriero and R. G. Baraniuk
Cited in the paper.
Sigmoid-weighted linear units for neural network function approximation in reinforcement learning
S. Elfwing, E. Uchibe, and K. Doya · 2018
Closest in time.