Fetching the paper…
Reading the bibliography…
Neural networks with the Rectified Linear Unit (ReLU) nonlinearity are described by a vector of parameters $\theta$, and realized as a piecewise linear continuous function $R_{\theta}: x \in \mathbb R^{d} \mapsto R_{\theta}(x) \in \mathbb R^{k}$.
Same, Same But Different - Recovering Neural Network Quantization Error Through Weight Factorization
Eldad Meller, Alexander Finkelstein, Uri Almog, and Mark Grobman · 1902
Earlier work this paper cites.
Positively Scale-Invariant Flatness of ReLU Neural Networks
Mingyang Yi, Qi Meng, Wei Chen, Zhi-ming Ma, and Tie-Yan Liu · 1903
Earlier work this paper cites.
Data-Free Quantization Through Weight Equalization and Bias Correction
Markus Nagel, Mart van Baalen, Tijmen Blankevoort, and Max Welling · 1906
Earlier work this paper cites.
Uniqueness of the weights for minimal feedforward nets with a given input-output map
Héctor J. Sussmann · 1992
Earlier work this paper cites.
Functionally equivalent feedforward neural networks
Věra Kůrková and Paul C. Kainen · 1993
Earlier work this paper cites.
Uniqueness of weights for neural networks
Francesca Albertini, Eduardo D. Sontag, and Vincent Maillot · 1993
Earlier work this paper cites.
Uniqueness of network parameterizations and faster learning
Paul Kainen, Vera Kurková, Vladik Kreinovich, and Ongard Sirisengtaksin · 1994
Earlier work this paper cites.
Reconstructing a neural net from its output
Charles Fefferman · 1994
Earlier work this paper cites.
Gradient-based learning applied to document recognition
Yann LeCun, Léon Bottou, Yoshua Bengio, and Patrick Haffner · 1998
Earlier work this paper cites.
Imagenet classification with deep convolutional neural networks
Alex Krizhevsky, Ilya Sutskever, and Geoffrey E Hinton · 2012
Earlier work this paper cites.
On the number of response regions of deep feed forward networks with piece-wise linear activations, 2013
Razvan Pascanu, Guido Montufar, and Yoshua Bengio · 2013
Earlier work this paper cites.
On the number of linear regions of deep neural networks
Guido Montúfar, Razvan Pascanu, Kyunghyun Cho, and Yoshua Bengio · 2014
Earlier work this paper cites.
Norm-Based Capacity Control in Neural Networks
Behnam Neyshabur, Ryota Tomioka, and Nathan Srebro · 2015
Cited alongside, same era.
Path-SGD - Path-Normalized Optimization in Deep Neural Networks
Behnam Neyshabur, Ruslan Salakhutdinov, and Nathan Srebro · 2015
Cited alongside, same era.
Path-sgd: Path-normalized optimization in deep neural networks
Behnam Neyshabur, Ruslan Salakhutdinov, and Nathan Srebro · 2015
Cited alongside, same era.
On the expressive power of deep neural networks
Maithra Raghu, Ben Poole, Jon Kleinberg, Surya Ganguli, and Jascha Sohl-Dickstein · 2017
Cited alongside, same era.
Multilinear compressive sensing and an application to convolutional linear networks
Francois Malgouyres and Joseph Landsberg · 2018
Cited alongside, same era.
𝒢 \mathcal{G} -sgd: Optimizing relu neural networks in its positively scale-invariant space, 2018
Scaling-Based Weight Normalization for Deep Neural Networks
Qunyong Yuan and Nanfeng Xiao · 2019
Later among the works it cites.
Data-free quantization through weight equalization and bias correction, 2019
Markus Nagel, Mart van Baalen, Tijmen Blankevoort, and Max Welling · 2019
Later among the works it cites.
Same, same but different - recovering neural network quantization error through weight factorization, 2019
Eldad Meller, Alexander Finkelstein, Uri Almog, and Mark Grobman · 2019
Later among the works it cites.
Positively scale-invariant flatness of ReLU neural networks, 2019
Mingyang Yi, Qi Meng, Wei Chen, Zhi ming Ma, and Tie-Yan Liu · 2019
Later among the works it cites.
Deep relu networks have surprisingly few activation patterns, 2019
Boris Hanin and David Rolnick · 2019
Later among the works it cites.
Functional vs. parametric equivalence of relu networks
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Qi Meng, Shuxin Zheng, Huishuai Zhang, Wei Chen, Zhi-Ming Ma, and Tie-Yan Liu · 2018
Cited alongside, same era.
The shattered gradients problem: If resnets are the answer, then what is the question?, 2018
David Balduzzi, Marcus Frean, Lennox Leary, JP Lewis, Kurt Wan-Duo Ma, and Brian McWilliams · 2018
Cited alongside, same era.
Nonlinear approximation and (deep) relu networks, 2019
I. Daubechies, R. DeVore, S. Foucart, B. Hanin, and G. Petrova · 2019
Cited alongside, same era.
Reverse-engineering deep relu networks, 2019
David Rolnick and Konrad P. Kording · 2019
Cited alongside, same era.
Robust and resource efficient identification of two hidden layer neural networks, 2019
Massimo Fornasier, Timo Klock, and Michael Rauchensteiner · 2019
Cited alongside, same era.
G-SGD - Optimizing ReLU Neural Networks in its Positively Scale-Invariant Space
Qi Meng, Shuxin Zheng, Huishuai Zhang, Wei Chen 0034, Qiwei Ye, Zhi-Ming Ma, Nenghai Yu, and Tie-Yan Liu · 2019
Cited alongside, same era.
Equi-normalization of Neural Networks
Pierre Stock, Benjamin Graham, Remi Gribonval, and Hervé Jégou · 2019
Cited alongside, same era.
Mary Phuong and Christoph H Lampert · 2019
Later among the works it cites.
Neural network approximation, 2020
Ronald DeVore, Boris Hanin, and Guergana Petrova · 2020
Later among the works it cites.
Functional vs. parametric equivalence of ReLU networks
Mary Phuong and Christoph H Lampert · 2020
Later among the works it cites.
On the stable recovery of deep structured linear networks under sparsity constraints
Francois Malgouyres · 2020
Later among the works it cites.
Cryptanalytic extraction of neural network models, 2020
Nicholas Carlini, Matthew Jagielski, and Ilya Mironov · 2020
Later among the works it cites.
Learning compositional functions via multiplicative weight updates
Jeremy Bernstein, Jiawei Zhao, Markus Meister, Ming-Yu Liu, Anima Anandkumar, and Yisong Yue · 2020
Later among the works it cites.