Fetching the paper…
Reading the bibliography…
Training a one-node neural network with ReLU activation function (One-Node-ReLU) is a fundamental optimization problem in deep learning.
Test criteria for pearson type iii distributions
MR Mickey, PB Mundle, DN Walker, and AM Glinsk · 1963
Earlier work this paper cites.
Asymptotic properties of non-linear least squares estimators
Robert I Jennrich · 1969
Earlier work this paper cites.
A fast learning algorithm for deep belief nets
Geoffrey E Hinton, Simon Osindero, and Yee-Whye Teh · 2006
Earlier work this paper cites.
Convex piecewise-linear fitting
Alessandro Magnani and Stephen P Boyd · 2009
Earlier work this paper cites.
Efficient learning of generalized linear and single index models with isotonic regression
Sham M Kakade, Varun Kanade, Ohad Shamir, and Adam Kalai · 2011
Earlier work this paper cites.
Fitting piecewise linear continuous functions
Alejandro Toriello and Juan Pablo Vielma · 2012
Earlier work this paper cites.
Understanding deep neural networks with rectified linear units
Raman Arora, Amitabh Basu, Poorya Mianjy, and Anirbit Mukherjee · 2016
Earlier work this paper cites.
Reliably learning the relu in polynomial time
Surbhi Goel, Varun Kanade, Adam Klivans, and Justin Thaler · 2016
Cited alongside, same era.
Understanding deep learning requires rethinking generalization
Chiyuan Zhang, Samy Bengio, Moritz Hardt, Benjamin Recht, and Oriol Vinyals · 2016
Cited alongside, same era.
Globally optimal gradient descent for a convnet with gaussian inputs
Alon Brutzkus and Amir Globerson · 2017
Cited alongside, same era.
When is a convolutional filter easy to learn?
Simon S Du, Jason D Lee, and Yuandong Tian · 2017
Cited alongside, same era.
Gradient descent learns one-hidden-layer cnn: Don’t be afraid of spurious local minima
Simon S Du, Jason D Lee, Yuandong Tian, Barnabas Poczos, and Aarti Singh · 2017
Complexity of training relu neural network
Digvijay Boob, Santanu S. Dey, and Guanghui Lan · 2018
Closest in time.
Relu regression: Complexity, exact and approximation algorithms
Santanu S Dey, Guanyi Wang, and Yao Xie · 2018
Closest in time.
Learning one convolutional layer with overlapping patches
Surbhi Goel, Adam Klivans, and Raghu Meka · 2018
Closest in time.
The computational complexity of training relu (s)
Pasin Manurangsi and Daniel Reichman · 2018
Closest in time.
On the complexity of training a neural network
L. Song, S. Vempala, J. Wilmes, and B. Xei · 2018
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Learning relus via gradient descent
Mahdi Soltanolkotabi · 2017
Cited alongside, same era.
Fitting relus via sgd and quantized sgd
Seyed Mohammadreza Mousavi Kalan, Mahdi Soltanolkotabi, and A Salman Avestimehr · 2019
Closest in time.