Fetching the paper…
Reading the bibliography…
This technical report describes an efficient technique for computing the norm of the gradient of the loss function for a neural network with respect to its parameters.
Stochastic optimization with importance sampling
Zhao, Peilin and Zhang, Tong · 2014
Earlier work this paper cites.
Nothing clear enough to list yet.
Nothing clear enough to list yet.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…