Fetching the paper…
Reading the bibliography…
A proper initialization of the weights in a neural network is critical to its convergence.
Elements of large-sample theory
Erich Leo Lehmann · 2004
Earlier work this paper cites.
A unified architecture for natural language processing: Deep neural networks with multitask learning
Ronan Collobert and Jason Weston · 2008
Earlier work this paper cites.
The tanh transformation
Michael D Godfrey · 2009
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Alex Krizhevsky and Geoffrey Hinton · 2009
Earlier work this paper cites.
Understanding the difficulty of training deep feedforward neural networks
Xavier Glorot and Yoshua Bengio · 2010
Earlier work this paper cites.
Parsing natural scenes and natural language with recursive neural networks
Richard Socher, Cliff C Lin, Chris Manning, and Andrew Y Ng · 2011
Earlier work this paper cites.
Speech recognition with deep recurrent neural networks
Alex Graves, Abdel-rahman Mohamed, and Geoffrey Hinton · 2013
Cited alongside, same era.
Rectifier nonlinearities improve neural network acoustic models
Andrew L Maas, Awni Y Hannun, and Andrew Y Ng · 2013
Cited alongside, same era.
Exact solutions to the nonlinear dynamics of learning in deep linear neural networks
Andrew M Saxe, James L McClelland, and Surya Ganguli · 2013
Cited alongside, same era.
Deep neural networks for object detection
Christian Szegedy, Alexander Toshev, and Dumitru Erhan · 2013
Cited alongside, same era.
Convolutional neural networks for sentence classification
Yoon Kim · 2014
Cited alongside, same era.
Sequence to sequence learning with neural networks
Ilya Sutskever, Oriol Vinyals, and Quoc V Le · 2014
Later among the works it cites.
Deeppose: Human pose estimation via deep neural networks
Alexander Toshev and Christian Szegedy · 2014
Later among the works it cites.
Delving deep into rectifiers: Surpassing human-level performance on imagenet classification
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun · 2015
Later among the works it cites.
Dmytro Mishkin and Jiri Matas · 2015
Later among the works it cites.
Imagenet large scale visual recognition challenge
Olga Russakovsky, Jia Deng, Hao Su, Jonathan Krause, Sanjeev Satheesh, Sean Ma, Zhiheng Huang, Andrej Karpathy, Aditya Khosla, Michael Bernstein, et al · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
David Sussillo and LF Abbott · 2014
Cited alongside, same era.
Cs 231n: Convolutional neural networks for visual recognition, lecture 5, slide 61
Andrej Karpathy, Justin Johnson, and Fei Fei Li · 2016
Later among the works it cites.