Fetching the paper…
Reading the bibliography…
Traditionally, multi-layer neural networks use dot product between the output vector of previous layer and the incoming weight vector as the input to activation function.
Self-organized formation of topologically correct feature maps
Kohonen, Teuvo · 1982
Earlier work this paper cites.
Fast learning in networks of locally-tuned processing units
Moody, John and Darken, Christian J · 1989
Earlier work this paper cites.
A simple weight decay can improve generalization
Krogh, Anders and Hertz, John A · 1991
Earlier work this paper cites.
Gradient-based learning applied to document recognition
LeCun, Yann, Bottou, Léon, Bengio, Yoshua, and Haffner, Patrick · 1998
Earlier work this paper cites.
Gradient flow in recurrent nets: the difficulty of learning long-term dependencies, 2001
Hochreiter, Sepp, Bengio, Yoshua, Frasconi, Paolo, and Schmidhuber, Jürgen · 2001
Earlier work this paper cites.
Modern information retrieval: A brief overview
Singhal, Amit · 2001
Earlier work this paper cites.
Rank, trace-norm and max-norm
Srebro, Nathan and Shraibman, Adi · 2005
Earlier work this paper cites.
Reducing the dimensionality of data with neural networks
Hinton, Geoffrey E and Salakhutdinov, Ruslan R · 2006
Earlier work this paper cites.
A fast learning algorithm for deep belief nets
Hinton, Geoffrey E, Osindero, Simon, and Teh, Yee-Whye · 2006
Earlier work this paper cites.
Introduction to data mining
Tan, Pang-Ning et al · 2006
Earlier work this paper cites.
Learning multiple layers of features from tiny images
Krizhevsky, Alex and Hinton, Geoffrey · 2009
Earlier work this paper cites.
Rectified linear units improve restricted boltzmann machines
Nair, Vinod and Hinton, Geoffrey E · 2010
Cited alongside, same era.
Reading digits in natural images with unsupervised feature learning
Netzer, Yuval, Wang, Tao, Coates, Adam, Bissacco, Alessandro, Wu, Bo, and Ng, Andrew Y · 2011
Cited alongside, same era.
Deep neural networks for acoustic modeling in speech recognition: The shared views of four research groups
Hinton, Geoffrey, Deng, Li, Yu, Dong, Dahl, George E, Mohamed, Abdel-rahman, Jaitly, Navdeep, Senior, Andrew, Vanhoucke, Vincent, Nguyen, Patrick, Sainath, Tara N, et al · 2012
Cited alongside, same era.
Imagenet classification with deep convolutional neural networks
Krizhevsky, Alex, Sutskever, Ilya, and Hinton, Geoffrey E · 2012
Cited alongside, same era.
Efficient backprop
LeCun, Yann A, Bottou, Léon, Orr, Genevieve B, and Müller, Klaus-Robert · 2012
Cited alongside, same era.
Dropout: a simple way to prevent neural networks from overfitting
Srivastava, Nitish, Hinton, Geoffrey E, Krizhevsky, Alex, Sutskever, Ilya, and Salakhutdinov, Ruslan · 2014
Later among the works it cites.
Batch normalization: Accelerating deep network training by reducing internal covariate shift
Ioffe, Sergey and Szegedy, Christian · 2015
Later among the works it cites.
Going deeper with convolutions
Szegedy, Christian, Liu, Wei, Jia, Yangqing, Sermanet, Pierre, Reed, Scott, Anguelov, Dragomir, Erhan, Dumitru, Vanhoucke, Vincent, and Rabinovich, Andrew · 2015
Later among the works it cites.
Tensorflow: Large-scale machine learning on heterogeneous distributed systems
Abadi, Martín, Agarwal, Ashish, Barham, Paul, Brevdo, Eugene, Chen, Zhifeng, Citro, Craig, Corrado, Greg S, Davis, Andy, Dean, Jeffrey, Devin, Matthieu, et al · 2016
Later among the works it cites.
Arpit, Devansh, Zhou, Yingbo, Kota, Bhargava U, and Govindaraju, Venu · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Lin, Min, Chen, Qiang, and Yan, Shuicheng · 2013
Cited alongside, same era.
Rectifier nonlinearities improve neural network acoustic models
Maas, Andrew L, Hannun, Awni Y, and Ng, Andrew Y · 2013
Cited alongside, same era.
Distributed representations of words and phrases and their compositionality
Mikolov, Tomas, Sutskever, Ilya, Chen, Kai, Corrado, Greg S, and Dean, Jeff · 2013
Cited alongside, same era.
Regularization of neural networks using dropconnect
Wan, Li, Zeiler, Matthew, Zhang, Sixin, Cun, Yann L, and Fergus, Rob · 2013
Cited alongside, same era.
Very deep convolutional networks for large-scale image recognition
Simonyan, Karen and Zisserman, Andrew · 2014
Cited alongside, same era.
Later among the works it cites.
Ba, Jimmy Lei, Kiros, Jamie Ryan, and Hinton, Geoffrey E · 2016
Later among the works it cites.
Deep residual learning for image recognition
He, Kaiming, Zhang, Xiangyu, Ren, Shaoqing, and Sun, Jian · 2016
Later among the works it cites.
Normalizing the normalizers: Comparing and extending network normalization schemes
Ren, Mengye, Liao, Renjie, Urtasun, Raquel, Sinz, Fabian H, and Zemel, Richard S · 2016
Later among the works it cites.
Weight normalization: A simple reparameterization to accelerate training of deep neural networks
Salimans, Tim and Kingma, Diederik P · 2016
Later among the works it cites.
Mastering the game of go with deep neural networks and tree search
Silver, David, Huang, Aja, Maddison, Chris J, Guez, Arthur, Sifre, Laurent, Van Den Driessche, George, Schrittwieser, Julian, Antonoglou, Ioannis, Panneershelvam, Veda, Lanctot, Marc, et al · 2016
Later among the works it cites.