Fetching the paper…
Reading the bibliography…
For most deep learning algorithms training is notoriously time consuming.
Multiplierless multilayer feedforward neural network design suitable for continuous input-output mapping
Kwan, Hon Keung and Tang, CZ · 1993
Earlier work this paper cites.
Fast neural networks without multipliers
Marchesi, Michele, Orlandi, Gianni, Piazza, Francesco, and Uncini, Aurelio · 1993
Earlier work this paper cites.
Device for generating binary sequences for stochastic computing
van Daalen, Max, Jeavons, Pete, Shawe-Taylor, John, and Cohen, Dave · 1993
Earlier work this paper cites.
Generating binary sequences for stochastic computing
Jeavons, Peter, Cohen, David A., and Shawe-Taylor, John · 1994
Earlier work this paper cites.
Backpropagation without multiplication
Simard, Patrice Y and Graf, Hans Peter · 1994
Earlier work this paper cites.
Gradient-based learning applied to document recognition
LeCun, Yann, Bottou, Léon, Bengio, Yoshua, and Haffner, Patrick · 1998
Earlier work this paper cites.
Stochastic bit-stream neural networks
Burge, Peter S., van Daalen, Max R., Rising, Barry J. P., and Shawe-Taylor, John S · 1999
Cited alongside, same era.
Learning multiple layers of features from tiny images, 2009
Krizhevsky, Alex and Hinton, Geoffrey · 2009
Cited alongside, same era.
Reading digits in natural images with unsupervised feature learning
Netzer, Yuval, Wang, Tao, Coates, Adam, Bissacco, Alessandro, Wu, Bo, and Ng, Andrew Y · 2011
Cited alongside, same era.
Theano: new features and speed improvements
Bastien, Frédéric, Lamblin, Pascal, Pascanu, Razvan, Bergstra, James, Goodfellow, Ian J., Bergeron, Arnaud, Bouchard, Nicolas, and Bengio, Yoshua · 2012
Cited alongside, same era.
Imagenet classification with deep convolutional neural networks
Krizhevsky, Alex, Sutskever, Ilya, and Hinton, Geoffrey E · 2012
Cited alongside, same era.
Building high-level features using large scale unsupervised learning
Le, Quoc V · 2013
Cheng, Zhiyong, Soudry, Daniel, Mao, Zexi, and Lan, Zhenzhong · 2015
Closest in time.
Binaryconnect: Training deep neural networks with binary weights during propagations
Courbariaux, Matthieu, Bengio, Yoshua, and David, Jean-Pierre · 2015
Closest in time.
On using monolingual corpora in neural machine translation
Gulcehre, Caglar, Firat, Orhan, Xu, Kelvin, Cho, Kyunghyun, Barrault, Loic, Lin, Huei-Chi, Bougares, Fethi, Schwenk, Holger, and Bengio, Yoshua · 2015
Closest in time.
Bitwise neural networks
Kim, Minje and Paris, Smaragdis · 2015
Closest in time.
Computational cost reduction in learned transform classifications
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Machado, Emerson Lopes, Miosso, Cristiano Jacques, von Borries, Ricardo, Coutinho, Murilo, Berger, Pedro de Azevedo, Marques, Thiago, and Jacobi, Ricardo Pezzuol · 2015
Closest in time.
Adding gradient noise improves learning for very deep networks
Neelakantan, Arvind, Vilnis, Luke, Le, Quoc V, Sutskever, Ilya, Kaiser, Lukasz, Kurach, Karol, and Martens, James · 2015
Closest in time.