Fetching the paper…
Reading the bibliography…
Neuromorphic hardware tends to pose limits on the connectivity of deep networks that one can run on them.
Learning internal representations by error propagation
Rumelhart, D. E., Hinton, G. E., and Williams, R. J. (1985) · 1985
Earlier work this paper cites.
Speaker-independent phone recognition using hidden markov models
K-F Lee and H-W Hon · 1989
Earlier work this paper cites.
Bayesian training of backpropagation networks by the hybrid monte carlo method
Radford M Neal · 1992
Earlier work this paper cites.
Bayesian training of backpropagation networks by the hybrid monte carlo method
Neal, R. M. (1992) · 1992
Earlier work this paper cites.
Regression shrinkage and selection via the lasso
Robert Tibshirani · 1996
Earlier work this paper cites.
Sequential monte carlo methods to train neural network models
Joao FG de Freitas, Mahesan Niranjan, Andrew H. Gee, and Arnaud Doucet · 2000
Earlier work this paper cites.
Handbook of Stochastic Methods
Gardiner, C.W. (2004) · 2004
Earlier work this paper cites.
Framewise phoneme classification with bidirectional LSTM and other neural network architectures
Alex Graves and Jürgen Schmidhuber · 2005
Earlier work this paper cites.
Transient and persistent dendritic spines in the neocortex in vivo
Anthony JGD Holtmaat, Joshua T Trachtenberg, Linda Wilbrecht, Gordon M Shepherd, Xiaoqun Zhang, Graham W Knott, and Karel Svoboda · 2005
Earlier work this paper cites.
Pattern recognition and machine learning
Christopher M Bishop · 2006
Earlier work this paper cites.
Axons and synaptic boutons are highly dynamic in adult visual cortex
Dan D Stettler, Homare Yamahachi, Wu Li, Winfried Denk, and Charles D Gilbert · 2006
Earlier work this paper cites.
A wafer-scale neuromorphic hardware system for large-scale neural modeling
Johannes Schemmel, Daniel Brüderle, Andreas Grübl, Matthias Hock, Karlheinz Meier, and Sebastian Millner · 2010
Earlier work this paper cites.
Bayesian learning via stochastic gradient langevin dynamics
Max Welling and Yee W Teh · 2011
Cited alongside, same era.
Lecture 6.5-RMSProp: Divide the gradient by a running average of its recent magnitude
Tijmen Tieleman and Geoffrey Hinton · 2012
Cited alongside, same era.
Exploiting sparseness in deep neural networks for large vocabulary speech recognition
D. Yu, F. Seide, G. Li, and L. Deng · 2012
Cited alongside, same era.
Speech recognition with deep recurrent neural networks
Alex Graves, Abdel-rahman Mohamed, and Geoffrey Hinton · 2013
Cited alongside, same era.
Stochastic gradient riemannian langevin dynamics on the probability simplex
Sam Patterson and Yee Whye Teh · 2013
Cited alongside, same era.
Deep fried convnets
Zichao Yang, Marcin Moczulski, Misha Denil, Nando de Freitas, Alex Smola, Le Song, and Ziyu Wang · 2015
Later among the works it cites.
Network plasticity as Bayesian inference
Kappel, D., Habenschuss, S., Legenstein, R., and Maass, W. (2015) · 2015
Later among the works it cites.
Bridging the gap between stochastic gradient mcmc and stochastic optimization
Changyou Chen, David Carlson, Zhe Gan, Chunyuan Li, and Lawrence Carin · 2016
Later among the works it cites.
SqueezeNet: AlexNet-level accuracy with 50x fewer parameters and < < 0.5 MB model size
Forrest N Iandola, Song Han, Matthew W Moskewicz, Khalid Ashraf, William J Dally, and Kurt Keutzer · 2016
Later among the works it cites.
Training skinny deep neural networks with iterative hard thresholding methods
Xiaojie Jin, Xiaotong Yuan, Jiashi Feng, and Shuicheng Yan · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Maxwell D. Collins and Pushmeet Kohli · 2014
Cited alongside, same era.
The spinnaker project
Steve B Furber, Francesco Galluppi, Steve Temple, and Luis A Plana · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik Kingma and Jimmy Ba · 2014
Cited alongside, same era.
A million spiking-neuron integrated circuit with a scalable communication network and interface
Paul A Merolla, John V Arthur, Rodrigo Alvarez-Icaza, Andrew S Cassidy, Jun Sawada, Filipp Akopyan, Bryan L Jackson, Nabil Imam, Chen Guo, Yutaka Nakamura, et al · 2014
Cited alongside, same era.
Impermanence of dendritic spines in live adult ca1 hippocampus
Alessio Attardo, James E Fitzgerald, and Mark J Schnitzer · 2015
Cited alongside, same era.
Network Plasticity as Bayesian Inference
David Kappel, Stefan Habenschuss, Robert Legenstein, and Wolfgang Maass · 2015
Cited alongside, same era.
Data-free parameter pruning for deep neural networks
Suraj Srinivas and R. Venkatesh Babu · 2015
Cited alongside, same era.
A stable brain from unstable components: Emerging concepts, implications for neural computation
Anna R Chambers and Simon Rumpel · 2017
Closest in time.
LSTM: A search space odyssey
Klaus Greff, Rupesh K Srivastava, Jan Koutník, Bas R Steunebrink, and Jürgen Schmidhuber · 2017
Closest in time.
ESE: Efficient speech recognition engine with sparse LSTM on FPGA
Song Han, Junlong Kang, Huizi Mao, Yiming Hu, Xin Li, Yubin Li, Dongliang Xie, Hong Luo, Song Yao, Yu Wang, et al · 2017
Closest in time.
In-datacenter performance analysis of a tensor processing unit
Norman P Jouppi, Cliff Young, Nishant Patil, David Patterson, Gaurav Agrawal, Raminder Bajwa, Sarah Bates, Suresh Bhatia, Nan Boden, Al Borchers, et al · 2017
Closest in time.
Intrinsic volatility of synaptic connections - a challenge to the synaptic trace theory of memory
Gianluigi Mongillo, Simon Rumpel, and Yonatan Loewenstein · 2017
Closest in time.
Exploring sparsity in recurrent neural networks
Sharan Narang, Gregory Diamos, Shubho Sengupta, and Erich Elsen · 2017
Closest in time.
Reward-based stochastic self-configuration of neural circuits
Kappel, D., Legenstein, R., Habenschuss, S., Hsieh, M., and Maass, W. (2017) · 2017
Closest in time.