Fetching the paper…
Reading the bibliography…
Recurrent neural networks (RNNs), including long short-term memory (LSTM) RNNs, have produced state-of-the-art results on a variety of speech recognition tasks.
“Long Short-term Memory,”
Sepp Hochreiter and Jürgen Schmidhuber, · 1997
Earlier work this paper cites.
“Learning to Forget: Continual Prediction with LSTM,”
Felix A Gers, Jürgen Schmidhuber, and Fred Cummins, · 2000
Earlier work this paper cites.
“Recurrent Nets that Time and Count,”
F. Gers, J. Schmidhuber, et al., · 2000
Earlier work this paper cites.
Structured Matrices and Polynomials: Unified Superfast Algorithms
V. Pan, · 2001
Earlier work this paper cites.
Introduction to Linear Algebra
G. Strang, · 2009
Earlier work this paper cites.
“Large Scale Distributed Deep Networks,”
J. Dean, G.S. Corrado, R. Monga, K. Chen, M. Devin, Q.V. Le, M.Z. Mao, M. Ranzato, A. Senior, P. Tucker, K. Yang, and A.Y. Ng, · 2012
Earlier work this paper cites.
“Low-Rank Matrix Factorization for Deep Belief Network Training,”
T. N. Sainath, B. Kingsbury, V. Sindhwani, E. Arisoy, and B. Ramabhadran, · 2013
Earlier work this paper cites.
“Restructuring of Deep Neural Network Acoustic Models with Singular Value Decomposition,”
Jian Xue, Jinyu Li, and Yifan Gong, · 2013
Cited alongside, same era.
“Memory-bounded Deep Convolutional Neural Networks,”
M. Collins and P. Kohli, · 2013
Cited alongside, same era.
“Unfolded Recurrent Neural Networks for Speech Recognition,”
G. Saon, H. Soltau, A. Emami, and M. Picheny, · 2014
Cited alongside, same era.
“Long Short-Term Memory Recurrent Neural Network Architectures for Large Scale Acoustic Modeling,”
H. Sak, A. Senior, and F. Beaufays, · 2014
Cited alongside, same era.
“Distilling the Knowledge in a Neural Network,”
G. Hinton, O. Vinyals, and J. Dean, · 2014
Cited alongside, same era.
“Convolutional, Long Short-Term Memory, Fully Connected Deep Neural Networks,”
T. N. Sainath, O. Vinyals, A. Senior, and H. Sak, · 2015
“Deep Learning with Limited Numerical Precision,”
Suyog Gupta, Ankur Agrawal, Kailash Gopalakrishnan, and Pritish Narayanan, · 2015
Later among the works it cites.
“Low-precision Storage for Deep Learning,”
M. Courbariaux, J.-P. David, and Y. Bengio, · 2015
Later among the works it cites.
“Compressing Neural Networks with the Hashing Trick,”
Wenlin Chen, James T Wilson, Stephen Tyree, Kilian Q Weinberger, and Yixin Chen, · 2015
Later among the works it cites.
“Structured Transforms for Small-footprint Deep Learning,”
V. Sindhwani, T. N. Sainath, and S. Kumar, · 2015
Later among the works it cites.
“LSTM: A Search Space Odyssey,”
Klaus Greff, Rupesh Kumar Srivastava, Jan Koutník, Bas R Steunebrink, and Jürgen Schmidhuber, · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Wojciech Zaremba, · 2015
Later among the works it cites.