Fetching the paper…
Reading the bibliography…
Deep learning has dramatically improved the performance of speech recognition systems through learning hierarchies of features optimized for the task at hand.
A. Graves, S. Fernández, F. Gomez, and J. Schmidhuber, “Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,” in
2006
Earlier work this paper cites.
M. Ranzato, F. J. Huang, Y. L. Boureau, and Y. LeCun, “Unsupervised learning of invariant feature hierarchies with applications to object recognition,” in
2007
Earlier work this paper cites.
N. Jaitly and G. E. Hinton, “Learning a better representation of speech soundwaves using restricted boltzmann machines.” in
2011
Earlier work this paper cites.
G. Hinton, L. Deng, D. Yu, G. Dahl, A. rahman Mohamed, N. Jaitly, A. Senior, V. Vanhoucke, P. Nguyen, T. Sainath, and B. Kingsbury, “Deep neural networks for acoustic modeling in speech recognition,”
2012
Earlier work this paper cites.
C. Farabet, C. Couprie, L. Najman, and Y. LeCun, “Learning hierarchical features for scene labeling,”
2013
Earlier work this paper cites.
K. Heafield, I. Pouzyrevsky, J. H. Clark, and P. Koehn, “Scalable modified Kneser-Ney language model estimation,” in
2013
Earlier work this paper cites.
Z. Tüske, P. Golik, R. Schlüter, and H. Ney, “Acoustic modeling with deep neural networks using raw time signal for lvcsr,” in
2014
Cited alongside, same era.
S. Dieleman and B. Schrauwen, “End-to-end learning for music audio,” in
2014
Cited alongside, same era.
2014
Cited alongside, same era.
N. Neverova, C. Wolf, G. W. Taylor, and F. Nebout, “Multi-scale deep learning for gesture detection and localization,” in
2014
Cited alongside, same era.
2014
Cited alongside, same era.
2015
Later among the works it cites.
D. Palaz, R. Collobert
2015
Later among the works it cites.
T. N. Sainath, R. J. Weiss, A. Senior, K. W. Wilson, and O. Vinyals, “Learning the speech front-end with raw waveform cldnns,” in
2015
Later among the works it cites.
P. Golik, Z. Tüske, R. Schlüter, and H. Ney, “Convolutional neural networks for acoustic modeling of raw time signal in lvcsr,” in
2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Koutník, K. Greff, F. J. Gomez, and J. Schmidhuber, “A clockwork RNN,”
2014
Cited alongside, same era.
2015
Later among the works it cites.
2015
Later among the works it cites.