Fetching the paper…
Reading the bibliography…
One of the most difficult speech recognition tasks is accurate recognition of human to human communication.
P. F. Brown, P. V. Desouza, R. L. Mercer, V. J. D. Pietra, and J. C. Lai, “Class-based n-gram models of natural language,”
1992
Earlier work this paper cites.
R. P. Lippmann, “Speech recognition by machines and humans,”
1997
Earlier work this paper cites.
J. Fiscus, W. M. Fisher, A. F. Martin, M. A. Przybocki, and D. S. Pallett, “2000 nist evaluation of conversational speech recognition over the telephone: English and mandarin performance results,” in
2000
Earlier work this paper cites.
D. Povey, D. Kanevsky, B. Kingsbury, B. Ramabhadran, G. Saon, and K. Visweswariah, “Boosted MMI for model and feature-space discriminative training,” in
2008
Earlier work this paper cites.
R. Collobert, K. Kavukcuoglu, and C. Farabet, “Torch7: A matlab-like environment for machine learning,” in
2011
Earlier work this paper cites.
H. Su, G. Li, D. Yu, and F. Seide, “Error back propagation for sequence training of context-dependent deep networks for conversational speech transcription,”
2013
Earlier work this paper cites.
K. Simonyan and A. Zisserman, “Very deep convolutional networks for large-scale image recognition,”
2014
Earlier work this paper cites.
H. Soltau, G. Saon, and T. N. Sainath, “Joint training of convolutional and non-convolutional neural networks,”
2014
Earlier work this paper cites.
W. Zaremba, I. Sutskever, and O. Vinyals, “Recurrent neural network regularization,”
2014
Earlier work this paper cites.
D. Kingma and J. Ba, “ADAM: A method for stochastic optimization,”
2014
Earlier work this paper cites.
G. Saon, H.-K. Kuo, S. Rennie, and M. Picheny, “The IBM 2015 English conversational speech recognition system,” in
2015
Earlier work this paper cites.
2015
Cited alongside, same era.
H. Sak, A. Senior, K. Rao, O. Irsoy, A. Graves, F. Beaufays, and J. Schalkwyk, “Learning acoustic frame labeling for speech recognition with recurrent neural networks,” in
2015
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,”
2015
Cited alongside, same era.
2015
Cited alongside, same era.
S. Zagoruyko and N. Komodakis, “Wide residual networks,”
2016
Later among the works it cites.
K. He, X. Zhang, S. Ren, and J. Sun, “Identity mappings in deep residual networks,”
2016
Later among the works it cites.
T. Sercu, C. Puhrsch, B. Kingsbury, and Y. LeCun, “Very deep multilingual convolutional neural networks for lvcsr,”
2016
Later among the works it cites.
T. Sercu and V. Goel, “Advances in very deep convolutional neural networks for lvcsr,”
2016
Later among the works it cites.
——, “Dense prediction on sequences with time-dilated convolutions for speech recognition,”
2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2016
Cited alongside, same era.
G. Saon, T. Sercu, S. Rennie, and H.-K. Kuo, “The IBM 2016 English conversational speech recognition system,” in
2016
Cited alongside, same era.
K. He, X. Zhang, S. Ren, and J. Sun, “Deep residual learning for image recognition,” in
2016
Cited alongside, same era.
Y. Ganin, E. Ustinova, H. Ajakan, P. Germain, H. Larochelle, F. Laviolette, M. Marchand, and V. Lempitsky, “Domain-adversarial training of neural networks,”
2016
Cited alongside, same era.
Y. Shinohara, “Adversarial multi-task learning of deep neural networks for robust speech recognition,”
2016
Cited alongside, same era.
2016
Cited alongside, same era.
2016
Later among the works it cites.
P. Ghahremani and J. Droppo, “Self-stabilized deep neural network,” in
2016
Later among the works it cites.
2017
Closest in time.
Y. Zhang, W. Chan, and N. Jaitly, “Very deep convolutional networks for end-to-end speech recognition,”
2017
Closest in time.
G. Kurata, A. Sethy, B. Ramabhadran, and G. Saon, “Empirical exploration of LSTM and CNN language models for speech recognition,”
2017
Closest in time.