Fetching the paper…
Reading the bibliography…
Sequence-to-sequence models have shown success in end-to-end speech recognition.
“A tutorial on hidden Markov models and selected applications in speech recognition”
Lawrence˜R. Rabiner · 1989
Earlier work this paper cites.
“Hierarchical Recurrent Neural Networks for Long-Term Dependencies”
Salah Hihi and Yoshua Bengio · 1996
Earlier work this paper cites.
“Long Short-Term Memory”
Sepp Hochreiter and Jurgen Schmidhuber · 1997
Earlier work this paper cites.
“Gradient-based learning applied to document recognition”
Yann LeCun, Leon Bottou, Yoshua Bengio and Patrick Haffner · 1998
Earlier work this paper cites.
“Practical Variational Inference for Neural Networks”
Alex Graves · 2011
Earlier work this paper cites.
“Sequence Transduction with Recurrent Neural Networks”
Alex Graves · 2012
Earlier work this paper cites.
“Deep Neural Networks for Acoustic Modeling in Speech Recognition: The Shared Views of Four Research Groups”
Geoffrey Hinton et al · 2012
Earlier work this paper cites.
“Exploring Convolutional Neural Network Structures and Optimization Techniques for Speech Recognition”
Ossama Abdel-Hamid, Li Deng and Dong Yu · 2013
Earlier work this paper cites.
“Deep Convolutional Neural Networks for LVCSR”
Tara Sainath, Abdel-rahman Mohamed, Brian Kingsbury and Bhuvana Ramabhadran · 2013
Earlier work this paper cites.
“Improvements to Deep Convolutional Neural Networks for LVCSR”
Tara Sainath et al · 2013
Earlier work this paper cites.
“Network In Network”
Min Lin, Qiang Chen and Shuicheng Yan · 2013
Earlier work this paper cites.
“Hybrid Speech Recognition with Bidirectional LSTM”
Alex Graves, Navdeep Jaitly and Abdel-rahman Mohamed · 2013
Cited alongside, same era.
“Towards End-to-End Speech Recognition with Recurrent Neural Networks”
Alex Graves and Navdeep Jaitly · 2014
Cited alongside, same era.
“Neural Machine Translation by Jointly Learning to Align and Translate”
Dzmitry Bahdanau, Kyunghyun Cho and Yoshua Bengio · 2015
Cited alongside, same era.
“Attention-Based Models for Speech Recognition”
Jan Chorowski et al · 2015
Cited alongside, same era.
“Very Deep Convolutional Networks for Large-Scale Image Recognition”
Karen Simonyan and Andrew Zisserman · 2015
Cited alongside, same era.
“Going Deeper with Convolutions”
Christian Szegedy et al · 2015
Cited alongside, same era.
“Listen, Attend and Spell: A Neural Network for Large Vocabulary Conversational Speech Recognition”
William Chan, Navdeep Jaitly, Quoc Le and Oriol Vinyals · 2016
Closest in time.
“End-to-end Attention-based Large Vocabulary Speech Recognition”
Dzmitry Bahdanau et al · 2016
Closest in time.
“Task Loss Estimation for Sequence Prediction”
Dzmitry Bahdanau et al · 2016
Closest in time.
“On Online Attention-based Speech Recognition and Joint Mandarin Character-Pinyin Training”
William Chan and Ian Lane · 2016
Closest in time.
“Deep Convolutional Neural Networks for Acoustic Modeling in Low Resource Languages”
William Chan and Ian Lane · 2016
Closest in time.
“Very deep multilingual convolutional neural networks for LVCSR”
Tom Sercu, Christian Puhrsch, Brian Kingsbury and Yann LeCun · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“Batch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Shift”
Sergey Ioffe and Christian Szegedy · 2015
Cited alongside, same era.
“Convolutional LSTM Network: A Machine Learning Approach for Precipitation Nowcasting”
Xingjian Shi et al · 2015
Cited alongside, same era.
“Training Very Deep Networks”
Rupesh Srivastava, Klaus Greff and J“”urgen Schmidhuber · 2015
Cited alongside, same era.
“Adam: A Method for Stochastic Optimization”
Diederik Kingma and Jimmy Ba · 2015
Cited alongside, same era.
“TensorFlow: Large-Scale Machine Learning on Heterogeneous Systems” Software available from tensorflow.org, 2015
Mart“’n Abadi et al · 2015
Cited alongside, same era.
Closest in time.
“Advances in Very Deep Convolutional Neural Networks for LVCSR”
Tom Sercu and Vaibhava Goel · 2016
Closest in time.
“Deep Speech 2: End-to-End Speech Recognition in English and Mandarin”
Dario Amodei et al · 2016
Closest in time.
“Deep Residual Learning for Image Recognition”
Kaiming He, Xiangyu Zhang, Shaoqing Ren and Jian Sun · 2016
Closest in time.
“Highway Long Short-Term Memory RNNs for Distant Speech Recognition”
Yu Zhang et al · 2016
Closest in time.
“Grid long short-term memory”
Nal Kalchbrenner, Ivo Danihelka and Alex Graves · 2016
Closest in time.