Fetching the paper…
Reading the bibliography…
Recurrent neural networks (RNNs) achieve cutting-edge performance on a variety of problems.
Comparison of parametric representations for monosyllabic word recognition in continuously spoken sentences
Steven B Davis and Paul Mermelstein · 1990
Earlier work this paper cites.
Building a large annotated corpus of english: The penn treebank
Mitchell P Marcus, Mary Ann Marcinkiewicz, and Beatrice Santorini · 1993
Earlier work this paper cites.
Regression shrinkage and selection via the lasso
Robert Tibshirani · 1996
Earlier work this paper cites.
Convergence of a block coordinate descent method for nondifferentiable minimization
Paul Tseng · 2001
Earlier work this paper cites.
Signal recovery by proximal forward-backward splitting
Patrick L Combettes and Valérie R Wajs · 2005
Earlier work this paper cites.
Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks
Alex Graves, Santiago Fernández, et al · 2006
Earlier work this paper cites.
Pathwise coordinate optimization
Jerome Friedman, Trevor Hastie, et al · 2007
Earlier work this paper cites.
Lattice-based optimization of sequence classification criteria for neural-network acoustic modeling
Brian Kingsbury · 2009
Earlier work this paper cites.
Online dictionary learning for sparse coding
Julien Mairal, Francis Bach, Jean Ponce, and Guillermo Sapiro · 2009
Earlier work this paper cites.
Extensions of recurrent neural network language model
Tomàš Mikolov, S. Kombrink, et al · 2011
Earlier work this paper cites.
Predicting parameters in deep learning
Misha Denil, Babak Shakibi, et al · 2013
Earlier work this paper cites.
Speech recognition with deep recurrent neural networks
Alex Graves, Abdel-rahman Mohamed, et al · 2013
Earlier work this paper cites.
Jitter-adaptive dictionary learning-application to multi-trial neuroelectric signals
Sebastian Hitziger, Maureen Clerc, et al · 2013
Earlier work this paper cites.
Accurate and compact large vocabulary speech recognition on mobile devices
Xin Lei, Andrew W Senior, et al · 2013
Cited alongside, same era.
Low-rank matrix factorization for deep neural network training with high-dimensional output targets
Tara N Sainath, Brian Kingsbury, et al · 2013
Cited alongside, same era.
Restructuring of deep neural network acoustic models with singular value decomposition
Jian Xue, Jinyu Li, and Yifan Gong · 2013
Cited alongside, same era.
Learning phrase representations using RNN encoder-decoder for statistical machine translation
Kyunghyun Cho, Bart Van Merriënboer, et al · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba · 2014
Cited alongside, same era.
Learning acoustic frame labeling for speech recognition with recurrent neural networks
Haşim Sak, Andrew Senior, et al · 2015
Later among the works it cites.
Tensorflow: A system for large-scale machine learning
Martín Abadi, Paul Barham, et al · 2016
Later among the works it cites.
Sequence-level knowledge distillation
Yoon Kim and Alexander M Rush · 2016
Later among the works it cites.
On the compression of recurrent neural networks with an application to lvcsr acoustic modeling for embedded speech recognition
Rohit Prabhavalkar, Ouais Alsharif, et al · 2016
Later among the works it cites.
Annealed sparsity via adaptive and dynamic shrinking
Kai Zhang, Shandian Zhe, et al · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Binbin Lin, Qingyang Li, et al · 2014
Cited alongside, same era.
Long short-term memory recurrent neural network architectures for large scale acoustic modeling
Hasim Sak, Andrew Senior, et al · 2014
Cited alongside, same era.
Recurrent neural network regularization
Wojciech Zaremba, Ilya Sutskever, and Oriol Vinyals · 2014
Cited alongside, same era.
Song Han, Huizi Mao, and William J Dally · 2015
Cited alongside, same era.
Learning both weights and connections for efficient neural network
Song Han, Jeff Pool, et al · 2015
Cited alongside, same era.
Distilling the knowledge in a neural network
Geoffrey Hinton, Oriol Vinyals, and Jeff Dean · 2015
Cited alongside, same era.
Librispeech: an asr corpus based on public domain audio books
Vassil Panayotov, Guoguo Chen, Daniel Povey, and Sanjeev Khudanpur · 2015
Cited alongside, same era.
Yuntao Chen, Naiyan Wang, and Zhaoxiang Zhang · 2017
Later among the works it cites.
A novel image tag completion method based on convolutional neural network
Yanyan Geng, Guohui Zhang, Weizhi Li, Yi Gu, Gaoyuan Liang, Jingbin Wang, Yanbin Wu, Nitin Patil, and Jing-Yan Wang · 2017
Later among the works it cites.
Ese: Efficient speech recognition engine with sparse lstm on fpga
Song Han, Junlong Kang, Huizi Mao, Yiming Hu, Xin Li, Yubin Li, Dongliang Xie, Hong Luo, Song Yao, Yu Wang, et al · 2017
Later among the works it cites.
Deeprebirth: Accelerating deep neural network execution on mobile devices
Dawei Li, Xiaolong Wang, and Deguang Kong · 2017
Later among the works it cites.
Exploring sparsity in recurrent neural networks
Sharan Narang, Gregory Diamos, Shubho Sengupta, and Erich Elsen · 2017
Later among the works it cites.
Show and tell: Lessons learned from the 2015 mscoco image captioning challenge
Oriol Vinyals, Alexander Toshev, et al · 2017
Later among the works it cites.
Alternating multi-bit quantization for recurrent neural networks
Chen Xu, Jianqiang Yao, et al · 2018
Closest in time.