Fetching the paper…
Reading the bibliography…
Most existing sequence labelling models rely on a fixed decomposition of a target sequence into a sequence of basic units.
Large-vocabulary speaker-independent continuous speech recognition using hmm
K-F Lee and H-W Hon · 1988
Earlier work this paper cites.
Segmenting speech without a lexicon: The roles of phonotactics and speech source
Timothy Andrew Cartwright and Michael R Brent · 1994
Earlier work this paper cites.
Phoneme-grapheme based speech recognition system
Mathew Magimai Doss, Todd A Stephenson, Hervé Bourlard, and Samy Bengio · 2003
Earlier work this paper cites.
Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks
Alex Graves, Santiago Fernández, Faustino Gomez, and Jürgen Schmidhuber · 2006
Earlier work this paper cites.
Contextual dependencies in unsupervised word segmentation
Sharon Goldwater, Thomas L Griffiths, and Mark Johnson · 2006
Earlier work this paper cites.
Learning phrase representations using rnn encoder-decoder for statistical machine translation
Kyunghyun Cho, Bart Van Merriënboer, Caglar Gulcehre, Dzmitry Bahdanau, Fethi Bougares, Holger Schwenk, and Yoshua Bengio · 2014
Earlier work this paper cites.
Sequence to sequence learning with neural networks
Ilya Sutskever, Oriol Vinyals, and Quoc V Le · 2014
Earlier work this paper cites.
Deep speech: Scaling up end-to-end speech recognition
Awni Y. Hannun, Carl Case, Jared Casper, Bryan Catanzaro, Greg Diamos, Erich Elsen, Ryan Prenger, Sanjeev Satheesh, Shubho Sengupta, Adam Coates, and Andrew Y. Ng · 2014
Earlier work this paper cites.
On using very large target vocabulary for neural machine translation
Jean Sébastien, Kyunghyun Cho, Roland Memisevic, and Yoshua Bengio · 2015
Earlier work this paper cites.
Grammar as a foreign language
Oriol Vinyals, Łukasz Kaiser, Terry Koo, Slav Petrov, Ilya Sutskever, and Geoffrey Hinton · 2015
Cited alongside, same era.
Attention-based models for speech recognition
Jan K Chorowski, Dzmitry Bahdanau, Dmitriy Serdyuk, Kyunghyun Cho, and Yoshua Bengio · 2015
Cited alongside, same era.
Deep speech 2: End-to-end speech recognition in english and mandarin
Dario Amodei, Rishita Anubhai, Eric Battenberg, Carl Case, Jared Casper, Bryan Catanzaro, Jingdong Chen, Mike Chrzanowski, Adam Coates, Greg Diamos, et al · 2015
Cited alongside, same era.
William Chan, Navdeep Jaitly, Quoc V Le, and Oriol Vinyals · 2015
Cited alongside, same era.
Fast and accurate recurrent neural network acoustic models for speech recognition
Hasim Sak, Andrew W. Senior, Kanishka Rao, and Françoise Beaufays · 2015
Wav2letter: an end-to-end convnet-based speech recognition system
Ronan Collobert, Christian Puhrsch, and Gabriel Synnaeve · 2016
Later among the works it cites.
Neural speech recognizer: Acoustic-to-word lstm model for large vocabulary speech recognition
Hagen Soltau, Hank Liao, and Hasim Sak · 2016
Later among the works it cites.
Dense prediction on sequences with time-dilated convolutions for speech recognition
Tom Sercu and Vaibhava Goel · 2016
Later among the works it cites.
Achieving open vocabulary neural machine translation with hybrid word-character models
Minh-Thang Luong and Christopher D Manning · 2016
Later among the works it cites.
Joint ctc-attention based end-to-end speech recognition using multi-task learning
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Eesen: End-to-end speech recognition using deep rnn models and wfst-based decoding
Yajie Miao, Mohammad Gowayyed, and Florian Metze · 2015
Cited alongside, same era.
Listen, attend and spell: A neural network for large vocabulary conversational speech recognition
William Chan, Navdeep Jaitly, Quoc Le, and Oriol Vinyals · 2016
Cited alongside, same era.
End-to-end attention-based large vocabulary speech recognition
Dzmitry Bahdanau, Jan Chorowski, Dmitriy Serdyuk, Yoshua Bengio, et al · 2016
Cited alongside, same era.
The microsoft 2016 conversational speech recognition system
W Xiong, J Droppo, X Huang, F Seide, M Seltzer, A Stolcke, D Yu, and G Zweig · 2016
Cited alongside, same era.
Google’s neural machine translation system: Bridging the gap between human and machine translation
Yonghui Wu, Mike Schuster, Zhifeng Chen, Quoc V. Le, Mohammad Norouzi, Wolfgang Macherey, Maxim Krikun, Yuan Cao, Qin Gao, Klaus Macherey, Jeff Klingner, Apurva Shah, Melvin Johnson, Xiaobing Liu, Lukasz Kaiser, Stephan Gouws, Yoshikiyo Kato, Taku Kudo, Hideto Kazawa, Keith Stevens, George Kurian, Nishant Patil, Wei Wang, Cliff Young, Jason Smith, Jason Riesa, Alex Rudnick, Oriol Vinyals, Gregory S. Corrado, Macduff Hughes, and Jeffrey Dean
Cited in the paper.
Latent sequence decompositions
William Chan, Yu Zhang, Quoc Le, and Navdeep Jaitly
Cited in the paper.
Google’s neural machine translation system: Bridging the gap between human and machine translation
Yonghui Wu, Mike Schuster, Zhifeng Chen, Quoc V Le, Mohammad Norouzi, Wolfgang Macherey, Maxim Krikun, Yuan Cao, Qin Gao, Klaus Macherey, et al
Cited in the paper.
Suyoun Kim, Takaaki Hori, and Shinji Watanabe · 2016
Later among the works it cites.
Sequence-level knowledge distillation
Yoon Kim and Alexander M Rush · 2016
Later among the works it cites.
Very deep convolutional networks for end-to-end speech recognition
Yu Zhang, William Chan, and Navdeep Jaitly · 2016
Later among the works it cites.
Purely sequence-trained neural networks for asr based on lattice-free mmi
Daniel Povey, Vijayaditya Peddinti, Daniel Galvez, Pegah Ghahrmani, Vimal Manohar, Xingyu Na, Yiming Wang, and Sanjeev Khudanpur · 2016
Later among the works it cites.
Towards better decoding and language model integration in sequence to sequence models
Jan Chorowski and Jaitly Navdeep · 2016
Later among the works it cites.