Fetching the paper…
Reading the bibliography…
Despite the success of sequence-to-sequence approaches in automatic speech recognition (ASR) systems, the models still suffer from several problems, mainly due to the mismatch between the training and inference conditions.
“A learning algorithm for continually running fully recurrent neural networks,”
Ronald J Williams and David Zipser, · 1989
Earlier work this paper cites.
“Simple statistical gradient-following algorithms for connectionist reinforcement learning,”
Ronald J Williams, · 1992
Earlier work this paper cites.
“The design for the Wall Street Journal-based CSR corpus,”
Douglas B. Paul and Janet M. Baker, · 1992
Earlier work this paper cites.
Introduction to Reinforcement Learning
Richard S. Sutton and Andrew G. Barto, · 1998
Earlier work this paper cites.
“Reinforcement learning for spoken dialogue systems,”
Satinder P Singh, Michael J Kearns, Diane J Litman, and Marilyn A Walker, · 2000
Earlier work this paper cites.
“Variance reduction techniques for gradient estimates in reinforcement learning,”
Evan Greensmith, Peter L Bartlett, and Jonathan Baxter, · 2004
Earlier work this paper cites.
“The Kaldi speech recognition toolkit,”
Daniel Povey, Arnab Ghoshal, Gilles Boulianne, Lukas Burget, Ondrej Glembek, Nagendra Goel, Mirko Hannemann, Petr Motlicek, Yanmin Qian, Petr Schwarz, Jan Silovsky, Georg Stemmer, and Karel Vesely, · 2011
Earlier work this paper cites.
Supervised sequence labelling with recurrent neural networks
Alex Graves et al., · 2012
Earlier work this paper cites.
“Reinforcement learning in robotics: A survey,”
Jens Kober and Jan Peters, · 2012
Earlier work this paper cites.
“Sequence to sequence learning with neural networks,”
Ilya Sutskever, Oriol Vinyals, and Quoc V Le, · 2014
Earlier work this paper cites.
“Neural machine translation by jointly learning to align and translate,”
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio, · 2014
Cited alongside, same era.
“Neural variational inference and learning in belief networks,”
Andriy Mnih and Karol Gregor, · 2014
Cited alongside, same era.
“First-pass large vocabulary continuous speech recognition using bi-directional recurrent DNNs,”
Awni Y Hannun, Andrew L Maas, Daniel Jurafsky, and Andrew Y Ng, · 2014
Cited alongside, same era.
“Adam: A method for stochastic optimization,”
Diederik Kingma and Jimmy Ba, · 2014
Cited alongside, same era.
“Show and tell: A neural image caption generator,”
Oriol Vinyals, Alexander Toshev, Samy Bengio, and Dumitru Erhan, · 2015
Cited alongside, same era.
“Listen, attend and spell: A neural network for large vocabulary conversational speech recognition,”
William Chan, Navdeep Jaitly, Quoc Le, and Oriol Vinyals, · 2016
Later among the works it cites.
Yonghui Wu, Mike Schuster, Zhifeng Chen, Quoc V Le, Mohammad Norouzi, Wolfgang Macherey, Maxim Krikun, Yuan Cao, Qin Gao, Klaus Macherey, et al., · 2016
Later among the works it cites.
“Minimum risk training for neural machine translation,”
Shiqi Shen, Yong Cheng, Zhongjun He, Wei He, Hua Wu, Maosong Sun, and Yang Liu, · 2016
Later among the works it cites.
“Mastering the game of go with deep neural networks and tree search,”
David Silver, Aja Huang, Chris J Maddison, Arthur Guez, Laurent Sifre, George Van Den Driessche, Julian Schrittwieser, Ioannis Antonoglou, Veda Panneershelvam, Marc Lanctot, et al., · 2016
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“Empirical evaluation of rectified activations in convolutional network,”
Bing Xu, Naiyan Wang, Tianqi Chen, and Mu Li, · 2015
Cited alongside, same era.
“Human-level control through deep reinforcement learning,”
Volodymyr Mnih, Koray Kavukcuoglu, David Silver, Andrei A. Rusu, Joel Veness, Marc G. Bellemare, Alex Graves, Martin Riedmiller, Andreas K. Fidjeland, Georg Ostrovski, Stig Petersen, Charles Beattie, Amir Sadik, Ioannis Antonoglou, Helen King, Dharshan Kumaran, Daan Wierstra, Shane Legg, and Demis Hassabis, · 2015
Cited alongside, same era.
“Sequence level training with recurrent neural networks,”
Marc Aurelio Ranzato, Sumit Chopra, Michael Auli, and Wojciech Zaremba, · 2015
Cited alongside, same era.
“End-to-end attention-based large vocabulary speech recognition,”
Dzmitry Bahdanau, Jan Chorowski, Dmitriy Serdyuk, Philemon Brakel, and Yoshua Bengio, · 2016
Cited alongside, same era.
Jiwei Li, Will Monroe, Alan Ritter, Michel Galley, Jianfeng Gao, and Dan Jurafsky, · 2016
Later among the works it cites.
“Attention-based wav2text with feature transfer learning,”
Andros Tjandra, Sakriani Sakti, and Satoshi Nakamura, · 2017
Closest in time.
“Joint CTC-attention based end-to-end speech recognition using multi-task learning,”
Suyoun Kim, Takaaki Hori, and Shinji Watanabe, · 2017
Closest in time.
“Optimizing expected word error rate via sampling for speech recognition,”
Matt Shannon, · 2017
Closest in time.
“Show, attend and tell: Neural image caption generation with visual attention,”
Kelvin Xu, Jimmy Ba, Ryan Kiros, Kyunghyun Cho, Aaron Courville, Ruslan Salakhudinov, Rich Zemel, and Yoshua Bengio, · 2057
Closest in time.