Fetching the paper…
Reading the bibliography…
There has recently been significant interest in hard attention models for tasks such as object recognition, visual captioning and speech recognition.
Learning from delayed rewards
Christopher John Cornish Hellaby Watkins, · 1989
Earlier work this paper cites.
“Speaker-independent phone recognition using hidden markov models,”
K-F Lee and H-W Hon, · 1989
Earlier work this paper cites.
“Simple statistical gradient-following algorithms for connectionist reinforcement learning,”
Ronald J Williams, · 1992
Earlier work this paper cites.
“Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,”
Alex Graves, Santiago Fernández, Faustino Gomez, and Jürgen Schmidhuber, · 2006
Earlier work this paper cites.
“Practical variational inference for neural networks,”
Alex Graves, · 2011
Earlier work this paper cites.
“Neural variational inference and learning in belief networks,”
Andriy Mnih and Karol Gregor, · 2014
Earlier work this paper cites.
“Learning generative models with visual attention,”
Yichuan Tang, Nitish Srivastava, and Ruslan R Salakhutdinov, · 2014
Earlier work this paper cites.
“Recurrent models of visual attention,”
Volodymyr Mnih, Nicolas Heess, Alex Graves, et al., · 2014
Cited alongside, same era.
“Neural machine translation by jointly learning to align and translate,”
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio, · 2014
Cited alongside, same era.
“Multiple object recognition with visual attention,”
Jimmy Ba, Volodymyr Mnih, and Koray Kavukcuoglu, · 2014
Cited alongside, same era.
“Importance weighted autoencoders,”
Yuri Burda, Roger B. Grosse, and Ruslan Salakhutdinov, · 2015
Cited alongside, same era.
“Show, attend and tell: Neural image caption generation with visual attention.,”
Kelvin Xu, Jimmy Ba, Ryan Kiros, Kyunghyun Cho, Aaron C Courville, Ruslan Salakhutdinov, Richard S Zemel, and Yoshua Bengio, · 2015
Cited alongside, same era.
“Reinforcement learning neural turing machines-revised,”
Wojciech Zaremba and Ilya Sutskever, · 2015
Later among the works it cites.
“Learning wake-sleep recurrent attention models,”
Jimmy Ba, Ruslan R Salakhutdinov, Roger B Grosse, and Brendan J Frey, · 2015
Later among the works it cites.
“Learning online alignments with continuous rewards policy gradient,”
Yuping Luo, Chung-Cheng Chiu, Navdeep Jaitly, and Ilya Sutskever, · 2016
Later among the works it cites.
“Variational inference for monte carlo objectives,”
Andriy Mnih and Danilo Jimenez Rezende, · 2016
Later among the works it cites.
“Learning to translate in real-time with neural machine translation,”
Jiatao Gu, Graham Neubig, Kyunghyun Cho, and Victor O. K. Li, · 2016
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
“Attention-based models for speech recognition,”
Jan K Chorowski, Dzmitry Bahdanau, Dmitriy Serdyuk, Kyunghyun Cho, and Yoshua Bengio, · 2015
Cited alongside, same era.
“An online sequence-to-sequence model using partial conditioning,”
Navdeep Jaitly, David Sussillo, Quoc V. Le, Oriol Vinyals, Ilya Sutskever, and Samy Bengio, · 2015
Cited alongside, same era.
Later among the works it cites.
“Online and linear-time attention by enforcing monotonic alignments,”
Colin Raffel, Thang Luong, Peter J Liu, Ron J Weiss, and Douglas Eck, · 2017
Closest in time.