Fetching the paper…
Reading the bibliography…
The Teacher Forcing algorithm trains recurrent networks by supplying observed sequence values as inputs during training and using the network's own one-step-ahead predictions to do multi-step sampling.
A learning algorithm for continually running fully recurrent neural networks
Williams, R. J. and Zipser, D. (1989) · 1989
Earlier work this paper cites.
Long short-term memory
Hochreiter, S. and Schmidhuber, J. (1997) · 1997
Earlier work this paper cites.
Iam-ondb - an on-line english sentence database acquired from handwritten text on a whiteboard
Liwicki, M. and Bunke, H. (2005) · 2005
Earlier work this paper cites.
Search-based Structured Prediction
Daumé, III, H., Langford, J., and Marcu, D. (2009) · 2009
Earlier work this paper cites.
Evaluating probabilities under high-dimensional latent variable models
Murray, I. and Salakhutdinov, R. R. (2009) · 2009
Earlier work this paper cites.
Recurrent neural network based language model
Mikolov, T. (2010) · 2010
Earlier work this paper cites.
A Reduction of Imitation Learning and Structured Prediction to No-Regret Online Learning
Ross, S., Gordon, G. J., and Bagnell, J. A. (2010) · 2010
Earlier work this paper cites.
The neural autoregressive distribution estimator
Larochelle, H. and Murray, I. (2011) · 2011
Earlier work this paper cites.
Supervised Sequence Labelling with Recurrent Neural Networks
Graves, A. (2012) · 2012
Earlier work this paper cites.
Context dependent recurrent neural network language model
Mikolov, T. and Zweig, G. (2012) · 2012
Earlier work this paper cites.
Estimating or Propagating Gradients Through Stochastic Neurons for Conditional Computation
Bengio, Y., Léonard, N., and Courville, A. (2013) · 2013
Earlier work this paper cites.
Generating sequences with recurrent neural networks
Graves, A. (2013) · 2013
Earlier work this paper cites.
Generating Sequences With Recurrent Neural Networks
Graves, A. (2013) · 2013
Cited alongside, same era.
Domain-Adversarial Neural Networks
Ajakan, H., Germain, P., Larochelle, H., Laviolette, F., and Marchand, M. (2014) · 2014
Cited alongside, same era.
Neural machine translation by jointly learning to align and translate
Bahdanau, D., Cho, K., and Bengio, Y. (2014) · 2014
Cited alongside, same era.
Learning phrase representations using RNN encoder–decoder for statistical machine translation
Cho, K., Van Merriënboer, B., Gülçehre, Ç., Bahdanau, D., Bougares, F., Schwenk, H., and Bengio, Y. (2014a) · 2014
Cited alongside, same era.
Generative adversarial networks
Goodfellow, I. J., Pouget-Abadie, J., Mirza, M., Xu, B., Warde-Farley, D., Ozair, S., Courville, A., and Bengio, Y. (2014) · 2014
Cited alongside, same era.
Attention-based models for speech recognition
Chorowski, J. K., Bahdanau, D., Serdyuk, D., Cho, K., and Bengio, Y. (2015) · 2015
Later among the works it cites.
Domain-Adversarial Training of Neural Networks
Ganin, Y., Ustinova, E., Ajakan, H., Germain, P., Larochelle, H., Laviolette, F., Marchand, M., and Lempitsky, V. (2015) · 2015
Later among the works it cites.
Made: Masked autoencoder for distribution estimation
Germain, M., Gregor, K., Murray, I., and Larochelle, H. (2015) · 2015
Later among the works it cites.
Draw: A recurrent neural network for image generation
Gregor, K., Danihelka, I., Graves, A., and Wierstra, D. (2015) · 2015
Later among the works it cites.
How (not) to Train your Generative Model: Scheduled Sampling, Likelihood, Adversary?
Huszár, F. (2015) · 2015
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Kingma, D. and Ba, J. (2014) · 2014
Cited alongside, same era.
Iterative neural autoregressive distribution estimator NADE-k
Raiko, T., Yao, L., Cho, K., and Bengio, Y. (2014) · 2014
Cited alongside, same era.
Markov chain monte carlo and variational inference: Bridging the gap
Salimans, T., Kingma, D. P., and Welling, M. (2014) · 2014
Cited alongside, same era.
Sequence to sequence learning with neural networks
Sutskever, I., Vinyals, O., and Le, Q. V. (2014) · 2014
Cited alongside, same era.
End-to-end attention-based large vocabulary speech recognition
Bahdanau, D., Chorowski, J., Serdyuk, D., Brakel, P., and Bengio, Y. (2015) · 2015
Cited alongside, same era.
Scheduled sampling for sequence prediction with recurrent neural networks
Bengio, S., Vinyals, O., Jaitly, N., and Shazeer, N. (2015) · 2015
Cited alongside, same era.
Mind’s eye: A recurrent visual representation for image caption generation
Chen, X. and Lawrence Zitnick, C. (2015) · 2015
Cited alongside, same era.
A note on the evaluation of generative models
Theis, L., van den Oord, A., and Bethge, M. (2015) · 2015
Later among the works it cites.
Show, attend and tell: Neural image caption generation with visual attention
Xu, K., Ba, J., Kiros, R., Courville, A., Salakhutdinov, R., Zemel, R., and Bengio, Y. (2015) · 2015
Later among the works it cites.
Theano: A python framework for fast computation of mathematical expressions
Al-Rfou, R., Alain, G., Almahairi, A., and et al. (2016) · 2016
Closest in time.
An Actor-Critic Algorithm for Sequence Prediction
Bahdanau, D., Brakel, P., Xu, K., Goyal, A., Lowe, R., Pineau, J., Courville, A., and Bengio, Y. (2016) · 2016
Closest in time.
Conditional handwriting generation in theano
Brebisson, A. (2016) · 2016
Closest in time.
Pixel Recurrent Neural Networks
van den Oord, A., Kalchbrenner, N., and Kavukcuoglu, K. (2016) · 2016
Closest in time.