Fetching the paper…
Reading the bibliography…
In natural language processing tasks performance of the models is often measured with some non-differentiable metric, such as BLEU score.
A learning algorithm for continually running fully recurrent neural networks
Ronald J. Williams and David Zipser · 1989
Earlier work this paper cites.
Simple statistical gradient-following algorithms for connectionist reinforcement learning
Ronald J. Williams · 1992
Earlier work this paper cites.
Long short-term memory
Sepp Hochreiter and Jurgen Schmidhuber · 1997
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu · 2002
Earlier work this paper cites.
Rouge: A package for automatic evaluation of summaries, 2004
Chin-Yew Lin · 2004
Earlier work this paper cites.
Variance reduction techniques for gradient estimates in reinforcement learning
Evan Greensmith, Peter L. Bartlett, and Jonathan Baxter · 2004
Earlier work this paper cites.
Sequence to sequence learning with neural networks, 2014
Ilya Sutskever, Oriol Vinyals, and Quoc V. Le · 2014
Cited alongside, same era.
Report on the 11th iwslt evaluation campaign, iwslt 201
Mauro Cettolo, Jan Niehues, Sebastian Stuker, Luisa Bentivogli, and Marcello Federico · 2014
Cited alongside, same era.
Adam: A method for stochastic optimization, 2014
Diederik P. Kingma and Jimmy Ba · 2014
Cited alongside, same era.
Gradient estimation using stochastic computation graphs, 2015
John Schulman, Nicolas Heess, Theophane Weber, and Pieter Abbeel · 2015
Cited alongside, same era.
Effective approaches to attention-based neural machine translation, 2015
Minh-Thang Luong, Hieu Pham, and Christopher D. Manning · 2015
Cited alongside, same era.
How not to evaluate your dialogue system: An empirical study of unsupervised evaluation metrics for dialogue response generation, 2016
Minimum risk training for neural machine translation, 2016
Shiqi Shen, Yong Cheng, Zhongjun He, Wei He, Hua Wu, Maosong Sun, and Yang Liu · 2016
Later among the works it cites.
Sequence-to-sequence learning as beam-search optimization, 2016
Sam Wiseman and Alexander M. Rush · 2016
Later among the works it cites.
A deep reinforced model for abstractive summarization, 2017
Romain Paulus, Caiming Xiong, and Richard Socher · 2017
Closest in time.
Rebar: Low-variance, unbiased gradient estimates for discrete latent variable models, 2017
George Tucker, Andriy Mnih, Chris J. Maddison, Dieterich Lawson, and Jascha Sohl-Dickstein · 2017
Closest in time.
Automatic differentiation in pytorch
Adam Paszke, Sam Gross, Soumith Chintala, Gregory Chanan, Edward Yang, Zachary DeVito, Zeming Lin, Alban Desmaison, Luca Antiga, and Adam Lerer · 2017
Closest in time.
A continuous relaxation of beam search for end-to-end training of neural sequence models, 2017
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Chia-Wei Liu, Ryan Lowe, Iulian V. Serban, Michael Noseworthy, Laurent Charlin, and Joelle Pineau · 2016
Cited alongside, same era.
Ninth workshop on statistical machine translation
ACL2014
Cited in the paper.
Kartik Goyal, Graham Neubig, Chris Dyer, and Taylor Berg-Kirkpatrick · 2017
Closest in time.