Fetching the paper…
Reading the bibliography…
Quite surprisingly, exact maximum a posteriori (MAP) decoding of neural language generators frequently leads to low-quality results.
Calibration of encoder decoder models for neural machine translation
Aviral Kumar and Sunita Sarawagi. 2019 · 1903
Earlier work this paper cites.
A note on two problems in connexion with graphs
Edsger W. Dijkstra. 1959 · 1959
Earlier work this paper cites.
Speech understanding systems: A summary of results of the five-year research effort at carnegie mellon university
Raj Reddy. 1977 · 1977
Earlier work this paper cites.
Improved alignment models for statistical machine translation
Franz Josef Och, Christoph Tillmann, and Hermann Ney. 1999 · 1999
Earlier work this paper cites.
A probabilistic Earley parser as a psycholinguistic model
John Hale. 2001 · 2001
Earlier work this paper cites.
BLEU: A method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Statistical phrase-based translation
Philipp Koehn, Franz J. Och, and Daniel Marcu. 2003 · 2003
Earlier work this paper cites.
Pharaoh: A beam search decoder for phrase-based statistical machine translation models
Philipp Koehn. 2004 · 2004
Earlier work this paper cites.
Algorithm Design
Jon Kleinberg and Éva Tardos. 2005 · 2005
Earlier work this paper cites.
Probabilistic Models of Word Order and Syntactic Discontinuity
Roger Levy. 2005 · 2005
Earlier work this paper cites.
Speakers optimize information density through syntactic reduction
Roger P. Levy and T. F. Jaeger. 2007 · 2007
Earlier work this paper cites.
Redundancy and reduction: Speakers manage syntactic information density
T. Florian Jaeger. 2010 · 2010
Earlier work this paper cites.
On language ’utility’: Processing complexity and communicative efficiency
T. Florian Jaeger and Harry Tily. 2011 · 2011
Earlier work this paper cites.
Wit 3 : Web inventory of transcribed and translated talks
Mauro Cettolo, Christian Girardi, and Marcello Federico. 2012 · 2012
Earlier work this paper cites.
Recurrent convolutional neural networks for discourse compositionality
Nal Kalchbrenner and Phil Blunsom. 2013 · 2013
Earlier work this paper cites.
Findings of the 2014 workshop on statistical machine translation
Ondřej Bojar, Christian Buck, Christian Federmann, Barry Haddow, Philipp Koehn, Johannes Leveling, Christof Monz, Pavel Pecina, Matt Post, Herve Saint-Amand, Radu Soricut, Lucia Specia, and Aleš Tamchyna. 2014 · 2014
Earlier work this paper cites.
On the properties of neural machine translation: Encoder–decoder approaches
Kyunghyun Cho, Bart van Merriënboer, Dzmitry Bahdanau, and Yoshua Bengio. 2014 · 2014
Earlier work this paper cites.
Sequence to sequence learning with neural networks
Ilya Sutskever, Oriol Vinyals, and Quoc V. Le. 2014 · 2014
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio. 2015 · 2015
Cited alongside, same era.
Scheduled sampling for sequence prediction with recurrent neural networks
Samy Bengio, Oriol Vinyals, Navdeep Jaitly, and Noam Shazeer. 2015 · 2015
Cited alongside, same era.
Montreal neural machine translation systems for WMT’15
Sébastien Jean, Orhan Firat, Kyunghyun Cho, Roland Memisevic, and Yoshua Bengio. 2015 · 2015
Cited alongside, same era.
Oriol Vinyals and Quoc V. Le. 2015 · 2015
Cited alongside, same era.
Show and tell: Lessons learned from the 2015 mscoco image captioning challenge
Oriol Vinyals, Alexander Toshev, Samy Bengio, and Dumitru Erhan. 2017 · 2015
Cited alongside, same era.
SGNMT – a flexible NMT decoding platform for quick prototyping of new models and search strategies
Felix Stahlberg, Eva Hasler, Danielle Saunders, and Bill Byrne. 2017 · 2017
Later among the works it cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Later among the works it cites.
Recurrent neural networks as weighted language recognizers
Yining Chen, Sorcha Gilroy, Andreas Maletti, Jonathan May, and Kevin Knight. 2018 · 2018
Later among the works it cites.
Understanding back-translation at scale
Sergey Edunov, Myle Ott, Michael Auli, and David Grangier. 2018 · 2018
Later among the works it cites.
Hierarchical neural story generation
Angela Fan, Mike Lewis, and Yann Dauphin. 2018 · 2018
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Improved neural machine translation with smt features
Wei He, Zhongjun He, Hua Wu, and Haifeng Wang. 2016 · 2016
Cited alongside, same era.
A diversity-promoting objective function for neural conversation models
Jiwei Li, Michel Galley, Chris Brockett, Jianfeng Gao, and Bill Dolan. 2016 · 2016
Cited alongside, same era.
Minimum risk training for neural machine translation
Shiqi Shen, Yong Cheng, Zhongjun He, Wei He, Hua Wu, Maosong Sun, and Yang Liu. 2016 · 2016
Cited alongside, same era.
Length bias in encoder decoder models and a case for global conditioning
Pavel Sountsov and Sunita Sarawagi. 2016 · 2016
Cited alongside, same era.
Modeling coverage for neural machine translation
Zhaopeng Tu, Zhengdong Lu, Yang Liu, Xiaohua Liu, and Hang Li. 2016 · 2016
Cited alongside, same era.
Google’s neural machine translation system: Bridging the gap between human and machine translation
Yonghui Wu, Mike Schuster, Zhifeng Chen, Quoc V. Le, Mohammad Norouzi, Wolfgang Macherey, Maxim Krikun, Yuan Cao, Qin Gao, Klaus Macherey, Jeff Klingner, Apurva Shah, Melvin Johnson, Xiaobing Liu, Lukasz Kaiser, Stephan Gouws, Yoshikiyo Kato, Taku Kudo, Hideto Kazawa, Keith Stevens, George Kurian, Nishant Patil, Wei Wang, Cliff Young, Jason Smith, Jason Riesa, Alex Rudnick, Oriol Vinyals, Gregory S. Corrado, Macduff Hughes, and Jeffrey Dean. 2016 · 2016
Cited alongside, same era.
Neural generative question answering
Jun Yin, Xin Jiang, Zhengdong Lu, Lifeng Shang, Hang Li, and Xiaoming Li. 2016 · 2016
Cited alongside, same era.
Ayush Jain, Vishal Singh, Sidharth Ranjan, Rajakrishnan Rajkumar, and Sumeet Agarwal. 2018 · 2018
Later among the works it cites.
Correcting length bias in neural machine translation
Kenton Murray and David Chiang. 2018 · 2018
Later among the works it cites.
A call for clarity in reporting BLEU scores
Matt Post. 2018 · 2018
Later among the works it cites.
Diverse beam search for improved description of complex scenes
Ashwin K. Vijayakumar, Michael Cogswell, Ramprasaath R. Selvaraju, Qing Sun, Stefan Lee, David J. Crandall, and Dhruv Batra. 2018 · 2018
Later among the works it cites.
Breaking the beam search curse: A study of (re-)scoring methods and stopping criteria for neural machine translation
Yilin Yang, Liang Huang, and Mingbo Ma. 2018 · 2018
Later among the works it cites.
Empirical analysis of beam search performance degradation in neural sequence models
Eldan Cohen and Christopher Beck. 2019 · 2019
Later among the works it cites.
Facebook FAIR’s WMT19 news translation task submission
Nathan Ng, Kyra Yee, Alexei Baevski, Myle Ott, Michael Auli, and Sergey Edunov. 2019 · 2019
Later among the works it cites.
fairseq: A fast, extensible toolkit for sequence modeling
Myle Ott, Sergey Edunov, Alexei Baevski, Angela Fan, Sam Gross, Nathan Ng, David Grangier, and Michael Auli. 2019 · 2019
Later among the works it cites.
On NMT Search Errors and Model Errors: Cat Got Your Tongue?
Felix Stahlberg and Bill Byrne. 2019 · 2019
Later among the works it cites.
Xlnet: Generalized autoregressive pretraining for language understanding
Zhilin Yang, Zihang Dai, Yiming Yang, Jaime Carbonell, Russ R Salakhutdinov, and Quoc V Le. 2019 · 2019
Later among the works it cites.
The curious case of neural text degeneration
Ari Holtzman, Jan Buys, Maxwell Forbes, and Yejin Choi. 2020 · 2020
Closest in time.
Clara Meister, Tim Vieira, and Ryan Cotterell. 2020 · 2020
Closest in time.