Fetching the paper…
Reading the bibliography…
Conventional neural autoregressive decoding commonly assumes a fixed left-to-right generation order, which may be sub-optimal.
Wanrong Zhu, Zhiting Hu, and Eric Xing. 2019 · 1901
Earlier work this paper cites.
Insertion transformer: Flexible sequence generation via insertion operations
Mitchell Stern, William Chan, Jamie Kiros, and Jakob Uszkoreit. 2019 · 1902
Earlier work this paper cites.
Non-monotonic sequential text generation
Sean Welleck, Kianté Brantley, Hal Daumé III, and Kyunghyun Cho. 2019 · 1902
Earlier work this paper cites.
A syntax-based statistical translation model
Kenji Yamada and Kevin Knight. 2001 · 2001
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Syntax-based language models for statistical machine translation
Eugene Charniak, Kevin Knight, and Kenji Yamada. 2003 · 2003
Earlier work this paper cites.
Meteor: An automatic metric for mt evaluation with improved correlation with human judgments
Satanjeev Banerjee and Alon Lavie. 2005 · 2005
Earlier work this paper cites.
A hierarchical phrase-based model for statistical machine translation
David Chiang. 2005 · 2005
Earlier work this paper cites.
A neural syntactic language model
Ahmad Emami and Frederick Jelinek. 2005 · 2005
Earlier work this paper cites.
Sparse forward-backward using minimum divergence beams for fast training of conditional random fields
Chris Pal, Charles Sutton, and Andrew McCallum. 2006 · 2006
Earlier work this paper cites.
A study of translation edit rate with targeted human annotation
Matthew Snover, Bonnie Dorr, Richard Schwartz, Linnea Micciulla, and John Makhoul. 2006 · 2006
Earlier work this paper cites.
Automatic evaluation of translation quality for distant language pairs
Hideki Isozaki, Tsutomu Hirao, Kevin Duh, Katsuhito Sudoh, and Hajime Tsukada. 2010 · 2010
Earlier work this paper cites.
The Kyoto free translation task
Graham Neubig. 2011 · 2011
Earlier work this paper cites.
Generating text with recurrent neural networks
Ilya Sutskever, James Martens, and Geoffrey E Hinton. 2011 · 2011
Earlier work this paper cites.
Statistical language models based on neural networks
Tomáš Mikolov. 2012 · 2012
Earlier work this paper cites.
Microsoft coco: Common objects in context
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick. 2014 · 2014
Earlier work this paper cites.
Sequence to sequence learning with neural networks
Ilya Sutskever, Oriol Vinyals, and Quoc V. Le. 2014 · 2014
Earlier work this paper cites.
Neural machine translation by jointly learning to align and translate
Dzmitry Bahdanau, Kyunghyun Cho, and Yoshua Bengio. 2015 · 2015
Cited alongside, same era.
Attention-based models for speech recognition
Jan K Chorowski, Dzmitry Bahdanau, Dmitriy Serdyuk, Kyunghyun Cho, and Yoshua Bengio. 2015 · 2015
Cited alongside, same era.
Deep visual-semantic alignments for generating image descriptions
Andrej Karpathy and Li Fei-Fei. 2015 · 2015
Cited alongside, same era.
Learning to generate pseudo-code from source code using statistical machine translation
Yusuke Oda, Hiroyuki Fudaba, Graham Neubig, Hideaki Hata, Sakriani Sakti, Tomoki Toda, and Satoshi Nakamura. 2015 · 2015
Cited alongside, same era.
A neural attention model for abstractive sentence summarization
Alexander M. Rush, Sumit Chopra, and Jason Weston. 2015 · 2015
Cited alongside, same era.
Towards string-to-tree neural machine translation
Roee Aharoni and Yoav Goldberg. 2017 · 2017
Later among the works it cites.
Learning to parse and translate improves neural machine translation
Akiko Eriguchi, Yoshimasa Tsuruoka, and Kyunghyun Cho. 2017 · 2017
Later among the works it cites.
spaCy 2: Natural language understanding with Bloom embeddings, convolutional neural networks and incremental parsing
Matthew Honnibal and Ines Montani. 2017 · 2017
Later among the works it cites.
Bidirectional beam search: Forward-backward inference in neural sequence models for fill-in-the-blank image captioning
Qing Sun, Stefan Lee, and Dhruv Batra. 2017 · 2017
Later among the works it cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Later among the works it cites.
A syntactic neural model for general-purpose code generation
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cider: Consensus-based image description evaluation
Ramakrishna Vedantam, C Lawrence Zitnick, and Devi Parikh. 2015 · 2015
Cited alongside, same era.
Pointer networks
Oriol Vinyals, Meire Fortunato, and Navdeep Jaitly. 2015 · 2015
Cited alongside, same era.
Oriol Vinyals and Quoc Le. 2015 · 2015
Cited alongside, same era.
Noisy parallel approximate decoding for conditional recurrent language model
Kyunghyun Cho. 2016 · 2016
Cited alongside, same era.
Recurrent neural network grammars
Chris Dyer, Adhiguna Kuncoro, Miguel Ballesteros, and Noah A. Smith. 2016 · 2016
Cited alongside, same era.
Incorporating copying mechanism in sequence-to-sequence learning
Jiatao Gu, Zhengdong Lu, Hang Li, and Victor O. K. Li. 2016 · 2016
Cited alongside, same era.
Deep residual learning for image recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016 · 2016
Cited alongside, same era.
Pengcheng Yin and Graham Neubig. 2017 · 2017
Later among the works it cites.
The importance of generation order in language modeling
Nicolas Ford, Daniel Duckworth, Mohammad Norouzi, and George E. Dahl. 2018 · 2018
Later among the works it cites.
Non-autoregressive neural machine translation
Jiatao Gu, James Bradbury, Caiming Xiong, Victor O.K. Li, and Richard Socher. 2018 · 2018
Later among the works it cites.
Deterministic non-autoregressive neural sequence modeling by iterative refinement
Jason Lee, Elman Mansimov, and Kyunghyun Cho. 2018 · 2018
Later among the works it cites.
Middle-out decoding
Shikib Mehri and Leonid Sigal. 2018 · 2018
Later among the works it cites.
Learning latent permutations with gumbel-sinkhorn networks
Gonzalo Mena, David Belanger, Scott Linderman, and Jasper Snoek. 2018 · 2018
Later among the works it cites.
Parallel WaveNet: Fast high-fidelity speech synthesis
Aaron van den Oord, Yazhe Li, Igor Babuschkin, Karen Simonyan, Oriol Vinyals, Koray Kavukcuoglu, George van den Driessche, Edward Lockhart, Luis Cobo, Florian Stimberg, Norman Casagrande, Dominik Grewe, Seb Noury, Sander Dieleman, Erich Elsen, Nal Kalchbrenner, Heiga Zen, Alex Graves, Helen King, Tom Walters, Dan Belov, and Demis Hassabis. 2018 · 2018
Later among the works it cites.
Self-attention with relative position representations
Peter Shaw, Jakob Uszkoreit, and Ashish Vaswani. 2018 · 2018
Later among the works it cites.
Blockwise parallel decoding for deep autoregressive models
Mitchell Stern, Noam Shazeer, and Jakob Uszkoreit. 2018 · 2018
Later among the works it cites.
Semi-autoregressive neural machine translation
Chunqi Wang, Ji Zhang, and Haiqing Chen. 2018a · 2018
Later among the works it cites.
A tree-based decoder for neural machine translation
Xinyi Wang, Hieu Pham, Pengcheng Yin, and Graham Neubig. 2018b · 2018
Later among the works it cites.
Beyond error propagation in neural machine translation: Characteristics of language also matter
Lijun Wu, Xu Tan, Di He, Fei Tian, Tao Qin, Jianhuang Lai, and Tie-Yan Liu. 2018 · 2018
Later among the works it cites.