Fetching the paper…
Reading the bibliography…
Non-autoregressive approaches aim to improve the inference speed of translation models by only requiring a single forward pass to generate the output sequence instead of iteratively producing each predicted token.
Error detecting and error correcting codes
Richard Wesley Hamming. 1950 · 1950
Earlier work this paper cites.
Binary codes capable of correcting deletions, insertions and reversals
Vladimir Iosifovich Levenshtein. 1966 · 1965
Earlier work this paper cites.
Bleu: a Method for Automatic Evaluation of Machine Translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Glancing Transformer for Non-Autoregressive Neural Machine Translation
Lihua Qian, Hao Zhou, Yu Bao, Mingxuan Wang, Lin Qiu, Weinan Zhang, Yong Yu, and Lei Li. 2021 · 2003
Earlier work this paper cites.
Statistical Significance Tests for Machine Translation Evaluation
Philipp Koehn. 2004 · 2004
Earlier work this paper cites.
Clustering with Bregman Divergences
Arindam Banerjee, Srujana Merugu, Inderjit S. Dhillon, and Joydeep Ghosh. 2005 · 2005
Earlier work this paper cites.
Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks
Alex Graves, Santiago Fernández, Faustino J. Gomez, and Jürgen Schmidhuber. 2006 · 2006
Earlier work this paper cites.
Continuous Space Language Models for Statistical Machine Translation
Holger Schwenk, Daniel Dechelotte, and Jean-Luc Gauvain. 2006 · 2006
Earlier work this paper cites.
A Study of Translation Edit Rate with Targeted Human Annotation
Matthew Snover, Bonnie Dorr, Rich Schwartz, Linnea Micciulla, and John Makhoul. 2006 · 2006
Earlier work this paper cites.
Moses: Open Source Toolkit for Statistical Machine Translation
Philipp Koehn, Hieu Hoang, Alexandra Birch, Chris Callison-Burch, Marcello Federico, Nicola Bertoldi, Brooke Cowan, Wade Shen, Christine Moran, Richard Zens, Chris Dyer, Ondřej Bojar, Alexandra Constantin, and Evan Herbst. 2007 · 2007
Earlier work this paper cites.
Adam: A Method for Stochastic Optimization
Diederik P. Kingma and Jimmy Ba. 2015 · 2015
Earlier work this paper cites.
Gaussian Error Linear Units (GELUs)
Dan Hendrycks and Kevin Gimpel. 2016 · 2016
Earlier work this paper cites.
Sequence-Level Knowledge Distillation
Yoon Kim and Alexander M. Rush. 2016 · 2016
Earlier work this paper cites.
Neural Machine Translation of Rare Words with Subword Units
Rico Sennrich, Barry Haddow, and Alexandra Birch. 2016 · 2016
Earlier work this paper cites.
chrF++: words helping character n-grams
Maja Popović. 2017 · 2017
Earlier work this paper cites.
Using the Output Embedding to Improve Language Models
Ofir Press and Lior Wolf. 2017 · 2017
Earlier work this paper cites.
Attention is All you Need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Lukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
Non-Autoregressive Neural Machine Translation
Jiatao Gu, James Bradbury, Caiming Xiong, Victor O. K. Li, and Richard Socher. 2018 · 2018
Earlier work this paper cites.
Marian: Fast Neural Machine Translation in C++
Marcin Junczys-Dowmunt, Roman Grundkiewicz, Tomasz Dwojak, Hieu Hoang, Kenneth Heafield, Tom Neckermann, Frank Seide, Ulrich Germann, Alham Fikri Aji, Nikolay Bogoychev, André F. T. Martins, and Alexandra Birch. 2018a · 2018
Earlier work this paper cites.
Deterministic Non-Autoregressive Neural Sequence Modeling by Iterative Refinement
Jason Lee, Elman Mansimov, and Kyunghyun Cho. 2018 · 2018
Earlier work this paper cites.
End-to-End Non-Autoregressive Neural Machine Translation with Connectionist Temporal Classification
Jindřich Libovický and Jindřich Helcl. 2018 · 2018
Cited alongside, same era.
A Call for Clarity in Reporting BLEU Scores
Matt Post. 2018 · 2018
Cited alongside, same era.
Blockwise Parallel Decoding for Deep Autoregressive Models
Mitchell Stern, Noam Shazeer, and Jakob Uszkoreit. 2018 · 2018
Cited alongside, same era.
Accelerating Neural Transformer via an Average Attention Network
Biao Zhang, Deyi Xiong, and Jinsong Su. 2018 · 2018
Cited alongside, same era.
Understanding Knowledge Distillation in Non-autoregressive Machine Translation
Chunting Zhou, Jiatao Gu, and Graham Neubig. 2020 · 2018
Cited alongside, same era.
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Non-Autoregressive Machine Translation with Latent Alignments
Chitwan Saharia, William Chan, Saurabh Saxena, and Mohammad Norouzi. 2020 · 2020
Later among the works it cites.
An EM Approach to Non-autoregressive Conditional Sequence Generation
Zhiqing Sun and Yiming Yang. 2020 · 2020
Later among the works it cites.
Can Multilinguality benefit Non-autoregressive Machine Translation?
Sweta Agrawal, Julia Kreutzer, and Colin Cherry. 2021 · 2021
Later among the works it cites.
Non-Autoregressive Translation by Learning Target Categorical Codes
Yu Bao, Shujian Huang, Tong Xiao, Dongqi Wang, Xinyu Dai, and Jiajun Chen. 2021 · 2021
Later among the works it cites.
Order-Agnostic Cross Entropy for Non-Autoregressive Machine Translation
Cunxiao Du, Zhaopeng Tu, and Jing Jiang. 2021 · 2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Jacob Devlin, Ming-Wei Chang, Kenton Lee, and Kristina Toutanova. 2019 · 2019
Cited alongside, same era.
Mask-Predict: Parallel Decoding of Conditional Masked Language Models
Marjan Ghazvininejad, Omer Levy, Yinhan Liu, and Luke Zettlemoyer. 2019 · 2019
Cited alongside, same era.
Levenshtein Transformer
Jiatao Gu, Changhan Wang, and Junbo Zhao. 2019 · 2019
Cited alongside, same era.
Hint-Based Training for Non-Autoregressive Machine Translation
Zhuohan Li, Zi Lin, Di He, Fei Tian, Tao Qin, Liwei Wang, and Tie-Yan Liu. 2019 · 2019
Cited alongside, same era.
FlowSeq: Non-Autoregressive Conditional Sequence Generation with Generative Flow
Xuezhe Ma, Chunting Zhou, Xian Li, Graham Neubig, and Eduard Hovy. 2019 · 2019
Cited alongside, same era.
compare-mt: A Tool for Holistic Comparison of Language Generation Systems
Graham Neubig, Zi-Yi Dou, Junjie Hu, Paul Michel, Danish Pruthi, and Xinyi Wang. 2019 · 2019
Cited alongside, same era.
Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks
Nils Reimers and Iryna Gurevych. 2019 · 2019
Cited alongside, same era.
Jiatao Gu and Xiang Kong. 2021 · 2021
Later among the works it cites.
Multi-Task Learning with Shared Encoder for Non-Autoregressive Machine Translation
Yongchang Hao, Shilin He, Wenxiang Jiao, Zhaopeng Tu, Michael Lyu, and Xing Wang. 2021 · 2021
Later among the works it cites.
Findings of the WMT 2021 Shared Task on Efficient Translation
Kenneth Heafield, Qianqian Zhu, and Roman Grundkiewicz. 2021 · 2021
Later among the works it cites.
Deep Encoder, Shallow Decoder: Reevaluating Non-autoregressive Machine Translation
Jungo Kasai, Nikolaos Pappas, Hao Peng, James Cross, and Noah A. Smith. 2021 · 2021
Later among the works it cites.
Scientific Credibility of Machine Translation Research: A Meta-Evaluation of 769 Papers
Benjamin Marie, Atsushi Fujita, and Raphael Rubino. 2021 · 2021
Later among the works it cites.
Guiding Non-Autoregressive Neural Machine Translation Decoding with Reordering Information
Qiu Ran, Yankai Lin, Peng Li, and Jie Zhou. 2021 · 2021
Later among the works it cites.
Descending through a Crowded Valley - Benchmarking Deep Learning Optimizers
Robin M. Schmidt, Frank Schneider, and Philipp Hennig. 2021 · 2021
Later among the works it cites.
Sequence-Level Training for Non-Autoregressive Neural Machine Translation
Chenze Shao, Yang Feng, Jinchao Zhang, Fandong Meng, and Jie Zhou. 2021 · 2021
Later among the works it cites.
AligNART: Non-autoregressive Neural Machine Translation by Jointly Learning to Estimate Alignment and Translate
Jongyoon Song, Sungwon Kim, and Sungroh Yoon. 2021 · 2021
Later among the works it cites.
Non-Autoregressive Text Generation with Pre-trained Language Models
Yixuan Su, Deng Cai, Yan Wang, David Vandyke, Simon Baker, Piji Li, and Nigel Collier. 2021 · 2021
Later among the works it cites.
How Length Prediction Influence the Performance of Non-Autoregressive Translation?
Minghan Wang, Guo Jiaxin, Yuxia Wang, Yimeng Chen, Su Chang, Hengchao Shang, Min Zhang, Shimin Tao, and Hao Yang. 2021 · 2021
Later among the works it cites.
How Does Distilled Data Complexity Impact the Quality and Confidence of Non-Autoregressive Machine Translation?
Weijia Xu, Shuming Ma, Dongdong Zhang, and Marine Carpuat. 2021 · 2021
Later among the works it cites.
Non-Autoregressive Machine Translation: It’s Not as Fast as it Seems
Jindřich Helcl, Barry Haddow, and Alexandra Birch. 2022 · 2022
Closest in time.
Non-Autoregressive Translation with Layer-Wise Prediction and Deep Supervision
Chenyang Huang, Hao Zhou, Osmar R. Zaïane, Lili Mou, and Lei Li. 2022 · 2022
Closest in time.