Fetching the paper…
Reading the bibliography…
Non-autoregressive neural machine translation (NAT) models suffer from the multi-modality problem that there may exist multiple possible translations of a source sentence, so the reference sentence may be inappropriate for the training when the NAT output is closer to other translations.
Non-autoregressive Transformer by Position Learning
Bao, Y.; Zhou, H.; Feng, J.; Wang, M.; Huang, S.; Chen, J.; and Lei, L. 2019 · 1911
Earlier work this paper cites.
Simple Statistical Gradient-Following Algorithms for Connectionist Reinforcement Learning
Williams, R. J. 1992 · 1992
Earlier work this paper cites.
The Optimal Reward Baseline for Gradient-Based Reinforcement Learning
Weaver, L.; and Tao, N. 2001 · 2001
Earlier work this paper cites.
Bleu: a Method for Automatic Evaluation of Machine Translation
Papineni, K.; Roukos, S.; Ward, T.; and Zhu, W.-J. 2002 · 2002
Earlier work this paper cites.
METEOR: An Automatic Metric for MT Evaluation with Improved Correlation with Human Judgments
Banerjee, S.; and Lavie, A. 2005 · 2005
Earlier work this paper cites.
Connectionist Temporal Classification: Labelling Unsegmented Sequence Data with Recurrent Neural Networks
Graves, A.; Fernández, S.; Gomez, F.; and Schmidhuber, J. 2006 · 2006
Earlier work this paper cites.
Adam: A Method for Stochastic Optimization
Kingma, D. P.; and Ba, J. 2014 · 2014
Earlier work this paper cites.
Neural Machine Translation by Jointly Learning to Align and Translate
Bahdanau, D.; Cho, K.; and Bengio, Y. 2015 · 2015
Earlier work this paper cites.
Sequence-Level Knowledge Distillation
Kim, Y.; and Rush, A. M. 2016 · 2016
Earlier work this paper cites.
Sequence Level Training with Recurrent Neural Networks
Ranzato, M.; Chopra, S.; Auli, M.; and Zaremba, W. 2016 · 2016
Earlier work this paper cites.
Neural Machine Translation of Rare Words with Subword Units
Sennrich, R.; Haddow, B.; and Birch, A. 2016 · 2016
Earlier work this paper cites.
An Actor-Critic Algorithm for Sequence Prediction
Bahdanau, D.; Brakel, P.; Xu, K.; Goyal, A.; Lowe, R.; Pineau, J.; Courville, A. C.; and Bengio, Y. 2017 · 2017
Earlier work this paper cites.
Convolutional Sequence to Sequence Learning
Gehring, J.; Auli, M.; Grangier, D.; Yarats, D.; and Dauphin, Y. N. 2017 · 2017
Earlier work this paper cites.
Trainable Greedy Decoding for Neural Machine Translation
Gu, J.; Cho, K.; and Li, V. O. 2017 · 2017
Earlier work this paper cites.
Improving Japanese-to-English Neural Machine Translation by Paraphrasing the Target Language
Sekizawa, Y.; Kajiwara, T.; and Komachi, M. 2017 · 2017
Earlier work this paper cites.
Attention is All You Need
Vaswani, A.; Shazeer, N.; Parmar, N.; Uszkoreit, J.; Jones, L.; Gomez, A. N.; Kaiser, u.; and Polosukhin, I. 2017 · 2017
Earlier work this paper cites.
Non-Autoregressive Neural Machine Translation
Gu, J.; Bradbury, J.; Xiong, C.; Li, V. O. K.; and Socher, R. 2018 · 2018
Earlier work this paper cites.
Fast Decoding in Sequence Models Using Discrete Latent Variables
Kaiser, L.; Bengio, S.; Roy, A.; Vaswani, A.; Parmar, N.; Uszkoreit, J.; and Shazeer, N. 2018 · 2018
Cited alongside, same era.
End-to-End Non-Autoregressive Neural Machine Translation with Connectionist Temporal Classification
Libovický, J.; and Helcl, J. 2018 · 2018
Cited alongside, same era.
A Study of Reinforcement Learning for Neural Machine Translation
Wu, L.; Tian, F.; Qin, T.; Lai, J.; and Liu, T.-Y. 2018 · 2018
Cited alongside, same era.
Syntactically Supervised Transformers for Faster Neural Machine Translation
Akoury, N.; Krishna, K.; and Iyyer, M. 2019 · 2019
Cited alongside, same era.
Mask-Predict: Parallel Decoding of Conditional Masked Language Models
Ghazvininejad, M.; Levy, O.; Liu, Y.; and Zettlemoyer, L. 2019 · 2019
Cited alongside, same era.
FlowSeq: Non-Autoregressive Conditional Sequence Generation with Generative Flow
ENGINE: Energy-Based Inference Networks for Non-Autoregressive Machine Translation
Tu, L.; Pang, R. Y.; Wiseman, S.; and Gimpel, K. 2020 · 2020
Later among the works it cites.
BERTScore: Evaluating Text Generation with BERT
Zhang*, T.; Kishore*, V.; Wu*, F.; Weinberger, K. Q.; and Artzi, Y. 2020 · 2020
Later among the works it cites.
Understanding Knowledge Distillation in Non-autoregressive Machine Translation
Zhou, C.; Gu, J.; and Neubig, G. 2020 · 2020
Later among the works it cites.
Non-Autoregressive Translation by Learning Target Categorical Codes
Bao, Y.; Huang, S.; Xiao, T.; Wang, D.; Dai, X.; and Chen, J. 2021 · 2021
Later among the works it cites.
Order-Agnostic Cross Entropy for Non-Autoregressive Machine Translation
Du, C.; Tu, Z.; and Jiang, J. 2021 · 2021
Later among the works it cites.
Fully Non-autoregressive Neural Machine Translation: Tricks of the Trade
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ma, X.; Zhou, C.; Li, X.; Neubig, G.; and Hovy, E. 2019 · 2019
Cited alongside, same era.
fairseq: A Fast, Extensible Toolkit for Sequence Modeling
Ott, M.; Edunov, S.; Baevski, A.; Fan, A.; Gross, S.; Ng, N.; Grangier, D.; and Auli, M. 2019 · 2019
Cited alongside, same era.
Retrieving Sequential Information for Non-Autoregressive Neural Machine Translation
Shao, C.; Feng, Y.; Zhang, J.; Meng, F.; Chen, X.; and Zhou, J. 2019 · 2019
Cited alongside, same era.
Non-Autoregressive Machine Translation with Auxiliary Regularization
Wang, Y.; Tian, F.; He, D.; Qin, T.; Zhai, C.; and Liu, T. 2019 · 2019
Cited alongside, same era.
Paraphrases as Foreign Languages in Multilingual Neural Machine Translation
Zhou, Z.; Sperber, M.; and Waibel, A. 2019 · 2019
Cited alongside, same era.
Human-Paraphrased References Improve Neural Machine Translation
Freitag, M.; Foster, G.; Grangier, D.; and Cherry, C. 2020 · 2020
Cited alongside, same era.
Aligned Cross Entropy for Non-Autoregressive Machine Translation
Ghazvininejad, M.; Karpukhin, V.; Zettlemoyer, L.; and Levy, O. 2020 · 2020
Cited alongside, same era.
Gu, J.; and Kong, X. 2021 · 2021
Later among the works it cites.
Enriching Non-Autoregressive Transformer with Syntactic and Semantic Structures for Neural Machine Translation
Liu, Y.; Wan, Y.; Zhang, J.; Zhao, W.; and Yu, P. 2021 · 2021
Later among the works it cites.
Glancing Transformer for Non-Autoregressive Neural Machine Translation
Qian, L.; Zhou, H.; Bao, Y.; Wang, M.; Qiu, L.; Zhang, W.; Yu, Y.; and Li, L. 2021 · 2021
Later among the works it cites.
Guiding Non-Autoregressive Neural Machine Translation Decoding with Reordering Information
Ran, Q.; Lin, Y.; Li, P.; and Zhou, J. 2021 · 2021
Later among the works it cites.
Sequence-Level Training for Non-Autoregressive Neural Machine Translation
Shao, C.; Feng, Y.; Zhang, J.; Meng, F.; and Zhou, J. 2021 · 2021
Later among the works it cites.
AligNART: Non-autoregressive Neural Machine Translation by Jointly Learning to Estimate Alignment and Translate
Song, J.; Kim, S.; and Yoon, S. 2021 · 2021
Later among the works it cites.
Duplex sequence-to-sequence learning for reversible machine translation
Zheng, Z.; Zhou, H.; Huang, S.; Chen, J.; Xu, J.; and Li, L. 2021 · 2021
Later among the works it cites.
Directed Acyclic Transformer for Non-Autoregressive Machine Translation
Huang, F.; Zhou, H.; Liu, Y.; Li, H.; and Huang, M. 2022b · 2022
Closest in time.
Non-Monotonic Latent Alignments for CTC-Based Non-Autoregressive Machine Translation
Shao, C.; and Feng, Y. 2022 · 2022
Closest in time.
Viterbi Decoding of Directed Acyclic Transformer for Non-Autoregressive Machine Translation
Shao, C.; Ma, Z.; and Feng, Y. 2022 · 2022
Closest in time.
One Reference Is Not Enough: Diverse Distillation with Reference Selection for Non-Autoregressive Translation
Shao, C.; Wu, X.; and Feng, Y. 2022 · 2022
Closest in time.