Fetching the paper…
Reading the bibliography…
State-of-the-art neural machine translation models generate a translation from left to right and every step is conditioned on the previously generated tokens.
Pay less attention with lightweight and dynamic convolutions
Wu, F., Fan, A., Baevski, A., Dauphin, Y., and Auli, M · 1901
Earlier work this paper cites.
Insertion-based decoding with automatically inferred generation order
Gu, J., Liu, Q., and Cho, K · 1902
Earlier work this paper cites.
Insertion transformer: Flexible sequence generation via insertion operations
Stern, M., Chan, W., Kiros, J. R., and Uszkoreit, J · 1902
Earlier work this paper cites.
Non-autoregressive machine translation with auxiliary regularization
Wang, Y., Tian, F., He, D., Qin, T., Zhai, C., and Liu, T.-Y · 1902
Earlier work this paper cites.
Cloze-driven pretraining of self-attention networks, 2019
Baevski, A., Edunov, S., Liu, Y., Zettlemoyer, L. S., and Auli, M · 1903
Earlier work this paper cites.
Mask-predict: Parallel decoding of conditional masked language models
Ghazvininejad, M., Levy, O., Liu, Y., and Zettlemoyer, L. S · 1904
Earlier work this paper cites.
fairseq: A fast, extensible toolkit for sequence modeling
Ott, M., Edunov, S., Baevski, A., Fan, A., Gross, S., Ng, N., Grangier, D., and Auli, M · 1904
Earlier work this paper cites.
Gu, J., Wang, C., and Zhao, J · 1905
Earlier work this paper cites.
A generalized framework of sequence generation with application to undirected sequence models, 2019
Mansimov, E., Wang, A., and Cho, K · 1905
Earlier work this paper cites.
KERMIT: Generative insertion-based modeling for sequences, 2019a
Chan, W., Kitaev, N., Guu, K., Stern, M., and Uszkoreit, J · 1906
Earlier work this paper cites.
XLNet: Generalized autoregressive pretraining for language understanding
Yang, Z., Dai, Z., Yang, Y., Carbonell, J. G., Salakhutdinov, R., and Le, Q. V · 1906
Earlier work this paper cites.
RoBERTa: A robustly optimized bert pretraining approach, 2019
Liu, Y., Ott, M., Goyal, N., Du, J., Joshi, M., Chen, D., Levy, O., Lewis, M., Zettlemoyer, L. S., and Stoyanov, V · 1907
Earlier work this paper cites.
Shu, R., Lee, J., Nakayama, H., and Cho, K · 1908
Earlier work this paper cites.
Hint-based training for non-autoregressive machine translation
Li, Z., Lin, Z., He, D., Tian, F., Qin, T., Wang, L., and Liu, T.-Y · 1909
Earlier work this paper cites.
FlowSeq: Non-autoregressive conditional sequence generation with generative flow
Ma, X., Zhou, C., Li, X., Neubig, G., and Hovy, E. H · 1909
Earlier work this paper cites.
An empirical study of generation order for machine translation, 2019b
Chan, W., Stern, M., Kiros, J. R., and Uszkoreit, J · 1910
Earlier work this paper cites.
Fast structured decoding for sequence models
Sun, Z., Li, Z., Wang, H., He, D., Lin, Z., and Deng, Z · 1910
Earlier work this paper cites.
Guiding non-autoregressive neural machine translation decoding with reordering information, 2019
Ran, Q., Lin, Y., Li, P., and Zhou, J · 1911
Cited alongside, same era.
Minimizing the bag-of-ngrams difference for non-autoregressive neural machine translation
Shao, C., Zhang, J., Feng, Y., Meng, F., and Zhou, J · 1911
Cited alongside, same era.
Non-autoregressive video captioning with iterative refinement, 2019a
Yang, B., Liu, F., and Zou, Y · 1911
Cited alongside, same era.
Understanding knowledge distillation in non-autoregressive machine translation
Zhou, C., Neubig, G., and Gu, J · 1911
Cited alongside, same era.
Semi-autoregressive training improves mask-predict decoding, 2020b
Sequence-level knowledge distillation
Kim, Y. and Rush, A. M · 2016
Later among the works it cites.
Neural machine translation of rare words with subword units
Sennrich, R., Haddow, B., and Birch, A · 2016
Later among the works it cites.
Convolutional sequence to sequence learning
Gehring, J., Auli, M., Grangier, D., Yarats, D., and Dauphin, Y · 2017
Later among the works it cites.
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, L., and Polosukhin, I · 2017
Later among the works it cites.
Non-autoregressive neural machine translation
Gu, J., Bradbury, J., Xiong, C., Li, V. O. K., and Socher, R · 2018
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Ghazvininejad, M., Levy, O., and Zettlemoyer, L · 2001
Cited alongside, same era.
Improving non-autoregressive neural machine translation with monolingual data
Zhou, J. and Keung, P · 2001
Cited alongside, same era.
Li, X., Meng, Y., Yuan, A., Wu, F., and Li, J · 2002
Cited alongside, same era.
BLEU: a method for automatic evaluation of machine translation
Papineni, K., Roukos, S., Ward, T., and Zhu, W.-J · 2002
Cited alongside, same era.
Aligned cross entropy for non-autoregressive machine translation, 2020a
Ghazvininejad, M., Karpukhin, V., Zettlemoyer, L., and Levy, O · 2004
Cited alongside, same era.
Non-autoregressive machine translation with latent alignments, 2020
Saharia, C., Chan, W., Saxena, S., and Norouzi, M · 2004
Cited alongside, same era.
ENGINE: Energy-based inference networks for non-autoregressive machine translation
Tu, L., Pang, R. Y., Wiseman, S., and Gimpel, K · 2005
Cited alongside, same era.
POINTER: Constrained text generation via insertion-based generative pre-training, 2020
Zhang, Y., Wang, G., Li, C., Gan, Z., Brockett, C., and Dolan, B · 2005
Cited alongside, same era.
Later among the works it cites.
Achieving human parity on automatic Chinese to English news translation, 2018
Hassan, H., Aue, A., Chen, C., Chowdhary, V., Clark, J., Federmann, C., Huang, X., Junczys-Dowmunt, M., Lewis, W., Li, M., Liu, S., Liu, T.-Y., Luo, R., Menezes, A., Qin, T., Seide, F., Tan, X., Tian, F., Wu, L., Wu, S., Xia, Y., Zhang, D., Zhang, Z., and Zhou, M · 2018
Later among the works it cites.
Fast decoding in sequence models using discrete latent variables
Kaiser, L., Roy, A., Vaswani, A., Parmar, N., Bengio, S., Uszkoreit, J., and Shazeer, N · 2018
Later among the works it cites.
Deterministic non-autoregressive neural sequence modeling by iterative refinement
Lee, J. D., Mansimov, E., and Cho, K · 2018
Later among the works it cites.
End-to-end non-autoregressive neural machine translation with connectionist temporal classification
Libovický, J. and Helcl, J · 2018
Later among the works it cites.
Micikevicius, P., Narang, S., Alben, J., Diamos, G., Elsen, E., Garcia, D., Ginsburg, B., Houston, M., Kuchaiev, O., Venkatesh, G., and Wu, H · 2018
Later among the works it cites.
Scaling neural machine translation
Ott, M., Edunov, S., Grangier, D., and Auli, M · 2018
Later among the works it cites.
A call for clarity in reporting BLEU scores
Post, M · 2018
Later among the works it cites.
Blockwise parallel decoding for deep autoregressive models
Stern, M., Shazeer, N., and Uszkoreit, J · 2018
Later among the works it cites.
Dehghani, M., Gouws, S., Vinyals, O., Uszkoreit, J., and Kaiser, L · 2019
Later among the works it cites.
BERT: Pre-training of deep bidirectional transformers for language understanding
Devlin, J., Chang, M.-W., Lee, K., and Toutanova, K · 2019
Later among the works it cites.
Recognition and translation of code-switching speech utterances
Nakayama, S., Kano, T., Tjandra, A., Sakti, S., and Nakamura, S · 2019
Later among the works it cites.