Fetching the paper…
Reading the bibliography…
We study lossless acceleration for seq2seq generation with a novel decoding algorithm -- Aggressive Decoding.
Semi-autoregressive training improves mask-predict decoding
Marjan Ghazvininejad, Omer Levy, and Luke Zettlemoyer · 2001
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu · 2002
Earlier work this paper cites.
Mining revision log of language learning sns for automated japanese error correction of second language learners
Tomoya Mizumoto, Mamoru Komachi, Masaaki Nagata, and Yuji Matsumoto · 2011
Earlier work this paper cites.
A new dataset and method for automatically grading esol texts
Helen Yannakoudakis, Ted Briscoe, and Ben Medlock · 2011
Earlier work this paper cites.
Better evaluation for grammatical error correction
Daniel Dahlmeier and Hwee Tou Ng · 2012
Earlier work this paper cites.
Multiple networks are more efficient than one: Fast and accurate models via ensembles and cascades
Xiaofang Wang, Dan Kondratyuk, Kris M. Kitani, Yair Movshovitz-Attias, and Elad Eban · 2012
Earlier work this paper cites.
Building a large annotated corpus of learner english: The nus corpus of learner english
Daniel Dahlmeier, Hwee Tou Ng, and Siew Mei Wu · 2013
Earlier work this paper cites.
The CoNLL-2013 shared task on grammatical error correction
Hwee Tou Ng, Siew Mei Wu, Yuanbin Wu, Christian Hadiwinoto, and Joel Tetreault · 2013
Earlier work this paper cites.
The conll-2014 shared task on grammatical error correction
Hwee Tou Ng, Siew Mei Wu, Ted Briscoe, Christian Hadiwinoto, Raymond Hendy Susanto, and Christopher Bryant · 2014
Earlier work this paper cites.
Teaching machines to read and comprehend
Karl Moritz Hermann, Tomás Kociský, Edward Grefenstette, Lasse Espeholt, Will Kay, Mustafa Suleyman, and Phil Blunsom · 2015
Earlier work this paper cites.
Sequence-level knowledge distillation
Yoon Kim and Alexander M Rush · 2016
Earlier work this paper cites.
Neural machine translation of rare words with subword units
Rico Sennrich, Barry Haddow, and Alexandra Birch · 2016
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
Born again neural networks
Tommaso Furlanello, Zachary Lipton, Michael Tschannen, Laurent Itti, and Anima Anandkumar · 2018
Earlier work this paper cites.
Non-autoregressive neural machine translation
Jiatao Gu, James Bradbury, Caiming Xiong, Victor O.K. Li, and Richard Socher · 2018
Earlier work this paper cites.
Multi-scale dense networks for resource efficient image classification
Gao Huang, Danlu Chen, Tianhong Li, Felix Wu, Laurens van der Maaten, and Kilian Q. Weinberger · 2018
Earlier work this paper cites.
Deterministic non-autoregressive neural sequence modeling by iterative refinement
Jason Lee, Elman Mansimov, and Kyunghyun Cho · 2018
Earlier work this paper cites.
End-to-end non-autoregressive neural machine translation with connectionist temporal classification
Jindrich Libovický and Jindrich Helcl · 2018
Cited alongside, same era.
Scaling neural machine translation
Myle Ott, Sergey Edunov, David Grangier, and Michael Auli · 2018
Cited alongside, same era.
Blockwise parallel decoding for deep autoregressive models
Mitchell Stern, Noam Shazeer, and Jakob Uszkoreit · 2018
Cited alongside, same era.
Approximation algorithms for cascading prediction models
Matthew Streeter · 2018
Cited alongside, same era.
Semi-autoregressive neural machine translation
Chunqi Wang, Ji Zhang, and Haiqing Chen · 2018
Cited alongside, same era.
Parallel iterative edit models for local sequence transduction
Abhijeet Awasthi, Sunita Sarawagi, Rasna Goyal, Sabyasachi Ghosh, and Vihari Piratla · 2019
Cited alongside, same era.
Learning to recover from multi-modality errors for non-autoregressive neural machine translation
Qiu Ran, Yankai Lin, Peng Li, and Jie Zhou · 2020
Later among the works it cites.
Non-autoregressive machine translation with latent alignments
Chitwan Saharia, William Chan, Saurabh Saxena, and Mohammad Norouzi · 2020
Later among the works it cites.
Minimizing the bag-of-ngrams difference for non-autoregressive neural machine translation
Chenze Shao, Jinchao Zhang, Yang Feng, Fandong Meng, and Jie Zhou · 2020
Later among the works it cites.
Seq2edits: Sequence transduction using span-level edit operations
Felix Stahlberg and Shankar Kumar · 2020
Later among the works it cites.
Deebert: Dynamic early exiting for accelerating BERT inference
Ji Xin, Raphael Tang, Jaejun Lee, Yaoliang Yu, and Jimmy Lin · 2020
Later among the works it cites.
BERT loses patience: Fast and robust inference with early exit
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
The bea-2019 shared task on grammatical error correction
Christopher Bryant, Mariano Felice, Øistein E Andersen, and Ted Briscoe · 2019
Cited alongside, same era.
Mask-predict: Parallel decoding of conditional masked language models
Marjan Ghazvininejad, Omer Levy, Yinhan Liu, and Luke Zettlemoyer · 2019
Cited alongside, same era.
Neural grammatical error correction systems with unsupervised pre-training on synthetic data
Roman Grundkiewicz, Marcin Junczys-Dowmunt, and Kenneth Heafield · 2019
Cited alongside, same era.
Levenshtein transformer
Jiatao Gu, Changhan Wang, and Junbo Zhao · 2019
Cited alongside, same era.
Sequence-to-sequence pre-training with data augmentation for sentence rewriting
Yi Zhang, Tao Ge, Furu Wei, Ming Zhou, and Xu Sun · 2019
Cited alongside, same era.
Mengyun Chen, Tao Ge, Xingxing Zhang, Furu Wei, and Ming Zhou · 2020
Cited alongside, same era.
Wangchunshu Zhou, Canwen Xu, Tao Ge, Julian J. McAuley, Ke Xu, and Furu Wei · 2020
Later among the works it cites.
Order-agnostic cross entropy for non-autoregressive machine translation
Cunxiao Du, Zhaopeng Tu, and Jing Jiang · 2021
Later among the works it cites.
Learning to rewrite for non-autoregressive neural machine translation
Xinwei Geng, Xiaocheng Feng, and Bing Qin · 2021
Later among the works it cites.
Fully non-autoregressive neural machine translation: Tricks of the trade
Jiatao Gu and Xiang Kong · 2021
Later among the works it cites.
Multi-task learning with shared encoder for non-autoregressive machine translation
Yongchang Hao, Shilin He, Wenxiang Jiao, Zhaopeng Tu, Michael R. Lyu, and Xing Wang · 2021
Later among the works it cites.
Non-autoregressive translation with layer-wise prediction and deep supervision
Chenyang Huang, Hao Zhou, Osmar R. Zaïane, Lili Mou, and Lei Li · 2021
Later among the works it cites.
Glancing transformer for non-autoregressive neural machine translation
Lihua Qian, Hao Zhou, Yu Bao, Mingxuan Wang, Lin Qiu, Weinan Zhang, Yong Yu, and Lei Li · 2021
Later among the works it cites.
Step-unrolled denoising autoencoders for text generation
Nikolay Savinov, Junyoung Chung, Mikolaj Binkowski, Erich Elsen, and Aäron van den Oord · 2021
Later among the works it cites.
Alignart: Non-autoregressive neural machine translation by jointly learning to estimate alignment and translate
Jongyoon Song, Sungwon Kim, and Sungroh Yoon · 2021
Later among the works it cites.
Early exiting with ensemble internal classifiers
Tianxiang Sun, Yunhua Zhou, Xiangyang Liu, Xinyu Zhang, Hao Jiang, Zhao Cao, Xuanjing Huang, and Xipeng Qiu · 2021
Later among the works it cites.
Improving non-autoregressive translation models without distillation
Xiao Shi Huang, Felipe Perez, and Maksims Volkovs · 2022
Closest in time.
Deepspeed-moe: Advancing mixture-of-experts inference and training to power next-generation ai scale
Samyam Rajbhandari, Conglong Li, Zhewei Yao, Minjia Zhang, Reza Yazdani Aminabadi, Ammar Ahmad Awan, Jeff Rasley, and Yuxiong He · 2022
Closest in time.