Fetching the paper…
Reading the bibliography…
In this work, we study hallucinations in Neural Machine Translation (NMT), which lie at an extreme end on the spectrum of NMT pathologies.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
METEOR: An automatic metric for MT evaluation with improved correlation with human judgments
Satanjeev Banerjee and Alon Lavie. 2005 · 2005
Earlier work this paper cites.
Analyzing the source and target contributions to predictions in neural machine translation
Elena Voita, Rico Sennrich, and Ivan Titov. 2020 · 2010
Earlier work this paper cites.
chrF: character n-gram F-score for automatic MT evaluation
Maja Popović. 2015 · 2015
Earlier work this paper cites.
Sequence-level knowledge distillation
Yoon Kim and Alexander M. Rush. 2016 · 2016
Earlier work this paper cites.
Neural machine translation of rare words with subword units
Rico Sennrich, Barry Haddow, and Alexandra Birch. 2016 · 2016
Earlier work this paper cites.
Modeling coverage for neural machine translation
Zhaopeng Tu, Zhengdong Lu, Yang Liu, Xiaohua Liu, and Hang Li. 2016 · 2016
Earlier work this paper cites.
Six challenges for neural machine translation
Philipp Koehn and Rebecca Knowles. 2017 · 2017
Earlier work this paper cites.
Understanding back-translation at scale
Sergey Edunov, Myle Ott, Michael Auli, and David Grangier. 2018 · 2018
Earlier work this paper cites.
Dual conditional cross-entropy filtering of noisy parallel corpora
Marcin Junczys-Dowmunt. 2018 · 2018
Earlier work this paper cites.
Hallucinations in neural machine translation
Katherine Lee, Orhan Firat, Ashish Agarwal, Clara Fannjiang, and David Sussillo. 2018 · 2018
Cited alongside, same era.
Filtering and mining parallel data in a joint multilingual space
Holger Schwenk. 2018 · 2018
Cited alongside, same era.
Neural machine translation incorporating named entity
Arata Ugawa, Akihiro Tamura, Takashi Ninomiya, Hiroya Takamura, and Manabu Okumura. 2018 · 2018
Cited alongside, same era.
Massively Multilingual Sentence Embeddings for Zero-Shot Cross-Lingual Transfer and Beyond
Mikel Artetxe and Holger Schwenk. 2019 · 2019
Cited alongside, same era.
Naver labs Europe’s systems for the WMT19 machine translation robustness task
Alexandre Berard, Ioan Calapodescu, and Claude Roux. 2019 · 2019
Cited alongside, same era.
Robust neural machine translation with doubly adversarial inputs
Yong Cheng, Lu Jiang, and Wolfgang Macherey. 2019 · 2019
Does learning require memorization? a short tale about a long tail
Vitaly Feldman. 2020 · 2020
Later among the works it cites.
What neural networks memorize and why: Discovering the long tail via influence estimation
Vitaly Feldman and Chiyuan Zhang. 2020 · 2020
Later among the works it cites.
Improved natural language generation via loss truncation
Daniel Kang and Tatsunori Hashimoto. 2020 · 2020
Later among the works it cites.
Domain robustness in neural machine translation
Mathias Müller, Annette Rios, and Rico Sennrich. 2020 · 2020
Later among the works it cites.
On long-tailed phenomena in neural machine translation
Vikas Raunak, Siddharth Dalmia, Vivek Gupta, and Florian Metze. 2020 · 2020
Later among the works it cites.
On exposure bias, hallucination and domain shift in neural machine translation
Chaojun Wang and Rico Sennrich. 2020 · 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Identifying fluently inadequate output in neural and statistical machine translation
Marianna Martindale, Marine Carpuat, Kevin Duh, and Paul McNamee. 2019 · 2019
Cited alongside, same era.
fairseq: A fast, extensible toolkit for sequence modeling
Myle Ott, Sergey Edunov, Alexei Baevski, Angela Fan, Sam Gross, Nathan Ng, David Grangier, and Michael Auli. 2019 · 2019
Cited alongside, same era.
Bertscore: Evaluating text generation with bert
Tianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q Weinberger, and Yoav Artzi. 2019 · 2019
Cited alongside, same era.
Later among the works it cites.
Parallel corpus filtering via pre-trained language models
Boliang Zhang, Ajay Nagesh, and Kevin Knight. 2020 · 2020
Later among the works it cites.
Tilted empirical risk minimization
Tian Li, Ahmad Beirami, Maziar Sanjabi, and Virginia Smith. 2021 · 2021
Closest in time.
Robust early-learning: Hindering the memorization of noisy labels
Xiaobo Xia, Tongliang Liu, Bo Han, Chen Gong, Nannan Wang, Zongyuan Ge, and Yi Chang. 2021 · 2021
Closest in time.