Fetching the paper…
Reading the bibliography…
Reference-based metrics that operate at the sentence-level typically outperform quality estimation metrics, which have access only to the source and system output.
Cross-lingual language model pretraining
Guillaume Lample and Alexis Conneau. 2019 · 1901
Earlier work this paper cites.
A survey on document-level machine translation: Methods and evaluation
Sameen Maruf, Fahimeh Saleh, and Gholamreza Haffari. 2019 · 1912
Earlier work this paper cites.
Evaluating discourse phenomena in neural machine translation
Rachel Bawden, Rico Sennrich, Alexandra Birch, and Barry Haddow. 2018 · 2018
Earlier work this paper cites.
A large-scale test set for the evaluation of context-aware pronoun translation in neural machine translation
Mathias Müller, Annette Rios, Elena Voita, and Rico Sennrich. 2018 · 2018
Earlier work this paper cites.
Elena Voita, Rico Sennrich, and Ivan Titov. 2019 · 2019
Earlier work this paper cites.
Document-level neural MT: A systematic comparison
António Lopes, M. Amin Farajian, Rachel Bawden, Michael Zhang, and André F. T. Martins. 2020 · 2020
Earlier work this paper cites.
Unbabel’s participation in the WMT20 metrics shared task
Ricardo Rei, Craig Stewart, Ana C Farinha, and Alon Lavie. 2020 · 2020
Cited alongside, same era.
InfoXLM: An information-theoretic framework for cross-lingual language model pre-training
Zewen Chi, Li Dong, Furu Wei, Nan Yang, Saksham Singhal, Wenhui Wang, Xia Song, Xian-Ling Mao, Heyan Huang, and Ming Zhou. 2021 · 2021
Cited alongside, same era.
Larger-scale transformers for multilingual masked language modeling
Naman Goyal, Jingfei Du, Myle Ott, Giri Anantharaman, and Alexis Conneau. 2021 · 2021
Cited alongside, same era.
To ship or not to ship: An extensive evaluation of automatic metrics for machine translation
Tom Kocmi, Christian Federmann, Roman Grundkiewicz, Marcin Junczys-Dowmunt, Hitokazu Matsushita, and Arul Menezes. 2021 · 2021
Cited alongside, same era.
Results of WMT22 metrics shared task: Stop using BLEU – neural metrics are better and more robust
Markus Freitag, Ricardo Rei, Nitika Mathur, Chi-kiu Lo, Craig Stewart, Eleftherios Avramidis, Tom Kocmi, George Foster, Alon Lavie, and André F. T. Martins. 2022 · 2022
Cited alongside, same era.
Findings of the 2022 conference on machine translation (WMT22)
Tom Kocmi, Rachel Bawden, Ondřej Bojar, Anton Dvorkovich, Christian Federmann, Mark Fishel, Thamme Gowda, Yvette Graham, Roman Grundkiewicz, Barry Haddow, Rebecca Knowles, Philipp Koehn, Christof Monz, Makoto Morishita, Masaaki Nagata, Toshiaki Nakazawa, Michal Novák, Martin Popel, and Maja Popović. 2022 · 2022
Later among the works it cites.
CometKiwi: IST-unbabel 2022 submission for the quality estimation shared task
Ricardo Rei, Marcos Treviso, Nuno M. Guerreiro, Chrysoula Zerva, Ana C Farinha, Christine Maroti, José G. C. de Souza, Taisiya Glushkova, Duarte Alves, Luisa Coheur, Alon Lavie, and André F. T. Martins. 2022 · 2022
Later among the works it cites.
Embarrassingly easy document-level MT metrics: How to convert any pretrained metric into a document-level metric
Giorgos Vernikos, Brian Thompson, Prashant Mathur, and Marcello Federico. 2022 · 2022
Later among the works it cites.
Training and meta-evaluating machine translation evaluation metrics at the paragraph level
Daniel Deutsch, Juraj Juraska, Mara Finkelstein, and Markus Freitag. 2023 · 2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
BlonDe: An automatic evaluation metric for document-level machine translation
Yuchen Jiang, Tianyu Liu, Shuming Ma, Dongdong Zhang, Jian Yang, Haoyang Huang, Rico Sennrich, Ryan Cotterell, Mrinmaya Sachan, and Ming Zhou. 2022 · 2022
Cited alongside, same era.
When does translation require context? a data-driven, multilingual exploration
Patrick Fernandes, Kayo Yin, Emmy Liu, André Martins, and Graham Neubig. 2023 · 2023
Closest in time.
How good are gpt models at machine translation? a comprehensive evaluation
Amr Hendy, Mohamed Abdelrehim, Amr Sharaf, Vikas Raunak, Mohamed Gabr, Hitokazu Matsushita, Young Jin Kim, Mohamed Afify, and Hany Hassan Awadalla. 2023 · 2023
Closest in time.